2 |
A morph-based and a word-based treebank for Beja
|
|
|
|
In: SyntaxFest ; TLT 2021 - 20th International Workshop on Treebanks and Linguistic Theories ; https://hal.archives-ouvertes.fr/hal-03494462 ; TLT 2021 - 20th International Workshop on Treebanks and Linguistic Theories, Mar 2022, Sofia, Bulgaria (2022)
|
|
BASE
|
|
Show details
|
|
3 |
Generación de flexión morfológica con UniMorph.: Evaluación con base de datos relacional y pautas de entrenamiento
|
|
|
|
In: Procesamiento del lenguaje natural, ISSN 1135-5948, Nº. 68, 2022, pags. 61-70 (2022)
|
|
BASE
|
|
Show details
|
|
9 |
A morph-based and a word-based treebank for Beja
|
|
|
|
In: SyntaxFest ; https://hal.archives-ouvertes.fr/hal-03494462 ; SyntaxFest, In press (2021)
|
|
BASE
|
|
Show details
|
|
10 |
Old Catalan Morphosyntax: developing an annotated corpus
|
|
|
|
In: EISSN: 2059-481X ; Journal of Open Humanities Data ; https://hal.archives-ouvertes.fr/hal-03617737 ; Journal of Open Humanities Data, Ubiquity Press, 2021, 7, pp.30. ⟨10.5334/johd.54⟩ (2021)
|
|
BASE
|
|
Show details
|
|
18 |
Prague Dependency Treebank - Consolidated 1.0 (PDT-C 1.0)
|
|
Hajič, Jan; Bejček, Eduard; Bémová, Alevtina; Buráňová, Eva; Fučíková, Eva; Hajičová, Eva; Havelka, Jiří; Hlaváčová, Jaroslava; Homola, Petr; Ircing, Pavel; Kárník, Jiří; Kettnerová, Václava; Klyueva, Natalia; Kolářová, Veronika; Kučová, Lucie; Lopatková, Markéta; Mareček, David; Mikulová, Marie; Mírovský, Jiří; Nedoluzhko, Anna; Novák, Michal; Pajas, Petr; Panevová, Jarmila; Peterek, Nino; Poláková, Lucie; Popel, Martin; Popelka, Jan; Romportl, Jan; Rysová, Magdaléna; Semecký, Jiří; Sgall, Petr; Spoustová, Johanka; Straka, Milan; Straňák, Pavel; Synková, Pavlína; Ševčíková, Magda; Šindlerová, Jana; Štěpánek, Jan; Štěpánková, Barbora; Toman, Josef; Urešová, Zdeňka; Vidová Hladká, Barbora; Zeman, Daniel; Zikánová, Šárka; Žabokrtský, Zdeněk. - : Charles University, Faculty of Mathematics and Physics, Institute of Formal and Applied Linguistics (UFAL), 2021
|
|
Abstract:
A richly annotated and genre-diversified language resource, The Prague Dependency Treebank – Consolidated 1.0 (PDT-C 1.0, or PDT-C in short in the sequel) is a consolidated release of the existing PDT-corpora of Czech data, uniformly annotated using the standard PDT scheme. PDT-corpora included in PDT-C: Prague Dependency Treebank (the original PDT contents, written newspaper and journal texts from three genres); Czech part of Prague Czech-English Dependency Treebank (translated financial texts, from English), Prague Dependency Treebank of Spoken Czech (spoken data, including audio and transcripts and multiple speech reconstruction annotation); PDT-Faust (user-generated texts). The difference from the separately published original treebanks can be briefly described as follows: it is published in one package, to allow easier data handling for all the datasets; the data is enhanced with a manual linguistic annotation at the morphological layer and new version of morphological dictionary is enclosed; a common valency lexicon for all four original parts is enclosed. Documentation provides two browsing and editing desktop tools (TrEd and MEd) and the corpus is also available online for searching using PML-TQ.
|
|
Keyword:
bridging relations; clauses; coreference; dependency; discourse; lemmatization; lexical semantics; lexicon; morphology; multiword expressions; semantic relations; speech recognition; speech reconstruction; spoken corpus; syntax; tectogrammatics; tokenization; topic-focus articulation; treebank; valency
|
|
URL: http://hdl.handle.net/11234/1-3185
|
|
BASE
|
|
Hide details
|
|
|
|