2 |
Quality and Efficiency of Manual Annotation: Data from the Pre-annotation Bias Experiment (part of the PDT-C 2.0 project)
|
|
|
|
BASE
|
|
Show details
|
|
5 |
PDT-Vallex: Czech Valency lexicon linked to treebanks 4.0 (PDT-Vallex 4.0)
|
|
Urešová, Zdeňka; Bémová, Alevtina; Fučíková, Eva; Hajič, Jan; Kolářová, Veronika; Mikulová, Marie; Pajas, Petr; Panevová, Jarmila; Štěpánek, Jan. - : Charles University, Faculty of Mathematics and Physics, Institute of Formal and Applied Linguistics (UFAL), 2021
|
|
Abstract:
The valency lexicon PDT-Vallex 4.0 has been built in close connection with the annotation of the Prague Dependency Treebank project (PDT) and its successors (mainly the Prague Czech-English Dependency Treebank project, PCEDT, the spoken language corpus (PDTSC) and corpus of user-generated texts in the project Faust). It contains over 14500 valency frames for almost 8500 verbs which occurred in the PDT, PCEDT, PDTSC and Faust corpora. In addition, there are nouns, adjectives and adverbs, linked from the PDT part only, increasing the total to over 17000 valency frames for 13000 words. All the corpora have been published in 2020 as the PDT-C 1.0 corpus with the PDT-Vallex 4.0 dictionary included; this is a copy of the dictionary published as a separate item for those not interested in the corpora themselves. It is available in electronically processable format (XML), and also in more human readable form including corpus examples (see the WEBSITE link below, and the links to its main publications elsewhere in this metadata). The main feature of the lexicon is its linking to the annotated corpora - each occurrence of each verb is linked to the appropriate valency frame with additional (generalized) information about its usage and surface morphosyntactic form alternatives. It replaces the previously published unversioned edition of PDT-Vallex from 2014.
|
|
Keyword:
annotation; Czech language; lexical semantics; lexicon; linguistic data; PDT; valency; verbal valency
|
|
URL: http://hdl.handle.net/11234/1-3499
|
|
BASE
|
|
Hide details
|
|
11 |
Search for the Relation of Form and Function Using the ForFun Database
|
|
|
|
In: Prague Bulletin of Mathematical Linguistics , Vol 110, Iss 1, Pp 71-84 (2018) (2018)
|
|
BASE
|
|
Show details
|
|
14 |
Difference between Written and Spoken Czech: The Case of Verbal Nouns Denoting an Action
|
|
|
|
In: Prague Bulletin of Mathematical Linguistics , Vol 107, Iss 1, Pp 19-38 (2017) (2017)
|
|
BASE
|
|
Show details
|
|
|
|