4 |
The Orange workflow for observing collocation trends ColTrend 1.0
|
|
|
|
BASE
|
|
Show details
|
|
5 |
Slovene ontology of semantic types for nouns SLONEST-noun 1.0
|
|
|
|
BASE
|
|
Show details
|
|
7 |
Multiword Expressions lexicon extracted from the Gigafida 2.1 corpus
|
|
|
|
BASE
|
|
Show details
|
|
8 |
The Orange workflow for observing collocation clusters ColEmbed 1.0
|
|
|
|
BASE
|
|
Show details
|
|
10 |
Frequency lists of collocations from the Gigafida 2.1 corpus
|
|
|
|
BASE
|
|
Show details
|
|
14 |
Frequency lists of character-level n-grams from the GOS 1.0 corpus 1.1
|
|
|
|
BASE
|
|
Show details
|
|
18 |
Consonant-vowel structures in the Gigafida 2.0 corpus
|
|
|
|
Abstract:
The lists contain consonant-vowel structures of all lemmas and word forms in the Gigafida 2.0 corpus. In each unit, its characters were converted as follows: C - consonant (in lists with finegrained character categorizations, consonants were divided into Z - sonorant, G - voiced obstruent, and K - voiceless obstruent), V - vowel, X - foreign consonant, Y - foreign vowel, S - symbol, P - punctuation, N - number, F - non-Latin-script character, ! - other. Each consonant-vowel structure also contains its frequency in the corpus (i.e. the total sum of the frequencies of all units corresponding to the consonant-vowel structure), as well as the set of all units (in the lists labeled "entire") or the set of its 30 most frequent units (in the lists labeled as "short"), along with their part-of-speech categories and their individual frequencies). They also contain the number of all unique units within the consonant-vowel structure. The lists were prepared based on frequency lists extracted from Gigafida 2.0 using LIST: http://hdl.handle.net/11356/1276 Note that there exists a related resource, "Consonant-vowel structures in the GOS 1.0 corpus", http://hdl.handle.net/11356/1290.
|
|
Keyword:
consonant-vowel structures; consonants; frequency list; Gigafida; obstruents; Slovenian language; sonorants; vowels
|
|
URL: http://hdl.handle.net/11356/1289
|
|
BASE
|
|
Hide details
|
|
20 |
Frequency lists of word-level n-grams from the GOS 1.0 corpus 1.1
|
|
|
|
BASE
|
|
Show details
|
|
|
|