DE eng

Search in the Catalogues and Directories

Page: 1 2 3
Hits 1 – 20 of 55

1
The Orange workflow for observing collocation trends ColTrend 1.0
Kosem, Iztok; Krek, Simon; Čibej, Jaka. - : Centre for Language Resources and Technologies, University of Ljubljana, 2021
BASE
Show details
2
Comprehensive Slovenian-Hungarian Dictionary 1.0
Kosem, Iztok; Bálint Čeh, Júlia; Ponikvar, Primož. - : Centre for Language Resources and Technologies, University of Ljubljana, 2021
BASE
Show details
3
Slovene ontology of semantic types for nouns SLONEST-noun 1.0
Kosem, Iztok; Pori, Eva; Gantar, Polona. - : Centre for Language Resources and Technologies, University of Ljubljana, 2021
BASE
Show details
4
Valency lexicon extracted from the Gigafida 2.1 corpus
Krek, Simon; Gantar, Polona; Krsnik, Luka. - : Centre for Language Resources and Technologies, University of Ljubljana, 2021
BASE
Show details
5
Morphological patterns from the Sloleks 2.0 lexicon 1.0
Arhar Holdt, Špela; Čibej, Jaka; Laskowski, Cyprian. - : Centre for Language Resources and Technologies, University of Ljubljana, 2021. : Jožef Stefan Institute, 2021
BASE
Show details
6
Multiword Expressions lexicon extracted from the Gigafida 2.1 corpus
Krek, Simon; Gantar, Apolonija; Laskowski, Cyprian. - : Centre for Language Resources and Technologies, University of Ljubljana, 2021
BASE
Show details
7
The Orange workflow for observing collocation clusters ColEmbed 1.0
Kosem, Iztok; Čibej, Jaka; Ljubešić, Nikola. - : Centre for Language Resources and Technologies, University of Ljubljana, 2021
BASE
Show details
8
Training corpus ssj500k 2.3
Krek, Simon; Dobrovoljc, Kaja; Erjavec, Tomaž. - : Centre for Language Resources and Technologies, University of Ljubljana, 2021
BASE
Show details
9
Frequency lists of collocations from the Gigafida 2.1 corpus
Krek, Simon; Gantar, Polona; Kosem, Iztok. - : Centre for Language Resources and Technologies, University of Ljubljana, 2021
BASE
Show details
10
Corpus of Written Standard Slovene Gigafida 2.0
Krek, Simon; Erjavec, Tomaž; Repar, Andraž. - : Centre for Language Resources and Technologies, University of Ljubljana, 2021
BASE
Show details
11
Language Teachers and Crowdsourcing: Insights from a Cross-European Survey
In: ISSN: 1331-6745 ; EISSN: 1849-0379 ; Rasprave Instituta za hrvatski jezik i jezikoslovlje ; https://hal.inria.fr/hal-02974069 ; Rasprave Instituta za hrvatski jezik i jezikoslovlje, 2020, 46 (1), pp.1-28. ⟨10.31724/rihjj.46.1.1⟩ (2020)
BASE
Show details
12
List of word relations from the Sloleks 2.0 lexicon 1.0
Čibej, Jaka; Arhar Holdt, Špela; Krek, Simon. - : Centre for Language Resources and Technologies, University of Ljubljana, 2020. : Jožef Stefan Institute, 2020
BASE
Show details
13
Slovene translation of SuperGLUE
Žagar, Aleš; Robnik-Šikonja, Marko; Goli, Teja. - : Faculty of Computer and Information Science, University of Ljubljana, 2020
BASE
Show details
14
Frequency lists of character-level n-grams from the GOS 1.0 corpus 1.1
Čibej, Jaka; Arhar Holdt, Špela; Dobrovoljc, Kaja. - : Centre for Language Resources and Technologies, University of Ljubljana, 2020. : Jožef Stefan Institute, 2020
BASE
Show details
15
Frequency lists of words from the GOS 1.0 corpus 1.1
Čibej, Jaka; Arhar Holdt, Špela; Dobrovoljc, Kaja. - : Centre for Language Resources and Technologies, University of Ljubljana, 2020. : Jožef Stefan Institute, 2020
BASE
Show details
16
Consonant-vowel structures in the GOS 1.0 corpus 1.1
Čibej, Jaka; Arhar Holdt, Špela; Dobrovoljc, Kaja. - : Centre for Language Resources and Technologies, University of Ljubljana, 2020. : Jožef Stefan Institute, 2020
BASE
Show details
17
Consonant-vowel structures in the Gigafida 2.0 corpus
Čibej, Jaka; Arhar Holdt, Špela; Dobrovoljc, Kaja. - : Centre for Language Resources and Technologies, University of Ljubljana, 2020. : Jožef Stefan Institute, 2020
BASE
Show details
18
Consonant-vowel structures in the GOS 1.0 corpus
Čibej, Jaka; Arhar Holdt, Špela; Dobrovoljc, Kaja. - : Centre for Language Resources and Technologies, University of Ljubljana, 2020. : Jožef Stefan Institute, 2020
BASE
Show details
19
Frequency lists of word-level n-grams from the GOS 1.0 corpus 1.1
Čibej, Jaka; Arhar Holdt, Špela; Dobrovoljc, Kaja; Krek, Simon. - : Centre for Language Resources and Technologies, University of Ljubljana, 2020. : Jožef Stefan Institute, 2020
Abstract: Frequency lists of word-level n-grams (or word sets) were extracted from the GOS 1.0 Corpus of Spoken Slovene (http://hdl.handle.net/11356/1040) using the LIST corpus extraction tool (http://hdl.handle.net/11356/1227). The lists contain all word-level 2-, 3-, 4- and 5-grams occurring in the corpus along with their absolute and relative frequencies, percentages, distribution across the text-types included in the corpus taxonomy, and five collocation measures: Dice, t-score, MI, MI3, logDice, and simple LL. The n-grams were extracted from lower-case word forms, standardized word forms, and morphosyntactic tags. For large lists, shortened versions with the first 150,000 lines were also prepared to facilitate further processing in spreadsheet analysis software. Compared to the previous version (http://hdl.handle.net/11356/1271), this one includes fixes of several typos and substitutes all instances of "normalized forms" with the more adequate term "standardized forms" (as used in the SSJ project).
Keyword: morphosyntactic tags; n-grams; Slovenian language; spoken corpus; standardized forms; word forms; word sets; words
URL: http://hdl.handle.net/11356/1365
BASE
Hide details
20
Reference List of Slovene Frequent Common Words
Pollak, Senja; Arhar Holdt, Špela; Krek, Simon. - : Jožef Stefan Institute, 2020. : Centre for Language Resources and Technologies, University of Ljubljana, 2020
BASE
Show details

Page: 1 2 3

Catalogues
0
0
0
0
1
0
0
Bibliographies
0
0
0
0
0
0
2
0
0
Linked Open Data catalogues
0
Online resources
0
0
0
0
Open access documents
52
0
0
0
0
© 2013 - 2024 Lin|gu|is|tik | Imprint | Privacy Policy | Datenschutzeinstellungen ändern