Home Catalogue search

eng

Refine your search:
- Keyword
- Creator / Publisher:
- Year
- Medium
- Type
- BLLDB-Access:
  - free (83)
  - subject to license (1)

Search in the Catalogues and Directories






	Sort by
Simple Search

Page: 1 2 3 4 5

Hits 1 – 20 of 83

1	The Orange workflow for observing collocation trends ColTrend 1.0
	Kosem, Iztok; Krek, Simon; Čibej, Jaka; Gantar, Polona; Arhar Holdt, Špela; Logar, Nataša; Laskowski, Cyprian; Klemenc, Bojan; Ljubešić, Nikola; Dobrovoljc, Kaja; Gorjanc, Vojko; Pori, Eva. - : Centre for Language Resources and Technologies, University of Ljubljana, 2021
	Abstract: The Orange workflow for observing collocation trends ColTrend 1.0 ColTrend is a workflow (.OWS file) for Orange Data Mining (an open-source machine learning and data visualization software: https://orangedatamining.com/) that allows the user to observe temporal collocation trends in corpora. The workflow consists of a series of Python scripts, data filters, and visualizers. As input, the workflow takes a .CSV file with data on collocations and their relative frequencies by year of publication extracted from a corpus. As output, it provides a .TSV file containing the same data (or a filtered selection thereof) enriched with four measures that indicate the collocation’s temporal trend in the corpus: (1) the slope (k) of a linear regression model fitted to the frequency data, which indicates whether the frequency of use of the collocation is increasing or declining; (2) the coefficient of determination (R2) of the linear regression model, indicating how linear the change in the collocation’s use is; (3) the ratio (m) of maximum relative frequency and average relative frequency, which indicates peaks in collocation usage; and (4) the coefficient of recent growth (t), which indicates an increased usage of the collocation in the last three years of the observed corpus data. The entry also contains three .CSV files that can be used to test the workflow. The files contain collocation candidates (along with their relative frequencies per year of publication) extracted from the Gigafida 2.0 Corpus of Written Slovene (https://viri.cjvt.si/gigafida/) with three different syntactic structures (as defined in http://hdl.handle.net/11356/1415): 1) p0-s0 (adjective + noun, e.g. rezervni sklad), 2) s0-s2 (noun + noun in the genitive case, e.g. ukinitev lastnine), and 3) gg-s4 (verb + noun in the accusative case, e.g. pripraviti besedilo). It should be noted that only collocation candidates with absolute frequency of 15 and above were extracted. Please note that the ColTrend workflow requires the installation of the Text Mining add-on for Orange. For installation instructions as well as a more detailed description of the different phases of the workflow and the measures used to observe the collocation trends, please consult the README file.
	Keyword: collocations; linear regression; relative frequency; temporal trends
	URL: http://hdl.handle.net/11356/1424
	BASE
	Hide details

2	Comprehensive Slovenian-Hungarian Dictionary 1.0
	Kosem, Iztok; Bálint Čeh, Júlia; Ponikvar, Primož. - : Centre for Language Resources and Technologies, University of Ljubljana, 2021
	BASE
	Show details

3	Slovene ontology of semantic types for nouns SLONEST-noun 1.0
	Kosem, Iztok; Pori, Eva; Gantar, Polona. - : Centre for Language Resources and Technologies, University of Ljubljana, 2021
	BASE
	Show details

4	Valency lexicon extracted from the Gigafida 2.1 corpus
	Krek, Simon; Gantar, Polona; Krsnik, Luka. - : Centre for Language Resources and Technologies, University of Ljubljana, 2021
	BASE
	Show details

5	Multiword Expressions lexicon extracted from the Gigafida 2.1 corpus
	Krek, Simon; Gantar, Apolonija; Laskowski, Cyprian. - : Centre for Language Resources and Technologies, University of Ljubljana, 2021
	BASE
	Show details

6	The Orange workflow for observing collocation clusters ColEmbed 1.0
	Kosem, Iztok; Čibej, Jaka; Ljubešić, Nikola. - : Centre for Language Resources and Technologies, University of Ljubljana, 2021
	BASE
	Show details

7	Frequency lists of collocations from the Gigafida 2.1 corpus
	Krek, Simon; Gantar, Polona; Kosem, Iztok. - : Centre for Language Resources and Technologies, University of Ljubljana, 2021
	BASE
	Show details

8	Corpus of Written Standard Slovene Gigafida 2.0
	Krek, Simon; Erjavec, Tomaž; Repar, Andraž. - : Centre for Language Resources and Technologies, University of Ljubljana, 2021
	BASE
	Show details

9	ELEXIS: Technical and social infrastructure for lexicography ...
	Woldrich, Anna; Goli, Teja; Kosem, Iztok. - : Zenodo, 2021
	BASE
	Show details

10	ELEXIS: Technical and social infrastructure for lexicography ...
	Woldrich, Anna; Goli, Teja; Kosem, Iztok. - : Zenodo, 2021
	BASE
	Show details

11	Creating Expert Knowledge by Relying on Language Learners: a Generic Approach for Mass-Producing Language Resources by Combining Implicit Crowdsourcing and Language Learning
	Nicolas, Lionel; Lyding, Verena; Borg, Claudia...
	In: LREC 2020 - Language Resources and Evaluation Conference ; https://hal.inria.fr/hal-02879883 ; LREC 2020 - Language Resources and Evaluation Conference, May 2020, Marseille, France (2020)
	BASE
	Show details

12	The image of the monolingual dictionary across Europe. Results of the European survey of dictionary use and culture
	Kosem, Iztok [Verfasser]; Lew, Robert [Verfasser]; Müller-Spitzer, Carolin [Verfasser]. - Mannheim : Leibniz-Institut für Deutsche Sprache (IDS), Bibliothek, 2019
	DNB Subject Category Language
	Show details

13	The microstructure of Online Linguistics Dictionaries: obligatory and facultative elements
	Flinz, Carolina [Verfasser]; Kosem, Iztok [Herausgeber]; Kosem, Karmen [Herausgeber]. - Mannheim : Leibniz-Institut für Deutsche Sprache (IDS), Bibliothek, 2019
	DNB Subject Category Language
	Show details

14	The Sintra variations – thinking outside the box in designing online dictionaries
	Michaelis, Frank [Verfasser]; Müller-Spitzer, Carolin [Verfasser]; Wolfer, Sascha [Verfasser]. - Mannheim : Leibniz-Institut für Deutsche Sprache (IDS), Bibliothek, 2019
	DNB Subject Category Language
	Show details

15	A corpus-based lexical resource of spoken German in interaction
	Meliss, Meike [Verfasser]; Möhrs, Christine [Verfasser]; Ribeiro Silveira, Maria [Verfasser]. - Mannheim : Leibniz-Institut für Deutsche Sprache (IDS), Bibliothek, 2019
	DNB Subject Category Language
	Show details

16	A corpus-based lexical resource of spoken German in interaction
	Meliss, Meike [Verfasser]; Möhrs, Christine [Verfasser]; Ribeiro Silveira, Maria [Verfasser]. - Mannheim : Leibniz-Institut für Deutsche Sprache (IDS), Bibliothek, 2019
	DNB Subject Category Language
	Show details

17	A web of loans: multilingual loanword lexicography with property graphs
	Meyer, Peter [Verfasser]; Eppinger, Mirjam [Verfasser]; Kosem, Iztok [Herausgeber]. - Mannheim : Leibniz-Institut für Deutsche Sprache (IDS), Bibliothek, 2019
	DNB Subject Category Language
	Show details

18	How much “tourism” is there in dictionary apps? An empirical study of lexicographical resources on mobile devices (German, Italian, Spanish)
	Flinz, Carolina [Verfasser]; Egido Vicente, Maria [Verfasser]; Kosem, Iztok [Herausgeber]. - Mannheim : Leibniz-Institut für Deutsche Sprache (IDS), Bibliothek, 2019
	DNB Subject Category Language
	Show details

19	Attitudes of Slovenian Language Users Towards General Monolingual Dictionaries: An International Perspective
	Wolfer, Sascha [Verfasser]; Kosem, Iztok [Verfasser]; Müller-Spitzer, Carolin [Verfasser]. - Mannheim : Institut für Deutsche Sprache, Bibliothek, 2019
	DNB Subject Category Language
	Show details

20	The Image of the Monolingual Dictionary Across Europe. Results of the European Survey of Dictionary use and Culture
	Kosem, Iztok; Lew, Robert; Müller-Spitzer, Carolin...
	In: ISSN: 0950-3846 ; EISSN: 1477-4577 ; International Journal of Lexicography ; https://hal.archives-ouvertes.fr/hal-03512668 ; International Journal of Lexicography, Oxford University Press (OUP), 2019 (2019)
	BASE
	Show details

Page: 1 2 3 4 5

© 2013 - 2024 Lin|gu|is|tik | Imprint | Privacy Policy | Datenschutzeinstellungen ändern