DE eng

Search in the Catalogues and Directories

Page: 1 2 3 4
Hits 1 – 20 of 79

1
Q-CAT Corpus Annotation Tool 1.3
Brank, Janez. - : Jožef Stefan Institute, 2022
BASE
Show details
2
Annotation manuelle des émotions dans des textes écrits avec la plateforme Glozz. ; Annotation manuelle des émotions dans des textes écrits avec la plateforme Glozz.: Guide d'annotation
In: https://hal.archives-ouvertes.fr/hal-03263194 ; [Rapport de recherche] MoDyCo; Université Paris Nanterre. 2021 (2021)
BASE
Show details
3
Manual annotation of occurrences of terms candidates in scientific papers ; Annotation manuelle d'occurrences de candidats termes et écrit scientifique
In: https://hal.archives-ouvertes.fr/hal-02005884 ; 2021 (2021)
BASE
Show details
4
Constituting the Democrat Corpus: Annotation and Evaluation Procedures ; Élaboration du corpus Democrat : procédures d’annotation et d’évaluation
In: ISSN: 0458-726X ; EISSN: 1958-9549 ; Langages ; https://hal.archives-ouvertes.fr/hal-03474329 ; Langages, Armand Colin (Larousse jusqu'en 2003), 2021, Un corpus annoté en chaînes de référence et son exploitation : le projet Democrat, pp.25-46 ; https://www.revues.armand-colin.com/lettres-langues/langages/langages-no-224-42021/elaboration-du-corpus-democrat-procedures-dannotation-devaluation (2021)
BASE
Show details
5
Choice of plausible alternatives dataset in Croatian COPA-HR
Ljubešić, Nikola. - : Jožef Stefan Institute, 2021
BASE
Show details
6
Annotated Corpus of Pre-Standardized Balkan Slavic Literature 1.1
Šimko, Ivan. - : Slavic Seminary, University of Zurich, 2021
BASE
Show details
7
Q-CAT Corpus Annotation Tool 1.2
Brank, Janez. - : Jožef Stefan Institute, 2021
BASE
Show details
8
Corpus of term-annotated texts RSDO5 1.1
BASE
Show details
9
Training corpus ssj500k 2.3
Abstract: The ssj500k training corpus contains about 500,000 tokens manually annotated on the levels of tokenisation, sentence segmentation, morphosyntactic tagging, and lemmatisation. About half of the corpus is also manually annotated with syntactic dependencies, named entities, and verbal multiword expressions. About a quarter of the corpus is also annotated with semantic role labels. The morphosyntactic tags and syntactic dependencies are included both in the JOS/MULTEXT-East framework, as well as in the framework of Universal Dependencies. The annotations of the ssj500k corpus follow (1) the MULTEXT-East V6 morphosyntactic specifications for Slovene, http://nl.ijs.si/ME/V6/msd/, (2) the JOS dependency schema, http://nl.ijs.si/jos/bib/jos-skladnja-navodila.pdf, the Universal Dependencies morphosyntactic specifications and syntactic dependencies for Slovene-SSJ, https://universaldependencies.org/, (4) the Janes annotation guidelines for Slovenian named entities, http://nl.ijs.si/janes/wp-content/uploads/2017/09/SlovenianNER-eng-v1.1.pdf, and (5) the Guidelines of the PARSEME shared task on verbal multiword expressions, http://parsemefr.lif.univ-mrs.fr/parseme-st-guidelines/1.1/ The vocabulary of (1) and (2) is provided in the back element and (3), (4), and (5) in the teiHeader of the TEI encoded corpus. The semantic role labels are also documented in the teiHeader. In contrast to the previous version 2.2, this version includes the corrected Universal Dependencies relations from UD version 2.8, updates the TEI encoding and adds UD annotations to the vertical file.
Keyword: CONLL-U; dependency treebank; manual annotation; named entities; parsing; part-of-speech tagging; semantic role labelling; TEI; tokenisation; verbal multiword expressions
URL: http://hdl.handle.net/11356/1434
BASE
Hide details
10
Corpus of term-annotated texts RSDO5 1.0
BASE
Show details
11
Slovenian Twitter hate speech dataset IMSyPP-sl
Kralj Novak, Petra; Mozetič, Igor; Ljubešić, Nikola. - : Jožef Stefan Institute, 2021
BASE
Show details
12
Crowdsourcing linguistic resources for natural non-standardised languages processing ; Myriadisation de ressources linguistiques pour le traitement automatique de langues non standardisées
Millour, Alice. - : HAL CCSD, 2020
In: https://hal.archives-ouvertes.fr/tel-03083213 ; Informatique et langage [cs.CL]. Sorbonne Universite, 2020. Français (2020)
BASE
Show details
13
List of formulaic sequences in spoken Slovenian
Dobrovoljc, Kaja; Roblek, Rebeka; Vianello, Chiara. - : Jožef Stefan Institute, 2020. : Centre for Language Resources and Technologies, University of Ljubljana, 2020
BASE
Show details
14
Annotated Corpus of Pre-Standardized Balkan Slavic Literature
Šimko, Ivan. - : Slavic Seminary, University of Zurich, 2020
BASE
Show details
15
Dataset of Slovene idiomatic expressions SloIE
Škvorc, Tadej; Gantar, Polona; Robnik-Šikonja, Marko. - : Faculty of Computer and Information Science, University of Ljubljana, 2020
BASE
Show details
16
Sentiment Annotated Dataset of Croatian News
Pelicon, Andraž; Pranjić, Marko; Miljković, Dragana. - : Jožef Stefan Institute, 2020
BASE
Show details
17
List of formulaic sequences in standard written Slovenian
Dobrovoljc, Kaja; Roblek, Rebeka; Vianello, Chiara. - : Jožef Stefan Institute, 2020. : Centre for Language Resources and Technologies, University of Ljubljana, 2020
BASE
Show details
18
Approximation et problèmes de dos : aspects quantitatifs et qualitatifs relatifs aux stratégies d’approximation d’un corpus issu d’un forum de santé
In: Langue française, N 207, 3, 2020-10-13, pp.41-56 (2020)
BASE
Show details
19
Analyse contrastive des noms sous-spécifiés à l’oral et à l’écrit à partir d’une extraction automatique
In: Langages, N 219, 3, 2020-08-11, pp.117-132 (2020)
BASE
Show details
20
Un corpus libre, évolutif et versionné en entités nommées du français
In: TALN 2019 - Traitement Automatique des Langues Naturelles ; https://hal.archives-ouvertes.fr/hal-02448590 ; TALN 2019 - Traitement Automatique des Langues Naturelles, Jul 2019, Toulouse, France (2019)
BASE
Show details

Page: 1 2 3 4

Catalogues
0
0
0
0
0
0
0
Bibliographies
0
0
0
0
0
0
0
0
0
Linked Open Data catalogues
0
Online resources
0
0
0
0
Open access documents
79
0
0
0
0
© 2013 - 2024 Lin|gu|is|tik | Imprint | Privacy Policy | Datenschutzeinstellungen ändern