DE eng

Search in the Catalogues and Directories

Hits 1 – 12 of 12

1
Unsupervised Cross-Lingual Information Retrieval using Monolingual Data Only ...
Abstract: We propose a fully unsupervised framework for ad-hoc cross-lingual information retrieval (CLIR) which requires no bilingual data at all. The framework leverages shared cross-lingual word embedding spaces in which terms, queries, and documents can be represented, irrespective of their actual language. The shared embedding spaces are induced solely on the basis of monolingual corpora in two languages through an iterative process based on adversarial neural networks. Our experiments on the standard CLEF CLIR collections for three language pairs of varying degrees of language similarity (English-Dutch/Italian/Finnish) demonstrate the usefulness of the proposed fully unsupervised approach. Our CLIR models with unsupervised cross-lingual embeddings outperform baselines that utilize cross-lingual embeddings induced relying on word-level and document-level alignments. We then demonstrate that further improvements can be achieved by unsupervised ensemble CLIR models. We believe that the proposed framework is the ... : accepted at SIGIR'18 (preprint) ...
Keyword: Computation and Language cs.CL; FOS Computer and information sciences
URL: https://arxiv.org/abs/1805.00879
https://dx.doi.org/10.48550/arxiv.1805.00879
BASE
Hide details
2
Adversarial Propagation and Zero-Shot Cross-Lingual Transfer of Word Vector Specialization ...
BASE
Show details
3
Post-Specialisation: Retrofitting Vectors of Words Unseen in Lexical Resources ...
BASE
Show details
4
A Resource-Light Method for Cross-Lingual Semantic Textual Similarity ...
BASE
Show details
5
Post-Specialisation: Retrofitting Vectors of Words Unseen in Lexical Resources ...
Vulic, Ivan; Glavaš, Goran; Mrkšić, Nikola. - : Apollo - University of Cambridge Repository, 2018
BASE
Show details
6
ArguminSci: a tool for analyzing argumentation and rhetorical aspects in scientific writing
Glavaš, Goran; Lauscher, Anne; Eckert, Kai. - : Association for Computational Linguistics, 2018
BASE
Show details
7
An argument-annotated corpus of scientific publications
Ponzetto, Simone Paolo; Lauscher, Anne; Glavaš, Goran. - : Association for Computational Linguistics, 2018
BASE
Show details
8
Investigating the role of argumentation in the rhetorical analysis of scientific publications with neural multi-task learning models
Ponzetto, Simone Paolo; Eckert, Kai; Lauscher, Anne. - : Association for Computational Linguistics, 2018
BASE
Show details
9
Post-specialisation: Retrofitting vectors of words unseen in lexical resources
Mrkšić, Nikola; Glavaš, Goran; Korhonen, Anna. - : Association for Computational Linguistics, 2018
BASE
Show details
10
Discriminating between lexico-semantic relations with the specialization tensor model
Vulić, Ivan; Glavaš, Goran. - : Association for Computational Linguistics, 2018
BASE
Show details
11
Adversarial propagation and zero-shot cross-lingual transfer of word vector specialization
Ponti, Edoardo Maria; Vulić, Ivan; Glavaš, Goran. - : Association for Computational Linguistics, 2018
BASE
Show details
12
Explicit retrofitting of distributional word vectors
Glavaš, Goran; Vulić, Ivan. - : Association for Computational Linguistics, 2018
BASE
Show details

Catalogues
0
0
0
0
0
0
0
Bibliographies
0
0
0
0
0
0
0
0
0
Linked Open Data catalogues
0
Online resources
0
0
0
0
Open access documents
12
0
0
0
0
© 2013 - 2024 Lin|gu|is|tik | Imprint | Privacy Policy | Datenschutzeinstellungen ändern