2 |
Indian Language Wordnets and their Linkages with Princeton WordNet ...
|
|
|
|
BASE
|
|
Show details
|
|
4 |
Techniques for Jointly Extracting Entities and Relations: A Survey ...
|
|
|
|
BASE
|
|
Show details
|
|
5 |
Knowledge-based Extraction of Cause-Effect Relations from Biomedical Text ...
|
|
|
|
BASE
|
|
Show details
|
|
6 |
Role of Language Relatedness in Multilingual Fine-tuning of Language Models: A Case Study in Indo-Aryan Languages ...
|
|
|
|
BASE
|
|
Show details
|
|
7 |
M2H2: A Multimodal Multiparty Hindi Dataset For Humor Recognition in Conversations ...
|
|
|
|
BASE
|
|
Show details
|
|
8 |
"So You Think You're Funny?": Rating the Humour Quotient in Standup Comedy ...
|
|
|
|
BASE
|
|
Show details
|
|
10 |
Crosslingual Embeddings are Essential in UNMT for Distant Languages: An English to IndoAryan Case Study ...
|
|
|
|
BASE
|
|
Show details
|
|
11 |
Harnessing Cross-lingual Features to Improve Cognate Detection for Low-resource Languages ...
|
|
|
|
Abstract:
Cognates are variants of the same lexical form across different languages; for example 'fonema' in Spanish and 'phoneme' in English are cognates, both of which mean 'a unit of sound'. The task of automatic detection of cognates among any two languages can help downstream NLP tasks such as Cross-lingual Information Retrieval, Computational Phylogenetics, and Machine Translation. In this paper, we demonstrate the use of cross-lingual word embeddings for detecting cognates among fourteen Indian Languages. Our approach introduces the use of context from a knowledge graph to generate improved feature representations for cognate detection. We, then, evaluate the impact of our cognate detection mechanism on neural machine translation (NMT), as a downstream task. We evaluate our methods to detect cognates on a challenging dataset of twelve Indian languages, namely, Sanskrit, Hindi, Assamese, Oriya, Kannada, Gujarati, Tamil, Telugu, Punjabi, Bengali, Marathi, and Malayalam. Additionally, we create evaluation datasets ... : Published at COLING 2020 ...
|
|
Keyword:
Computation and Language cs.CL; FOS Computer and information sciences
|
|
URL: https://dx.doi.org/10.48550/arxiv.2112.08789 https://arxiv.org/abs/2112.08789
|
|
BASE
|
|
Hide details
|
|
12 |
Extracting N-ary Cross-sentence Relations using Constrained Subsequence Kernel ...
|
|
|
|
BASE
|
|
Show details
|
|
13 |
Related Tasks can Share! A Multi-task Framework for Affective language ...
|
|
|
|
BASE
|
|
Show details
|
|
14 |
Utilizing Language Relatedness to improve Machine Translation: A Case Study on Languages of the Indian Subcontinent ...
|
|
|
|
BASE
|
|
Show details
|
|
15 |
Addressing word-order Divergence in Multilingual Neural Machine Translation for extremely Low Resource Languages ...
|
|
|
|
BASE
|
|
Show details
|
|
16 |
Morphology Generation for Statistical Machine Translation ...
|
|
|
|
BASE
|
|
Show details
|
|
17 |
Role of Morphology Injection in Statistical Machine Translation ...
|
|
|
|
BASE
|
|
Show details
|
|
19 |
SMPOST: Parts of Speech Tagger for Code-Mixed Indic Social Media Text ...
|
|
|
|
BASE
|
|
Show details
|
|
20 |
Indowordnets help in Indian Language Machine Translation ...
|
|
|
|
BASE
|
|
Show details
|
|
|
|