1 |
AUTOLEX: An Automatic Framework for Linguistic Exploration ...
|
|
|
|
Abstract:
Each language has its own complex systems of word, phrase, and sentence construction, the guiding principles of which are often summarized in grammar descriptions for the consumption of linguists or language learners. However, manual creation of such descriptions is a fraught process, as creating descriptions which describe the language in "its own terms" without bias or error requires both a deep understanding of the language at hand and linguistics as a whole. We propose an automatic framework AutoLEX that aims to ease linguists' discovery and extraction of concise descriptions of linguistic phenomena. Specifically, we apply this framework to extract descriptions for three phenomena: morphological agreement, case marking, and word order, across several languages. We evaluate the descriptions with the help of language experts and propose a method for automated evaluation when human evaluation is infeasible. ... : 9 pages ...
|
|
Keyword:
Computation and Language cs.CL; FOS Computer and information sciences
|
|
URL: https://dx.doi.org/10.48550/arxiv.2203.13901 https://arxiv.org/abs/2203.13901
|
|
BASE
|
|
Hide details
|
|
2 |
SD-QA: Spoken Dialectal Question Answering for the Real World ...
|
|
|
|
BASE
|
|
Show details
|
|
3 |
Phoneme Recognition through Fine Tuning of Phonetic Representations: a Case Study on Luhya Language Varieties ...
|
|
|
|
BASE
|
|
Show details
|
|
4 |
Machine Translation into Low-resource Language Varieties ...
|
|
|
|
BASE
|
|
Show details
|
|
5 |
Code to Comment Translation: A Comparative Study on Model Effectiveness & Errors ...
|
|
|
|
BASE
|
|
Show details
|
|
6 |
Systematic Inequalities in Language Technology Performance across the World's Languages ...
|
|
|
|
BASE
|
|
Show details
|
|
7 |
Multilingual Code-Switching for Zero-Shot Cross-Lingual Intent Prediction and Slot Filling ...
|
|
|
|
BASE
|
|
Show details
|
|
8 |
Investigating Post-pretraining Representation Alignment for Cross-Lingual Question Answering ...
|
|
|
|
BASE
|
|
Show details
|
|
9 |
Towards More Equitable Question Answering Systems: How Much More Data Do You Need? ...
|
|
|
|
BASE
|
|
Show details
|
|
10 |
Cross-Lingual Text Classification of Transliterated Hindi and Malayalam ...
|
|
|
|
BASE
|
|
Show details
|
|
11 |
Evaluating the Morphosyntactic Well-formedness of Generated Texts ...
|
|
|
|
BASE
|
|
Show details
|
|
12 |
Lexically Aware Semi-Supervised Learning for OCR Post-Correction ...
|
|
|
|
BASE
|
|
Show details
|
|
13 |
When is Wall a Pared and when a Muro? -- Extracting Rules Governing Lexical Selection ...
|
|
|
|
BASE
|
|
Show details
|
|
15 |
Towards Minimal Supervision BERT-based Grammar Error Correction ...
|
|
|
|
BASE
|
|
Show details
|
|
16 |
SIGMORPHON 2020 Shared Task 0: Typologically Diverse Morphological Inflection ...
|
|
|
|
BASE
|
|
Show details
|
|
17 |
It's not a Non-Issue: Negation as a Source of Error in Machine Translation ...
|
|
|
|
BASE
|
|
Show details
|
|
18 |
Automatic Extraction of Rules Governing Morphological Agreement ...
|
|
|
|
BASE
|
|
Show details
|
|
19 |
A Summary of the First Workshop on Language Technology for Language Documentation and Revitalization ...
|
|
|
|
BASE
|
|
Show details
|
|
20 |
Universal Phone Recognition with a Multilingual Allophone System ...
|
|
|
|
BASE
|
|
Show details
|
|
|
|