DE eng

Search in the Catalogues and Directories

Hits 1 – 8 of 8

1
A generic and open framework for multiword expressions treatment : from acquisition to applications
BASE
Show details
2
A generic and open framework for multiword expressions treatment : from acquisition to applications
Abstract: The treatment of multiword expressions (MWEs), like take off, bus stop and big deal, is a challenge for NLP applications. This kind of linguistic construction is not only arbitrary but also much more frequent than one would initially guess. This thesis investigates the behaviour of MWEs across different languages, domains and construction types, proposing and evaluating an integrated methodological framework for their acquisition. There have been many theoretical proposals to define, characterise and classify MWEs. We adopt generic definition stating that MWEs are word combinations which must be treated as a unit at some level of linguistic processing. They present a variable degree of institutionalisation, arbitrariness, heterogeneity and limited syntactic and semantic variability. There has been much research on automatic MWE acquisition in the recent decades, and the state of the art covers a large number of techniques and languages. Other tasks involving MWEs, namely disambiguation, interpretation, representation and applications, have received less emphasis in the field. The first main contribution of this thesis is the proposal of an original methodological framework for automatic MWE acquisition from monolingual corpora. This framework is generic, language independent, integrated and contains a freely available implementation, the mwetoolkit. It is composed of independent modules which may themselves use multiple techniques to solve a specific sub-task in MWE acquisition. The evaluation of MWE acquisition is modelled using four independent axes. We underline that the evaluation results depend on parameters of the acquisition context, e.g., nature and size of corpora, language and type of MWE, analysis depth, and existing resources. The second main contribution of this thesis is the application-oriented evaluation of our methodology proposal in two applications: computer-assisted lexicography and statistical machine translation. For the former, we evaluate the usefulness of automatic MWE acquisition with the mwetoolkit for creating three lexicons: Greek nominal expressions, Portuguese complex predicates and Portuguese sentiment expressions. For the latter, we test several integration strategies in order to improve the treatment given to English phrasal verbs when translated by a standard statistical MT system into Portuguese. Both applications can benefit from automatic MWE acquisition, as the expressions acquired automatically from corpora can both speed up and improve the quality of the results. The promising results of previous and ongoing experiments encourage further investigation about the optimal way to integrate MWE treatment into other applications. Thus, we conclude the thesis with an overview of the past, ongoing and future work.
Keyword: Computational linguistics; Corpus linguistics; Lexical acquisition; Lexicography; Linguagem natural; Linguística computacional; Machine translation; Multiword expressions; Natural language processing
URL: http://hdl.handle.net/10183/65777
BASE
Hide details
3
Extração de expressões multipalavra em corpora técnicos ; Extraction of multiword expressions in technical domains
BASE
Show details
4
Extração de expressões multipalavra em corpora técnicos ; Extraction of multiword expressions in technical domains
BASE
Show details
5
Aprimorando o tratamento de expressões multipalavras em um tradutor automatico baseado em regras ; Improving the multiword expression treatment in a rule-based machine translator
BASE
Show details
6
Aprimorando o tratamento de expressões multipalavras em um tradutor automatico baseado em regras ; Improving the multiword expression treatment in a rule-based machine translator
BASE
Show details
7
A verb learning model driven by syntactic constructions ; Um modelo de aquisição de verbos guiado por construções sintáticas
BASE
Show details
8
A verb learning model driven by syntactic constructions ; Um modelo de aquisição de verbos guiado por construções sintáticas
BASE
Show details

Catalogues
0
0
0
0
0
0
0
Bibliographies
0
0
0
0
0
0
0
0
0
Linked Open Data catalogues
0
Online resources
0
0
0
0
Open access documents
8
0
0
0
0
© 2013 - 2024 Lin|gu|is|tik | Imprint | Privacy Policy | Datenschutzeinstellungen ändern