DE eng

Search in the Catalogues and Directories

Hits 1 – 20 of 20

1
Bootstrapping Techniques for Polysynthetic Morphological Analysis ...
Lane, William; Bird, Steven. - : arXiv, 2020
BASE
Show details
2
Bayesian phylogenetics, sequence alignment and the genetic structure of the Kainji languages. ...
Bacon, Geoff; Bird, Steven. - : Monash University, 2016
BASE
Show details
3
Learning Crosslingual Word Embeddings without Bilingual Corpora ...
BASE
Show details
4
Language Preservation 2.0: Crowdsourcing oral language documentation using mobile devices
Bird, Steven. - 2015
BASE
Show details
5
Practical Natural Language Processing for Low-Resource Languages.
BASE
Show details
6
Language Preservation 2.0: Crowdsourcing oral language documentation using mobile devices
Bird, Steven. - 2015
BASE
Show details
7
Tone in Usarufa: Field Recordings
Bird, Steven. - 2011
BASE
Show details
8
Equipping university students to document their ancestral languages
BASE
Show details
9
Curating lexical databases for minority languages
BASE
Show details
10
OLAC: Accessing the world's language resources
BASE
Show details
11
OLAC: Accessing the world's language resources
BASE
Show details
12
Curating lexical databases for minority languages
Abstract: One of the biggest challenges in compiling a dictionary of a minority language is managing the large quantity of lexical data. Decisions about the format and content of the dictionary or the orthography typically evolve over the years that such projects usually take. This results in inconsistencies between older and newer entries. Revising the data for publication as a dictionary introduces further inconsistencies as does having multiple contributors and/or editors. Proofreading a lexical database takes a great deal of time and the richer its structure the more this is the case. The tools described in this presentation significantly reduce this effort. Tools developed for checking the consistency of the lexical database in the Iu Mien—Chinese—English dictionary project have proven extremely helpful. Two basic approaches are used: 1) use of a program written to check for likely errors that scans the lexical database and produces an error report that is used by a lexicographer to make appropriate corrections. 2) outputting the lexical data in alternate forms that make it easier for the lexicographer to spot problem areas. These alternative forms include the reverse indexes and views structured according to semantic domains. The Iu Mien—Chinese—English dictionary project, like many minority language dictionary projects, uses SIL's Toolbox software. It is very flexible software but its capabilities to enforce consistency are quite limited. Some parts of the approach described here are specific to MDF (Multi-Dictionary Formatter) lexical databases in Toolbox but will be equally useful for other MDF databases. Other parts are specific to each of the three languages involved but will be useful for non-Toolbox lexical databases. Every dictionary is unique and this applies not only to content of the entries but also the decisions about how entries should be arranged to suit the languages involved. Other decisions about the structure are likely to be made differently even in other dictionaries of the same languages. It is the way that each dictionary combines themes that are found in many dictionaries that makes them unique, e.g. to be root based or not, to have include subentries. Therefore our approach is to use a toolkit based approach to curating lexical databases. This allows checking techniques to be mixed and matched to suit the unique aspects of a lexical project. The checking software is written in Python and relies on the toolbox module in NLTK (The Natural Language Toolkit http://nltk.sourceforge.net).
URL: http://hdl.handle.net/10125/5094
BASE
Hide details
13
Grid-Enabling Natural Language Engineering By Stealth ...
Hughes, Baden; Bird, Steven. - : arXiv, 2003
BASE
Show details
14
A Grid Based Architecture for High-Performance NLP ...
Hughes, Baden; Bird, Steven. - : arXiv, 2003
BASE
Show details
15
Grassfields Bantu Fieldwork: Dschang Lexicon ...
Bird, Steven. - : Linguistic Data Consortium, 2003
BASE
Show details
16
NLTK: The Natural Language Toolkit ...
Loper, Edward; Bird, Steven. - : arXiv, 2002
BASE
Show details
17
A Formal Framework for Linguistic Annotation (revised version) ...
Bird, Steven; Liberman, Mark. - : arXiv, 2000
BASE
Show details
18
Annotation graphs as a framework for multidimensional linguistic data analysis ...
Bird, Steven; Liberman, Mark. - : arXiv, 1999
BASE
Show details
19
A Formal Framework for Linguistic Annotation ...
Bird, Steven; Liberman, Mark. - : arXiv, 1999
BASE
Show details
20
Key aspects of declarative phonology
Bird, Steven; Coleman, John S.; Scobbie, James M.. - : European Studies Research Institute, University of Salford., 1996
BASE
Show details

Catalogues
0
0
0
0
0
0
0
Bibliographies
0
0
0
0
0
0
0
0
0
Linked Open Data catalogues
0
Online resources
0
0
0
0
Open access documents
20
0
0
0
0
© 2013 - 2024 Lin|gu|is|tik | Imprint | Privacy Policy | Datenschutzeinstellungen ändern