Home
Catalogue search
Refine your search:
Keyword:
*AFGHANISTAN (2)
*ENGLISH LANGUAGE (2)
*FOREIGN LANGUAGES (2)
*MACHINE TRANSLATION (2)
Linguistics (2)
*AUTOMATION (1)
*BILINGUAL DATA PRODUCTION (1)
*BILINGUAL PARALLEL TEXT (1)
*CRIMINAL JUSTICE SYSTEM (1)
*DARI (1)
more
Creator / Publisher:
ARMY RESEARCH LAB ADELPHI MD COMPUTATIONAL AND INFORMATION SCIENCES DIRECTORATE (2)
Jahed, Ghulam H (1)
LaRocca, Steven A (1)
Morgan, John J (1)
Tanenbaum, Will (1)
Vanni, Michelle (1)
Year:
2012 (1)
2011 (1)
Medium:
Online (2)
Type:
Article (2)
BLLDB-Access:
free (2)
subject to license (0)
Search in the Catalogues and Directories
All fields
Title
Creator / Publisher
Keyword
Year
AND
OR
AND NOT
All fields
Title
Creator / Publisher
Keyword
Year
AND
OR
AND NOT
All fields
Title
Creator / Publisher
Keyword
Year
AND
OR
AND NOT
All fields
Title
Creator / Publisher
Keyword
Year
AND
OR
AND NOT
All fields
Title
Creator / Publisher
Keyword
Year
Sort by
creator [A → Z]
'
creator [Z → A]
'
publishing year ↑ (asc)
'
publishing year ↓ (desc)
'
title [A → Z]
'
title [Z → A]
'
Simple Search
Hits 1 – 2 of 2
1
Method to Select Technical Terms for Glossaries in Support of Joint Task Force Operations
Vanni, Michelle
In: DTIC (2012)
BASE
Show details
2
Introduction of Automation for the Production of Bilingual, Parallel-Aligned Text
Tanenbaum, Will
;
LaRocca, Steven A
;
Morgan, John J
;
Jahed, Ghulam H
In: DTIC (2011)
Abstract:
As the study and application of statistical machine translation (SMT) grows, progress is often circumscribed by a lack of data. The statistical models that govern statistical machine translation (SMT) engines rely on many large bilingual text corpora, each comprised of vast numbers of bilingual text segments. For certain languages, corpora already exist and help to power translation engines. Regrettably, this is not the case for every language the Army is interested in, making the creation or acquisition of such data a priority. To this end, a language expert in Dari and Pashto was hired, who collected, prepared, and ensured the quality of bilingual text. To explore ways in which to aid the expert, a variety of the steps performed by the expert and necessary to the process were automated. The hypothesis was that automation of selected processes would improve efficiency, measured in terms of both speed of production and quantity of data produced, even when time to correct automation-caused errors was accounted for. As predicted, the net result of introducing automation was an increase in both the rate of producing correct bilingual segments and the number produced. The implications of these results for improving larger bilingual data creation and acquisition efforts are discussed. ; The original document contains color images.
Keyword:
*AFGHANISTAN
;
*AUTOMATION
;
*BILINGUAL DATA PRODUCTION
;
*BILINGUAL PARALLEL TEXT
;
*DATA MINING
;
*ENGLISH LANGUAGE
;
*FOREIGN LANGUAGES
;
*MACHINE TRANSLATION
;
*STATISTICAL ANALYSIS
;
*STATISTICAL MACHINE TRANSLATION
;
ACCURACY
;
ALIGNMENT
;
ARMY OPERATIONS
;
ARMY PERSONNEL
;
Computer Programming and Software
;
DARI LANGUAGE
;
DARI-ENGLISH TRANSLATION
;
EFFICIENCY
;
Linguistics
;
NATURAL LANGUAGE
;
PARSERS
;
PASHTO LANGUAGE
;
PASHTO-ENGLISH TRANSLATION
;
PIPELINE PROJECT
;
PRODUCTION
;
SEGMENTATION
;
SOFTWARE TOOLS
;
Statistics and Probability
URL:
http://www.dtic.mil/docs/citations/ADA552756
http://oai.dtic.mil/oai/oai?&verb=getRecord&metadataPrefix=html&identifier=ADA552756
BASE
Hide details
Mobile view
All
Catalogues
UB Frankfurt Linguistik
0
IDS Mannheim
0
OLC Linguistik
0
UB Frankfurt Retrokatalog
0
DNB Subject Category Language
0
Institut für Empirische Sprachwissenschaft
0
Leibniz-Centre General Linguistics (ZAS)
0
Bibliographies
BLLDB
0
BDSL
0
IDS Bibliografie zur deutschen Grammatik
0
IDS Bibliografie zur Gesprächsforschung
0
IDS Konnektoren im Deutschen
0
IDS Präpositionen im Deutschen
0
IDS OBELEX meta
0
MPI-SHH Linguistics Collection
0
MPI for Psycholinguistics
0
Linked Open Data catalogues
Annohub
0
Online resources
Link directory
0
Journal directory
0
Database directory
0
Dictionary directory
0
Open access documents
BASE
2
Linguistik-Repository
0
IDS Publikationsserver
0
Online dissertations
0
Language Description Heritage
0
© 2013 - 2024 Lin|gu|is|tik
|
Imprint
|
Privacy Policy
|
Datenschutzeinstellungen ändern