Machine Translation for Triage and Exploitation of Massive Text Data
James E. Andrews, Kristen Summers · 2008
The National Ground Intelligence Center (NGIC) collects massive quantities of textual data in foreign languages. To support exploi-tation in light of intelligence requirements, a triage process must be applied to this data as those requirements emerge, to identify the most useful data for further exploitation. Ma-chine translation provides critical support for this triage. This paper outlines the types of collected data and the different challenges they present for machine translation, as well as the types of triage to support for collections of this nature, and the issues raised for ma-chine translation by those uses. 1