International Workshop on Spoken Language Translation (IWSLT) 2012

Hong Kong
December 6-7, 2012

Active Error Detection and Resolution for Speech-to-Speech Translation

Rohit Prasad, Rohit Kumar, Sankaranarayanan Ananthakrishnan, Wei Chen, Sanjika Hewavitharana, Matthew Roy, Frederick Choi, Aaron Challenner, Enoch Kan, Arvind Neelakantan, Prem Natarajan

Speech, Language, and Multimedia Business Unit, Raytheon BBN Technologies Cambridge, MA, USA

We describe a novel two-way speech-to-speech (S2S) translation system that actively detects a wide variety of common error types and resolves them through user-friendly dialog with the user(s). We present algorithms for detecting out-of-vocabulary (OOV) named entities and terms, sense ambiguities, homophones, idioms, ill-formed input, etc. and discuss novel, interactive strategies for recovering from such errors. We also describe our approach for prioritizing different error types and an extensible architecture for implementing these decisions. We demonstrate the efficacy of our system by presenting analysis on live interactions in the English-to-Iraqi Arabic direction that are designed to invoke different error types for spoken language translation. Our analysis shows that the system can successfully resolve 47% of the errors, resulting in a dramatic improvement in the transfer of problematic concepts.

Full Paper    Presentation

Bibliographic reference.  Prasad, Rohit / Kumar, Rohit / Ananthakrishnan, Sankaranarayanan / Chen, Wei / Hewavitharana, Sanjika / Roy, Matthew / Choi, Frederick / Challenner, Aaron / Kan, Enoch / Neelakantan, Arvind / Natarajan, Prem (2012): "Active error detection and resolution for speech-to-speech translation", In IWSLT-2012, 150-157.