Interspeech'2005 - Eurospeech

Lisbon, Portugal
September 4-8, 2005

Development of a Conversational Telephone Speech Recognizer for Levantine Arabic

Dimitra Vergyri (1), Katrin Kirchhoff (2), R. Gadde (1), Andreas Stolcke (1), Jing Zheng (1)

(1) SRI International, USA; (2) University of Washington, USA

Many languages, including Arabic, are characterized by a wide variety of different dialects that often differ strongly from each other. When developing speech technology for dialect-rich languages, the portability and reusability of data, algorithms, and systemcomponents becomes extremely important. In this paper, we describe the development of a large-vocabulary speech recognition system for Levantine Arabic, which was a new dialectal recognition task for our existing system. We discuss the dialect-specific modeling choices (grapheme vs. phoneme based acoustic models, automatic vowelization techniques, and morphological language models) and investigate to what extent techniques previously tested on other languages are portable to the present task. We present state-of-the-art recognition results on the 2004 Levantine Arabic Rich Transcription evaluation.

Full Paper

Bibliographic reference.  Vergyri, Dimitra / Kirchhoff, Katrin / Gadde, R. / Stolcke, Andreas / Zheng, Jing (2005): "Development of a conversational telephone speech recognizer for Levantine Arabic", In INTERSPEECH-2005, 1613-1616.