ODYSSEY 2004 - The Speaker and Language Recognition Workshop

May 31 - June 3, 2004
Toledo, Spain

The MMSR Bilingual and Crosschannel Corpora for Speaker Recognition Research and Evaluation

Joseph P. Campbell (1), Hirotaka Nakasone (2), Christopher Cieri (3), David Miller (3), Kevin Walker (3), Alvin F. Martin (4), Mark A. Przybocki (4)

(1) MIT Lincoln Laboratory, Lexington, MA, USA
(2) Federal Bureau of Investigation, Quantico, VA, USA
(3) University of Pennsylvania, Linguistic Data Consortium, Philadelphia, PA, USA
(4) National Institute of Standards and Technology, Gaithersburg, MD, USA

We describe efforts to create corpora to support and evaluate systems that meet the challenge of speaker recognition in the face of both channel and language variation. In addition to addressing ongoing evaluation of speaker recognition systems, these corpora are aimed at the bilingual and crosschannel dimensions. We report on specific data collection efforts at the Linguistic Data Consortium, the 2004 speaker recognition evaluation program organized by the National Institute of Standards and Technology (NIST), and the research ongoing at the US Federal Bureau of Investigation and MIT Lincoln Laboratory. We cover the design and requirements, the collections and evaluation integrating discussions of the data preparation, research, technology development and evaluation on a grand scale.

Full Paper

Bibliographic reference.  Campbell, Joseph P. / Nakasone, Hirotaka / Cieri, Christopher / Miller, David / Walker, Kevin / Martin, Alvin F. / Przybocki, Mark A. (2004): "The MMSR bilingual and crosschannel corpora for speaker recognition research and evaluation", In ODYS-2004, 29-32.