Sixth European Conference on Speech Communication and Technology

Budapest, Hungary
September 5-9, 1999

Semi-Automatic Acquisition of Domain-Specific Semantic Structures

Kai-Chung Siu, Helen M. Meng

Human-Computer Communications Laboratory, Department of Systems Engineering and Engineering Management, The Chinese University of Hong Kong, Shatin, N.T., Hong Kong

This paper describes a methodology for semi-automatic grammar induction from unannotated corpora belonging to a restricted domain. The grammar contains both semantic and syntactic structures, which are conducive towards language understanding. Our work aims to ameliorate the reliance of grammar development on expert handcrafting or the availability of annotated corpora. To strive for a reasonable model for real data, as well as portability across domain and languages, we adopt a statistical approach. Our approach is also amenable to the optional injection of prior knowledge to aid grammar induction, and subsequent hand editing for grammar refinement. This constitutes the semi-automatic nature of the approach. Experiments with the ATIS corpus showed positive results in semantic parsing, when compared to an entirely handcrafted grammar.

Keywords: grammar induction, semantic processing.

Full Paper (PDF)   Gnu-Zipped Postscript

Bibliographic reference.  Siu, Kai-Chung / Meng, Helen M. (1999): "Semi-automatic acquisition of domain-specific semantic structures", In EUROSPEECH'99, 2039-2042.