Using semantic analysis to improve speech recognition performance

Erdoğan, Hakan and Sarıkaya, Ruhi and Chen, Stanley F. and Gao, Yuqing and Picheny, Michael (2005) Using semantic analysis to improve speech recognition performance. Computer Speech & Language, 19 (3). pp. 321-343. ISSN 0885-2308

[thumbnail of 3011800001082.pdf] PDF
Restricted to Repository staff only

Download (248kB) | Request a copy


Although syntactic structure has been used in recent work in language modeling, there has not been much effort in using semantic analysis for language models. In this study, we propose three new language modeling techniques that use semantic analysis for spoken dialog systems. We call these methods concept sequence modeling, two-level semantic-lexical modeling, and joint semantic-lexical modeling. These models combine lexical information with varying amounts of semantic information, using annotation supplied by either a shallow semantic parser or full hierarchical parser. These models also differ in how the lexical and semantic information is combined, ranging from simple interpolation to tight integration using maximum entropy modeling. We obtain improvements in recognition accuracy over word and class N-gram language models in three different task domains. Interpolation of the proposed models with class N-gram language models provides additional improvement in the air travel reservation domain. We show that as we increase the semantic information utilized and as we increase the tightness of integration between lexical and semantic items, we obtain improved performance when interpolating with class language models, indicating that the two types of models become more complementary in nature.
Item Type: Article
Subjects: Q Science > QA Mathematics > QA075 Electronic computers. Computer science
Divisions: Faculty of Engineering and Natural Sciences
Depositing User: Hakan Erdoğan
Date Deposited: 19 Oct 2005 03:00
Last Modified: 25 May 2011 14:03

Actions (login required)

View Item
View Item