Show simple item record

IEEE Transactions on Audio, Speech and Language Processing

dc.contributor.authorDielmann, Alfred
dc.contributor.authorRenals, Steve
dc.date.accessioned2010-10-04T14:29:37Z
dc.date.available2010-10-04T14:29:37Z
dc.date.issued2008
dc.identifier.urihttp://ieeexplore.ieee.org/xpls/abs_all.jsp?isnumber=4599391&arnumber=4497831&count=18&index=9en
dc.identifier.urihttp://hdl.handle.net/1842/3823
dc.description.abstractThis paper is concerned with the automatic recognition of dialogue acts (DAs) in multiparty conversational speech. We present a joint generative model for DA recognition in which segmentation and classification of DAs are carried out in parallel. Our approach to DA recognition is based on a switching dynamic Bayesian network (DBN) architecture. This generative approach models a set of features, related to lexical content and prosody, and incorporates a weighted interpolated factored language model. The switching DBN coordinates the recognition process by integrating the component models. The factored language model, which is estimated from multiple conversational data corpora, is used in conjunction with additional task-specific language models. In conjunction with this joint generative model, we have also investigated the use of a discriminative approach, based on conditional random fields, to perform a reclassification of the segmented DAs. We have carried out experiments on the AMI corpus of multimodal meeting recordings, using both manually transcribed speech, and the output of an automatic speech recognizer, and using different configurations of the generative model. Our results indicate that the system performs well both on reference and fully automatic transcriptions. A further significant improvement in recognition accuracy is obtained by the application of the discriminative reranking approach based on conditional random fields.en
dc.titleRecognition of Dialogue Acts in Multiparty Meetings using a Switching DBNen
dc.typeArticleen
dc.identifier.doi10.1109/TASL.2008.922463en
rps.issue7en
rps.volume16en
rps.titleIEEE Transactions on Audio, Speech and Language Processingen
dc.extent.pageNumbers1303--1314en
dc.date.updated2010-10-04T14:29:38Z


Files in this item

This item appears in the following Collection(s)

Show simple item record