Spoken document retrieval with unsupervised query modeling techniques

Berlin Chen*, Kuan Yu Chen, Pei Ning Chen, Yi Wen Chen

*此作品的通信作者

研究成果: 雜誌貢獻期刊論文同行評審

25 引文 斯高帕斯(Scopus)

摘要

Ever-increasing amounts of publicly available multimedia associated with speech information have motivated spoken document retrieval (SDR) to be an active area of intensive research in the speech processing community. Much work has been dedicated to developing elaborate indexing and modeling techniques for representing spoken documents, but only little to improving query formulations for better representing the information needs of users. The latter is critical to the success of a SDR system. In view of this, we present in this paper a novel use of a relevance language modeling framework for SDR. It not only inherits the merits of several existing techniques but also provides a principled way to render the lexical and topical relationships between a query and a spoken document. We further explore various ways to glean both relevance and non-relevance cues from the spoken document collection so as to enhance query modeling in an unsupervised fashion. In addition, we also investigate representing the query and documents with different granularities of index features to work in conjunction with the various relevance and/or non-relevance cues. Empirical evaluations performed on the TDT (Topic Detection and Tracking) collections reveal that the methods derived from our modeling framework hold good promise for SDR and are very competitive with existing retrieval methods.

原文英語
文章編號6239571
頁(從 - 到)2602-2612
頁數11
期刊IEEE Transactions on Audio, Speech and Language Processing
20
發行號9
DOIs
出版狀態已發佈 - 2012

ASJC Scopus subject areas

  • 聲學與超音波
  • 電氣與電子工程

指紋

深入研究「Spoken document retrieval with unsupervised query modeling techniques」主題。共同形成了獨特的指紋。

引用此