Improved Spoken Term Detection by Feature Space Pseudo-Relevance Feedback
Journal
Proceedings of the 11th Annual Conference of the International Speech Communication Association, INTERSPEECH 2010
Pages
1672-1675
Start Page
1672
End Page
1675
Date Issued
2010-09
Author(s)
Abstract
In this paper, we propose an improved approach for spoken term detection using pseudo-relevance feedback. To remedy the problem of unmatched acoustic models with respect to spoken utterances produced under different acoustic conditions, which may give relatively poor recognition output, we integrate the relevance scores derived from the lattices with the DTW distances derived from the feature space of MFCC parameters or phonetic posteriorgrams. These DTW distances are evaluated for a carefully selected set of pseudo-relevant utterances, which obtained from the first-pass returned list given by the search engine. The utterances on the first-pass returned list are then reranked accordingly and finally shown to the user. Very encouraging, performance improvements were obtained in the preliminary experiments, especially when the acoustic models are poorly matched to the spoken utterances.
Event(s)
11th Annual Conference of the International Speech Communication Association, INTERSPEECH 2010
Subjects
Pseudo-relevance feedback
Spoken term detection
Publisher
International Speech Communication Association
Type
conference paper
