Term-Dependent Confidence Normalisation for Out-of-Vocabulary Spoken Term Detection

Dong Wang; Javier Tejedor; Simon King; Joe Frankel

doi:10.1007/s11390-012-1228-x

Dong Wang, Javier Tejedor, Simon King, Joe Frankel. Term-Dependent Confidence Normalisation for Out-of-Vocabulary Spoken Term DetectionJ. Journal of Computer Science and Technology, 2012, (2): 358-375. DOI: 10.1007/s11390-012-1228-x

Citation:

Term-Dependent Confidence Normalisation for Out-of-Vocabulary Spoken Term Detection

Abstract

Abstract

An important component of a spoken term detection (STD) system involves estimating confidence measures of hypothesised detections. A potential problem of the widely used lattice-based confidence estimation, however, is that the confidence scores are treated uniformly for all search terms, regardless of how much they may differ in terms of phonetic or linguistic properties. This problem is particularly evident for out-of-vocabulary (OOV) terms which tend to exhibit high intra-term diversity. To address the impact of term diversity on confidence measures, we propose in this work a term-dependent normalisation technique which compensates for term diversity in confidence estimation. We first derive an evaluation-metric-oriented normalisation that optimises the evaluation metric by compensating for the diverse occurrence rates among terms, and then propose a linear bias compensation and a discriminative compensation to deal with the bias problem that is inherent in lattice-based confidence measurement and from which the Term Specific Threshold (TST) approach suffers. We tested the proposed technique on speech data from the multi-party meeting domain with two state-of-the-art STD systems based on phonemes and words respectively. The experimental results demonstrate that the confidence normalisation approach leads to a significant performance improvement in STD, particularly for OOV terms with phoneme-based systems.

FullText(HTML)

References (68)

Relative Articles

Supplements (0)

Cited By

Term-Dependent Confidence Normalisation for Out-of-Vocabulary Spoken Term Detection

Abstract

Catalog

Export File

Citation

Format

Content