Extension of the remos concept to frequency-filtering-based features for reverberation-robust speech recognition

Maas R, Wolf M, Sehr A, Nadeu C, Kellermann W (2011)

Publication Language: English

Publication Status: Published

Publication Type: Conference contribution, Conference Contribution

Publication year: 2011

Pages Range: 13-16

Article Number: 5942381

Event location: Edinburgh

ISBN: 9781457709999

DOI: 10.1109/HSCMA.2011.5942381

Abstract

The introduction of partly decorrelated features into the REMOS (REverberationMOdeling for Speech recognition) concept for distant-talking speech recognition [1] is discussed. REMOS combines a hidden Markov model (HMM), trained on clean speech, with a reverberation model capturing certain room characteristics. The most likely contributions of both models to a reverberant observation are determined by an inner optimization problem. In HMM frameworks, decorrelated features are assumed when diagonal covariance matrices are used in the output densities. However, in REMOS, only highly correlated logmelspec (logarithmic mel-spectral) features have been used so far, which has been limiting the recognition performance. In this work, we extend the RE-MOS concept and introduce a new set of partly decorrelated features derived from the frequency filtering [2]. Recognition experiments with connected digits show a consistent relative reduction in word error rate of up to 29% compared to the former logmelspec implementation. © 2011 IEEE.

Authors with CRIS profile

Roland Maas Lehrstuhl für Multimediakommunikation und Signalverarbeitung Armin Sehr Professur für Signalverarbeitung Walter Kellermann Professur für Signalverarbeitung

Involved external institutions

Universitat Politècnica de Catalunya · BarcelonaTech (UPC)

Spain (ES)

How to cite

APA:

Maas, R., Wolf, M., Sehr, A., Nadeu, C., & Kellermann, W. (2011). Extension of the remos concept to frequency-filtering-based features for reverberation-robust speech recognition. In Proceedings of the 2011 Joint Workshop on Hands-free Speech Communication and Microphone Arrays, HSCMA'11 (pp. 13-16). Edinburgh, GB.

MLA:

Maas, Roland, et al. "Extension of the remos concept to frequency-filtering-based features for reverberation-robust speech recognition." Proceedings of the 2011 Joint Workshop on Hands-free Speech Communication and Microphone Arrays, HSCMA'11, Edinburgh 2011. 13-16.

BibTeX: Download