Extension of the remos concept to frequency-filtering-based features for reverberation-robust speech recognition

Maas, Roland; Wolf, Martin; Sehr, Armin; Nadeu Camprubí, Climent; Kellermann, Walter

doi:10.1109/HSCMA.2011.5942381

Visualitza/Obre

Article (283,7Kb) (Accés restringit) Sol·licita una còpia a l'autor

Veure estadístiques d'ús d'UPCommons

Estadístiques de LA Referencia / Recolecta

Cita com:

Mostra el registre d'ítem complet

Maas, Roland

Wolf, Martin

Sehr, Armin

Nadeu Camprubí, Climent

Kellermann, Walter

Tipus de documentText en actes de congrés

Data publicació2011

Condicions d'accésAccés restringit per política de l'editorial

Tots els drets reservats. Aquesta obra està protegida pels drets de propietat intel·lectual i industrial corresponents. Sense perjudici de les exempcions legals existents, queda prohibida la seva reproducció, distribució, comunicació pública o transformació sense l'autorització del titular dels drets

Abstract

The introduction of partly decorrelated features into the REMOS (REverberationMOdeling for Speech recognition) concept for distant-talking speech recognition [1] is discussed. REMOS combines a hidden Markov model (HMM), trained on clean speech, with a reverberation model capturing certain room characteristics. The most likely contributions of both models to a reverberant observation are determined by an inner optimization problem. In HMM frameworks, decorrelated features are assumed when diagonal covariance matrices are used in the output densities. However, in REMOS, only highly correlated logmelspec (logarithmic mel-spectral) features have been used so far, which has been limiting the recognition performance. In this work, we extend the RE-MOS concept and introduce a new set of partly decorrelated features derived from the frequency filtering [2]. Recognition experiments with connected digits show a consistent relative reduction in word error rate of up to 29% compared to the former logmelspec implementation.

CitacióMaas, R. [et al.]. Extension of the remos concept to frequency-filtering-based features for reverberation-robust speech recognition. A: Workshop on Hands-free Speech Communication and Microphone Arrays (HSCMA),. "2011 Joint Workshop on Hands-free Speech Communication and Microphone Arrays". Edinburgh: 2011, p. 13-18.

URIhttp://hdl.handle.net/2117/15468

DOI10.1109/HSCMA.2011.5942381

Versió de l'editorhttp://ieeexplore.ieee.org/xpls/abs_all.jsp?arnumber=5942381&tag=1

Col·leccions

Veure estadístiques d'ús d'UPCommons

Mostra el registre d'ítem complet

Fitxers	Descripció	Mida	Format	Visualitza
HSCMA2011.pdf	Article	283,7Kb	PDF	Accés restringit

UPCommons. Portal del coneixement obert de la UPC

Extension of the remos concept to frequency-filtering-based features for reverberation-robust speech recognition

Visualitza/Obre

Explora