Robust ASR using neural network based speech enhancement and feature simulation - INRIA - Institut National de Recherche en Informatique et en Automatique Accéder directement au contenu
Communication Dans Un Congrès Année : 2015

Robust ASR using neural network based speech enhancement and feature simulation

Résumé

We consider the problem of robust automatic speech recognition (ASR) in the context of the CHiME-3 Challenge. The proposed system combines three contributions. First, we propose a deep neural network (DNN) based multichannel speech enhancement technique, where the speech and noise spectra are estimated using a DNN based regressor and the spatial parameters are derived in an expectation-maximization (EM) like fashion. Second, a conditional restricted Boltz-mann machine (CRBM) model is trained using the obtained enhanced speech and used to generate simulated training and development datasets. The goal is to increase the similarity between simulated and real data, so as to increase the benefit of multicondition training. Finally, we make some changes to the ASR backend. Our system ranked 4th among 25 entries
Fichier principal
Vignette du fichier
INRIA.pdf (388.81 Ko) Télécharger le fichier
Origine : Fichiers produits par l'(les) auteur(s)
Loading...

Dates et versions

hal-01204553 , version 1 (24-09-2015)

Identifiants

  • HAL Id : hal-01204553 , version 1

Citer

Sunit Sivasankaran, Aditya A Nugraha, Emmanuel Vincent, Juan Andrés Morales Cordovilla, Siddharth Dalmia, et al.. Robust ASR using neural network based speech enhancement and feature simulation. ASRU, Dec 2015, Arizona, United States. ⟨hal-01204553⟩
532 Consultations
1379 Téléchargements

Partager

Gmail Facebook X LinkedIn More