We gratefully acknowledge support from
the Simons Foundation and member institutions.
Full-text links:

Download:

Current browse context:

eess.AS

Change to browse by:

References & Citations

Bookmark

(what is this?)
CiteULike logo BibSonomy logo Mendeley logo del.icio.us logo Digg logo Reddit logo

Electrical Engineering and Systems Science > Audio and Speech Processing

Title: Robust Raw Waveform Speech Recognition Using Relevance Weighted Representations

Abstract: Speech recognition in noisy and channel distorted scenarios is often challenging as the current acoustic modeling schemes are not adaptive to the changes in the signal distribution in the presence of noise. In this work, we develop a novel acoustic modeling framework for noise robust speech recognition based on relevance weighting mechanism. The relevance weighting is achieved using a sub-network approach that performs feature selection. A relevance sub-network is applied on the output of first layer of a convolutional network model operating on raw speech signals while a second relevance sub-network is applied on the second convolutional layer output. The relevance weights for the first layer correspond to an acoustic filterbank selection while the relevance weights in the second layer perform modulation filter selection. The model is trained for a speech recognition task on noisy and reverberant speech. The speech recognition experiments on multiple datasets (Aurora-4, CHiME-3, VOiCES) reveal that the incorporation of relevance weighting in the neural network architecture improves the speech recognition word error rates significantly (average relative improvements of 10% over the baseline systems)
Comments: arXiv admin note: text overlap with arXiv:2001.07067
Subjects: Audio and Speech Processing (eess.AS); Sound (cs.SD)
Journal reference: Proc. Interspeech 2020, 1649-1653 (2020)
DOI: 10.21437/Interspeech.2020-2301
Cite as: arXiv:2011.00721 [eess.AS]
  (or arXiv:2011.00721v1 [eess.AS] for this version)

Submission history

From: Purvi Agrawal [view email]
[v1] Thu, 29 Oct 2020 19:32:50 GMT (1241kb,D)

Link back to: arXiv, form interface, contact.