Unified Modeling of Multi-Domain Multi-Device ASR Systems

Mitra, Soumyajit; Ray, Swayambhu Nath; Padi, Bharat; Sen, Arunasish; Bilgi, Raghavendra; Arsikere, Harish; Ghosh, Shalini; Srinivasamurthy, Ajay; Garimella, Sri

Full-text links:

Download:

Source

Computer Science > Computation and Language

Title: Unified Modeling of Multi-Domain Multi-Device ASR Systems

Authors: Soumyajit Mitra, Swayambhu Nath Ray, Bharat Padi, Arunasish Sen, Raghavendra Bilgi, Harish Arsikere, Shalini Ghosh, Ajay Srinivasamurthy, Sri Garimella

(Submitted on 13 May 2022 (v1), last revised 13 Oct 2022 (this version, v3))

Abstract: Modern Automatic Speech Recognition (ASR) systems often use a portfolio of domain-specific models in order to get high accuracy for distinct user utterance types across different devices. In this paper, we propose an innovative approach that integrates the different per-domain per-device models into a unified model, using a combination of domain embedding, domain experts, mixture of experts and adversarial training. We run careful ablation studies to show the benefit of each of these innovations in contributing to the accuracy of the overall unified model. Experiments show that our proposed unified modeling approach actually outperforms the carefully tuned per-domain models, giving relative gains of up to 10% over a baseline model with negligible increase in the number of parameters.

Comments:	We will update the paper completely with our latest experiments and analysis
Subjects:	Computation and Language (cs.CL); Sound (cs.SD); Audio and Speech Processing (eess.AS)
Cite as:	arXiv:2205.06655 [cs.CL]
	(or arXiv:2205.06655v3 [cs.CL] for this version)

Submission history

From: Swayambhu Nath Ray [view email]
[v1] Fri, 13 May 2022 14:07:22 GMT (9158kb,D)
[v2] Thu, 2 Jun 2022 15:35:00 GMT (9156kb,D)
[v3] Thu, 13 Oct 2022 16:15:31 GMT (0kb,I)

Which authors of this paper are endorsers? | Disable MathJax (What is MathJax?)

Link back to: arXiv, form interface, contact.

> cs > arXiv:2205.06655

Download:

Current browse context:

Change to browse by:

References & Citations

DBLP - CS Bibliography

Bookmark

Computer Science > Computation and Language

Title: Unified Modeling of Multi-Domain Multi-Device ASR Systems

Submission history