ViSQOL v3: An Open Source Production Ready Objective Speech and Audio Metric

Chinen, Michael; Lim, Felicia S. C.; Skoglund, Jan; Gureev, Nikita; O'Gorman, Feargus; Hines, Andrew

Full-text links:

Download:

Current browse context:

eess.AS

< prev | next >

new | recent | 2004

Electrical Engineering and Systems Science > Audio and Speech Processing

Title: ViSQOL v3: An Open Source Production Ready Objective Speech and Audio Metric

Authors: Michael Chinen, Felicia S. C. Lim, Jan Skoglund, Nikita Gureev, Feargus O'Gorman, Andrew Hines

(Submitted on 20 Apr 2020)

Abstract: Estimation of perceptual quality in audio and speech is possible using a variety of methods. The combined v3 release of ViSQOL and ViSQOLAudio (for speech and audio, respectively,) provides improvements upon previous versions, in terms of both design and usage. As an open source C++ library or binary with permissive licensing, ViSQOL can now be deployed beyond the research context into production usage. The feedback from internal production teams at Google has helped to improve this new release, and serves to show cases where it is most applicable, as well as to highlight limitations. The new model is benchmarked against real-world data for evaluation purposes. The trends and direction of future work is discussed.

Comments:	2020 Twelfth International Conference on Quality of Multimedia Experience (QoMEX)
Subjects:	Audio and Speech Processing (eess.AS); Sound (cs.SD); Signal Processing (eess.SP)
Cite as:	arXiv:2004.09584 [eess.AS]
	(or arXiv:2004.09584v1 [eess.AS] for this version)

Submission history

From: Michael Chinen [view email]
[v1] Mon, 20 Apr 2020 19:19:26 GMT (1156kb,D)

Which authors of this paper are endorsers? | Disable MathJax (What is MathJax?)

Link back to: arXiv, form interface, contact.

> eess > arXiv:2004.09584

Download:

Current browse context:

Change to browse by:

References & Citations

Bookmark

Electrical Engineering and Systems Science > Audio and Speech Processing

Title: ViSQOL v3: An Open Source Production Ready Objective Speech and Audio Metric

Submission history