Current browse context:
cs
Change to browse by:
References & Citations
Computer Science > Sound
Title: Using instantaneous frequency and aperiodicity detection to estimate F0 for high-quality speech synthesis
(Submitted on 25 May 2016 (v1), last revised 22 Jul 2016 (this version, v2))
Abstract: This paper introduces a general and flexible framework for F0 and aperiodicity (additive non periodic component) analysis, specifically intended for high-quality speech synthesis and modification applications. The proposed framework consists of three subsystems: instantaneous frequency estimator and initial aperiodicity detector, F0 trajectory tracker, and F0 refinement and aperiodicity extractor. A preliminary implementation of the proposed framework substantially outperformed (by a factor of 10 in terms of RMS F0 estimation error) existing F0 extractors in tracking ability of temporally varying F0 trajectories. The front end aperiodicity detector consists of a complex-valued wavelet analysis filter with a highly selective temporal and spectral envelope. This front end aperiodicity detector uses a new measure that quantifies the deviation from periodicity. The measure is less sensitive to slow FM and AM and closely correlates with the signal to noise ratio.
Submission history
From: Hideki Kawahara [view email][v1] Wed, 25 May 2016 10:20:07 GMT (476kb)
[v2] Fri, 22 Jul 2016 20:56:20 GMT (1098kb)
Link back to: arXiv, form interface, contact.