We gratefully acknowledge support from
the Simons Foundation and member institutions.
Full-text links:

Download:

Current browse context:

cs.CL

Change to browse by:

References & Citations

DBLP - CS Bibliography

Bookmark

(what is this?)
CiteULike logo BibSonomy logo Mendeley logo del.icio.us logo Digg logo Reddit logo ScienceWISE logo

Computer Science > Computation and Language

Title: GIPFA: Generating IPA Pronunciation from Audio

Authors: Xavier Marjou
Abstract: Transcribing spoken audio samples into the International Phonetic Alphabet (IPA) has long been reserved for experts. In this study, we examine the use of an Artificial Neural Network (ANN) model to automatically extract the IPA phonemic pronunciation of a word based on its audio pronunciation, hence its name Generating IPA Pronunciation From Audio (GIPFA). Based on the French Wikimedia dictionary, we trained our model which then correctly predicted 75% of the IPA pronunciations tested. Interestingly, by studying inference errors, the model made it possible to highlight possible errors in the dataset as well as to identify the closest phonemes in French.
Comments: 10 pages, 2 figures, 7 tables
Subjects: Computation and Language (cs.CL); Sound (cs.SD); Audio and Speech Processing (eess.AS)
Journal reference: Proceedings of the eLex 2021 conference, page 588
Cite as: arXiv:2006.07573 [cs.CL]
  (or arXiv:2006.07573v2 [cs.CL] for this version)

Submission history

From: Xavier Marjou [view email]
[v1] Sat, 13 Jun 2020 06:14:11 GMT (46kb,D)
[v2] Tue, 21 Sep 2021 19:53:39 GMT (76kb,D)

Link back to: arXiv, form interface, contact.