We gratefully acknowledge support from
the Simons Foundation and member institutions.

Audio and Speech Processing

Authors and titles for eess.AS in Nov 2018, skipping first 150

[ total of 166 entries: 1-50 | 51-100 | 101-150 | 151-166 ]
[ showing 50 entries per page: fewer | more | all ]
[151]  arXiv:1811.10169 (cross-list from cs.CL) [pdf, ps, other]
Title: Improving Gated Recurrent Unit Based Acoustic Modeling with Batch Normalization and Enlarged Context
Comments: ISCSLP 2018
Subjects: Computation and Language (cs.CL); Audio and Speech Processing (eess.AS)
[152]  arXiv:1811.10376 (cross-list from cs.LG) [pdf, other]
Title: Robustness against the channel effect in pathological voice detection
Comments: Machine Learning for Health (ML4H) Workshop at NeurIPS 2018 arXiv:1811.07216
Subjects: Machine Learning (cs.LG); Sound (cs.SD); Audio and Speech Processing (eess.AS); Machine Learning (stat.ML)
[153]  arXiv:1811.10561 (cross-list from cs.CL) [pdf, other]
Title: CLEAR: A Dataset for Compositional Language and Elementary Acoustic Reasoning
Comments: NeurIPS 2018 Visually Grounded Interaction and Language (ViGIL) Workshop
Subjects: Computation and Language (cs.CL); Machine Learning (cs.LG); Sound (cs.SD); Audio and Speech Processing (eess.AS); Machine Learning (stat.ML)
[154]  arXiv:1811.10708 (cross-list from cs.SD) [pdf, other]
Title: Combining High-Level Features of Raw Audio Waves and Mel-Spectrograms for Audio Tagging
Comments: Detection and Classification of Acoustic Scenes and Events 2018 (DCASE 2018), 19-20 November 2018, Surrey, UK
Subjects: Sound (cs.SD); Machine Learning (cs.LG); Audio and Speech Processing (eess.AS)
[155]  arXiv:1811.10736 (cross-list from cs.LG) [pdf, other]
Title: DONUT: CTC-based Query-by-Example Keyword Spotting
Comments: Accepted to NeurIPS 2018 Workshop on Interpretability and Robustness for Audio, Speech, and Language
Subjects: Machine Learning (cs.LG); Sound (cs.SD); Audio and Speech Processing (eess.AS); Machine Learning (stat.ML)
[156]  arXiv:1811.10813 (cross-list from cs.CV) [pdf, other]
Title: Noise-tolerant Audio-visual Online Person Verification using an Attention-based Neural Network Fusion
Subjects: Computer Vision and Pattern Recognition (cs.CV); Audio and Speech Processing (eess.AS)
[157]  arXiv:1811.10988 (cross-list from cs.IR) [pdf, other]
Title: Facilitating the Manual Annotation of Sounds When Using Large Taxonomies
Comments: 5 pages, 5 figures, IEEE FRUCT International Workshop on Semantic Audio and the Internet of Things
Journal-ref: Proceedings of the 23rd Conference of Open Innovations Association FRUCT, Bologna, Italy. 2018. ISSN 2305-7254, ISBN 978-952-68653-6-2, FRUCT Oy, e-ISSN 2343-0737 (license CC BY-ND)
Subjects: Information Retrieval (cs.IR); Human-Computer Interaction (cs.HC); Machine Learning (cs.LG); Sound (cs.SD); Audio and Speech Processing (eess.AS)
[158]  arXiv:1811.11307 (cross-list from cs.SD) [pdf, other]
Title: Improved Speech Enhancement with the Wave-U-Net
Comments: 5 pages (including 1 for References), 1 figure, 2 tables
Subjects: Sound (cs.SD); Machine Learning (cs.LG); Neural and Evolutionary Computing (cs.NE); Audio and Speech Processing (eess.AS); Signal Processing (eess.SP)
[159]  arXiv:1811.11663 (cross-list from cs.SD) [pdf, other]
Title: Multiple source direction of arrival estimation using subspace pseudointensity vectors
Comments: In Proceedings of the LOCATA Challenge Workshop - a satellite event of IWAENC 2018 (arXiv:1811.08482 )
Subjects: Sound (cs.SD); Audio and Speech Processing (eess.AS)
[160]  arXiv:1811.12208 (cross-list from cs.SD) [pdf, other]
Title: UFANS: U-shaped Fully-Parallel Acoustic Neural Structure For Statistical Parametric Speech Synthesis With 20X Faster
Subjects: Sound (cs.SD); Audio and Speech Processing (eess.AS)
[161]  arXiv:1811.12214 (cross-list from cs.SD) [pdf, other]
Title: Play as You Like: Timbre-enhanced Multi-modal Music Style Transfer
Subjects: Sound (cs.SD); Audio and Speech Processing (eess.AS)
[162]  arXiv:1811.12254 (cross-list from cs.LG) [pdf, other]
Title: The Effect of Heterogeneous Data for Alzheimer's Disease Detection from Speech
Comments: Machine Learning for Health (ML4H) Workshop at NeurIPS 2018 arXiv:1811.07216
Subjects: Machine Learning (cs.LG); Computation and Language (cs.CL); Sound (cs.SD); Audio and Speech Processing (eess.AS); Machine Learning (stat.ML)
[163]  arXiv:1811.12408 (cross-list from cs.SD) [pdf, other]
Title: From Context to Concept: Exploring Semantic Relationships in Music with Word2Vec
Comments: Accepted for publication in Neural Computing and Applications, Springer. In Press
Journal-ref: Neural Computing and Applications, Springer. 2019
Subjects: Sound (cs.SD); Information Retrieval (cs.IR); Machine Learning (cs.LG); Audio and Speech Processing (eess.AS); Machine Learning (stat.ML)
[164]  arXiv:1811.12802 (cross-list from cs.IR) [pdf, other]
Title: Naive Dictionary On Musical Corpora: From Knowledge Representation To Pattern Recognition
Comments: 25 pages
Subjects: Information Retrieval (cs.IR); Machine Learning (cs.LG); Sound (cs.SD); Audio and Speech Processing (eess.AS); Machine Learning (stat.ML)
[165]  arXiv:1811.12739 (cross-list from cs.LG) [pdf, other]
Title: Neural separation of observed and unobserved distributions
Comments: ICML'19
Subjects: Machine Learning (cs.LG); Computation and Language (cs.CL); Audio and Speech Processing (eess.AS); Machine Learning (stat.ML)
[166]  arXiv:1811.03700 (cross-list from cs.LG) [pdf, ps, other]
Title: A Comparison of Lattice-free Discriminative Training Criteria for Purely Sequence-Trained Neural Network Acoustic Models
Authors: Chao Weng, Dong Yu
Comments: under review ICASSP2019
Subjects: Machine Learning (cs.LG); Artificial Intelligence (cs.AI); Computation and Language (cs.CL); Audio and Speech Processing (eess.AS); Machine Learning (stat.ML)
[ total of 166 entries: 1-50 | 51-100 | 101-150 | 151-166 ]
[ showing 50 entries per page: fewer | more | all ]

Disable MathJax (What is MathJax?)

Links to: arXiv, form interface, find, eess, 2303, contact, help  (Access key information)