We gratefully acknowledge support from
the Simons Foundation and member institutions.

Audio and Speech Processing

Authors and titles for eess.AS in Nov 2018, skipping first 50

[ total of 166 entries: 1-25 | 26-50 | 51-75 | 76-100 | 101-125 | 126-150 | 151-166 ]
[ showing 25 entries per page: fewer | more | all ]
[51]  arXiv:1811.09725 [pdf, other]
Title: Interpretable Convolutional Filters with SincNet
Comments: In Proceedings of NIPS@IRASL 2018. arXiv admin note: substantial text overlap with arXiv:1808.00158
Subjects: Audio and Speech Processing (eess.AS); Computation and Language (cs.CL); Machine Learning (cs.LG); Neural and Evolutionary Computing (cs.NE)
[52]  arXiv:1811.09919 [pdf, other]
Title: A Method for Analysis of Patient Speech in Dialogue for Dementia Detection
Comments: 8 pages, Resources and ProcessIng of linguistic, paralinguistic and extra-linguistic Data from people with various forms of cognitive impairment, LREC 2018
Subjects: Audio and Speech Processing (eess.AS); Machine Learning (cs.LG); Sound (cs.SD)
[53]  arXiv:1811.10812 [pdf, other]
Title: Large-scale Speaker Retrieval on Random Speaker Variability Subspace
Comments: Interspeech 2019
Subjects: Audio and Speech Processing (eess.AS); Information Retrieval (cs.IR)
[54]  arXiv:1811.11078 [pdf, other]
Title: Refined WaveNet Vocoder for Variational Autoencoder Based Voice Conversion
Comments: 5 pages, 7 figures, 1 table. Accepted to EUSIPCO 2019
Subjects: Audio and Speech Processing (eess.AS); Computation and Language (cs.CL); Sound (cs.SD)
[55]  arXiv:1811.11517 [pdf, other]
Title: Acoustics-guided evaluation (AGE): a new measure for estimating performance of speech enhancement algorithms for robust ASR
Comments: Submitted to ICASSP 2019
Subjects: Audio and Speech Processing (eess.AS); Sound (cs.SD)
[56]  arXiv:1811.11785 [pdf, ps, other]
Title: SVD-PHAT: A Fast Sound Source Localization Method
Journal-ref: Proceedings of the 2019 IEEE International Conference on Acoustics, Speech, and Signal Processing
Subjects: Audio and Speech Processing (eess.AS); Sound (cs.SD); Signal Processing (eess.SP)
[57]  arXiv:1811.11787 [pdf, ps, other]
Title: A Study of the Complexity and Accuracy of Direction of Arrival Estimation Methods Based on GCC-PHAT for a Pair of Close Microphones
Subjects: Audio and Speech Processing (eess.AS); Sound (cs.SD); Signal Processing (eess.SP)
[58]  arXiv:1811.11913 [pdf, other]
Title: LP-WaveNet: Linear Prediction-based WaveNet Speech Synthesis
Comments: Submitted to EUSIPCO 2020
Subjects: Audio and Speech Processing (eess.AS); Sound (cs.SD)
[59]  arXiv:1811.12290 [pdf, other]
Title: Tuplemax Loss for Language Identification
Comments: Submitted to ICASSP 2019
Subjects: Audio and Speech Processing (eess.AS); Machine Learning (cs.LG); Sound (cs.SD); Machine Learning (stat.ML)
[60]  arXiv:1811.02489 (cross-list from eess.SP) [pdf, other]
Title: Unifying Probabilistic Models for Time-Frequency Analysis
Comments: Accepted to International Conference on Acoustics, Speech and Signal Processing (ICASSP) 2019
Subjects: Signal Processing (eess.SP); Machine Learning (cs.LG); Sound (cs.SD); Audio and Speech Processing (eess.AS); Machine Learning (stat.ML)
[61]  arXiv:1811.08783 (cross-list from eess.SP) [pdf, other]
Title: Designing nearly tight window for improving time-frequency masking
Subjects: Signal Processing (eess.SP); Sound (cs.SD); Audio and Speech Processing (eess.AS)
[62]  arXiv:1811.00002 (cross-list from cs.SD) [pdf, other]
Title: WaveGlow: A Flow-based Generative Network for Speech Synthesis
Comments: 5 pages, 1 figure, 1 table, 13 equations
Subjects: Sound (cs.SD); Artificial Intelligence (cs.AI); Machine Learning (cs.LG); Audio and Speech Processing (eess.AS); Machine Learning (stat.ML)
[63]  arXiv:1811.00003 (cross-list from cs.SD) [src]
Title: Deep Net Features for Complex Emotion Recognition
Comments: Conflict of interest
Subjects: Sound (cs.SD); Machine Learning (cs.LG); Audio and Speech Processing (eess.AS); Machine Learning (stat.ML)
[64]  arXiv:1811.00078 (cross-list from cs.SD) [pdf, other]
Title: On Single-Channel Speech Enhancement and On Non-Linear Modulation-Domain Kalman Filtering
Comments: 13 pages
Subjects: Sound (cs.SD); Audio and Speech Processing (eess.AS)
[65]  arXiv:1811.00162 (cross-list from cs.AI) [pdf, other]
Title: Modeling Melodic Feature Dependency with Modularized Variational Auto-Encoder
Comments: The first three authors contributed equally
Subjects: Artificial Intelligence (cs.AI); Machine Learning (cs.LG); Multimedia (cs.MM); Sound (cs.SD); Audio and Speech Processing (eess.AS)
[66]  arXiv:1811.00183 (cross-list from stat.ML) [pdf, other]
Title: Designing an Effective Metric Learning Pipeline for Speaker Diarization
Subjects: Machine Learning (stat.ML); Machine Learning (cs.LG); Sound (cs.SD); Audio and Speech Processing (eess.AS)
[67]  arXiv:1811.00223 (cross-list from cs.SD) [pdf, other]
Title: Neural Music Synthesis for Flexible Timbre Control
Subjects: Sound (cs.SD); Audio and Speech Processing (eess.AS); Machine Learning (stat.ML)
[68]  arXiv:1811.00301 (cross-list from cs.SD) [pdf]
Title: Weakly supervised CRNN system for sound event detection with large-scale unlabeled in-domain data
Comments: Submitted to ICASSP 2019
Subjects: Sound (cs.SD); Audio and Speech Processing (eess.AS)
[69]  arXiv:1811.00348 (cross-list from cs.SD) [pdf, ps, other]
Title: Sequence-to-sequence Models for Small-Footprint Keyword Spotting
Comments: Submitted to ICASSP 2019
Subjects: Sound (cs.SD); Audio and Speech Processing (eess.AS)
[70]  arXiv:1811.00350 (cross-list from cs.SD) [pdf, ps, other]
Title: End-to-end Models with auditory attention in Multi-channel Keyword Spotting
Comments: Submitted to ICASSP 2019
Subjects: Sound (cs.SD); Audio and Speech Processing (eess.AS)
[71]  arXiv:1811.00403 (cross-list from cs.CL) [pdf, other]
Title: Truly unsupervised acoustic word embeddings using weak top-down constraints in encoder-decoder models
Authors: Herman Kamper
Comments: 5 pages, 3 figures, 2 tables; accepted to ICASSP 2019
Subjects: Computation and Language (cs.CL); Machine Learning (cs.LG); Sound (cs.SD); Audio and Speech Processing (eess.AS)
[72]  arXiv:1811.00454 (cross-list from cs.SD) [pdf, ps, other]
Title: Referenceless Performance Evaluation of Audio Source Separation using Deep Neural Networks
Journal-ref: This paper will be presented at EUSIPCO 2019
Subjects: Sound (cs.SD); Machine Learning (cs.LG); Multimedia (cs.MM); Audio and Speech Processing (eess.AS)
[73]  arXiv:1811.00707 (cross-list from cs.CL) [pdf, other]
Title: Training Neural Speech Recognition Systems with Synthetic Speech Augmentation
Comments: Pre-print. Work in progress, 5 pages, 1 figure
Subjects: Computation and Language (cs.CL); Machine Learning (cs.LG); Sound (cs.SD); Audio and Speech Processing (eess.AS)
[74]  arXiv:1811.00936 (cross-list from cs.SD) [pdf, other]
Title: Acoustic Features Fusion using Attentive Multi-channel Deep Architecture
Comments: Accepted in CHiME'18 (Interspeech Workshop)
Subjects: Sound (cs.SD); Audio and Speech Processing (eess.AS)
[75]  arXiv:1811.01092 (cross-list from cs.LG) [pdf, ps, other]
Title: Unifying Isolated and Overlapping Audio Event Detection with Multi-Label Multi-Task Convolutional Recurrent Neural Networks
Comments: Accepted for the 44th International Conference on Acoustics, Speech, and Signal Processing (ICASSP 2019)
Subjects: Machine Learning (cs.LG); Sound (cs.SD); Audio and Speech Processing (eess.AS); Machine Learning (stat.ML)
[ total of 166 entries: 1-25 | 26-50 | 51-75 | 76-100 | 101-125 | 126-150 | 151-166 ]
[ showing 25 entries per page: fewer | more | all ]

Disable MathJax (What is MathJax?)

Links to: arXiv, form interface, find, eess, 2301, contact, help  (Access key information)