Sound

Authors and titles for cs.SD in Jun 2022, skipping first 45

[ total of 221 entries: 1-10 | ... | 16-25 | 26-35 | 36-45 | 46-55 | 56-65 | 66-75 | 76-85 | ... | 216-221 ]
[ showing 10 entries per page: fewer | more | all ]

[46] arXiv:2206.08007 [pdf, ps, other]: Title: DCASE 2022: Comparative Analysis Of CNNs For Acoustic Scene Classification Under Low-Complexity Considerations

Authors: Josep Zaragoza-Paredes, Javier Naranjo-Alcazar, Valery Naranjo, Pedro Zuccarello

Subjects: Sound (cs.SD); Machine Learning (cs.LG); Audio and Speech Processing (eess.AS)
[47] arXiv:2206.08039 [pdf, ps, other]: Title: Acoustic Modeling for End-to-End Empathetic Dialogue Speech Synthesis Using Linguistic and Prosodic Contexts of Dialogue History

Authors: Yuto Nishimura, Yuki Saito, Shinnosuke Takamichi, Kentaro Tachibana, Hiroshi Saruwatari

Comments: 5 pages, 3 figures, Accepted for INTERSPEECH2022

Subjects: Sound (cs.SD); Computation and Language (cs.CL); Machine Learning (cs.LG); Audio and Speech Processing (eess.AS)
[48] arXiv:2206.08170 [pdf, other]: Title: Adversarial Privacy Protection on Speech Enhancement

Authors: Mingyu Dong, Diqun Yan, Rangding Wang

Comments: 5 pages, 6 figures

Subjects: Sound (cs.SD); Cryptography and Security (cs.CR); Machine Learning (cs.LG); Audio and Speech Processing (eess.AS)
[49] arXiv:2206.08189 [pdf, other]: Title: Censer: Curriculum Semi-supervised Learning for Speech Recognition Based on Self-supervised Pre-training

Authors: Bowen Zhang, Songjun Cao, Xiaoming Zhang, Yike Zhang, Long Ma, Takahiro Shinozaki

Subjects: Sound (cs.SD); Audio and Speech Processing (eess.AS)
[50] arXiv:2206.08233 [pdf, other]: Title: Event-related data conditioning for acoustic event classification

Authors: Yuanbo Hou, Dick Botteldooren

Comments: Accepted by INTERSPEECH 2022

Subjects: Sound (cs.SD); Audio and Speech Processing (eess.AS)
[51] arXiv:2206.08297 [pdf, other]: Title: A Language Model With Million Sample Context For Raw Audio Using Transformer Architectures

Authors: Prateek Verma

Comments: 12 pages, 1 figure. Technical Report at Stanford University

Subjects: Sound (cs.SD); Machine Learning (cs.LG); Audio and Speech Processing (eess.AS)
[52] arXiv:2206.08312 [pdf, other]: Title: SoundSpaces 2.0: A Simulation Platform for Visual-Acoustic Learning

Authors: Changan Chen, Carl Schissler, Sanchit Garg, Philip Kobernik, Alexander Clegg, Paul Calamia, Dhruv Batra, Philip W Robinson, Kristen Grauman

Comments: Camera-ready version. Website: this https URL Project page: this https URL

Subjects: Sound (cs.SD); Computer Vision and Pattern Recognition (cs.CV); Multimedia (cs.MM); Audio and Speech Processing (eess.AS)
[53] arXiv:2206.08317 [pdf, other]: Title: Paraformer: Fast and Accurate Parallel Transformer for Non-autoregressive End-to-End Speech Recognition

Authors: Zhifu Gao, Shiliang Zhang, Ian McLoughlin, Zhijie Yan

Comments: 5 pages, 3 figures, accepted by INTERSPEECH 2022

Subjects: Sound (cs.SD); Computation and Language (cs.CL); Audio and Speech Processing (eess.AS)
[54] arXiv:2206.09131 [pdf, other]: Title: Tackling Spoofing-Aware Speaker Verification with Multi-Model Fusion

Authors: Haibin Wu, Jiawen Kang, Lingwei Meng, Yang Zhang, Xixin Wu, Zhiyong Wu, Hung-yi Lee, Helen Meng

Comments: Accepted by Odyssey 2022

Subjects: Sound (cs.SD); Machine Learning (cs.LG); Audio and Speech Processing (eess.AS)
[55] arXiv:2206.09142 [pdf, other]: Title: Redundancy Reduction Twins Network: A Training framework for Multi-output Emotion Regression

Authors: Xin Jing, Meishu Song, Andreas Triantafyllopoulos, Zijiang Yang, Björn W. Schuller

Comments: 5 pages, accepted by ICML Exvo workshop

Subjects: Sound (cs.SD); Audio and Speech Processing (eess.AS)

[ total of 221 entries: 1-10 | ... | 16-25 | 26-35 | 36-45 | 46-55 | 56-65 | 66-75 | 76-85 | ... | 216-221 ]
[ showing 10 entries per page: fewer | more | all ]

Disable MathJax (What is MathJax?)

Links to: arXiv, form interface, find, cs, 2404, contact, help (Access key information)

> cs > cs.SD

Sound

Authors and titles for cs.SD in Jun 2022, skipping first 45