Burst2Vec: An Adversarial Multi-Task Approach for Predicting Emotion, Age, and Origin from Vocal Bursts

Anuchitanukul, Atijit; Specia, Lucia

Full-text links:

Download:

Current browse context:

cs.SD

< prev | next >

new | recent | 2206

Computer Science > Sound

Title: Burst2Vec: An Adversarial Multi-Task Approach for Predicting Emotion, Age, and Origin from Vocal Bursts

Authors: Atijit Anuchitanukul, Lucia Specia

(Submitted on 24 Jun 2022 (v1), last revised 18 Oct 2022 (this version, v2))

Abstract: We present Burst2Vec, our multi-task learning approach to predict emotion, age, and origin (i.e., native country/language) from vocal bursts. Burst2Vec utilises pre-trained speech representations to capture acoustic information from raw waveforms and incorporates the concept of model debiasing via adversarial training. Our models achieve a relative 30 % performance gain over baselines using pre-extracted features and score the highest amongst all participants in the ICML ExVo 2022 Multi-Task Challenge.

Subjects:	Sound (cs.SD); Computation and Language (cs.CL); Audio and Speech Processing (eess.AS)
Cite as:	arXiv:2206.12469 [cs.SD]
	(or arXiv:2206.12469v2 [cs.SD] for this version)

Submission history

From: Atijit Anuchitanukul [view email]
[v1] Fri, 24 Jun 2022 18:57:41 GMT (1380kb,D)
[v2] Tue, 18 Oct 2022 05:48:23 GMT (1380kb,D)

Which authors of this paper are endorsers? | Disable MathJax (What is MathJax?)

Link back to: arXiv, form interface, contact.

> cs > arXiv:2206.12469

Download:

Current browse context:

Change to browse by:

References & Citations

DBLP - CS Bibliography

Bookmark

Computer Science > Sound

Title: Burst2Vec: An Adversarial Multi-Task Approach for Predicting Emotion, Age, and Origin from Vocal Bursts

Submission history