TASK	DATASET	MODEL	METRIC NAME	METRIC VALUE	GLOBAL RANK
Vocal Bursts Type Prediction	HUME-VB	w2v2-r-er	Average Recall	0.4902	# 1
Cultural Vocal Bursts Intensity Prediction	HUME-VB	w2v2-r-er	Concordance correlation coefficient (CCC)	0.5199	# 2
Vocal Bursts Valence Prediction	HUME-VB	w2v2-r-er	Concordance correlation coefficient (CCC)	0.6290	# 1
Vocal Bursts Intensity Prediction	HUME-VB	w2v2-r-vad	Concordance correlation coefficient (CCC)	0.6554	# 2

Badge	Markdown
	`[![PWC](https://img.shields.io/endpoint.svg?url=https://paperswithcode.com/badge/evaluating-variants-of-wav2vec-2-0-on/type-on-hume-vb)](https://paperswithcode.com/sota/type-on-hume-vb?p=evaluating-variants-of-wav2vec-2-0-on)`
	`[![PWC](https://img.shields.io/endpoint.svg?url=https://paperswithcode.com/badge/evaluating-variants-of-wav2vec-2-0-on/two-on-hume-vb)](https://paperswithcode.com/sota/two-on-hume-vb?p=evaluating-variants-of-wav2vec-2-0-on)`
	`[![PWC](https://img.shields.io/endpoint.svg?url=https://paperswithcode.com/badge/evaluating-variants-of-wav2vec-2-0-on/culture-on-hume-vb)](https://paperswithcode.com/sota/culture-on-hume-vb?p=evaluating-variants-of-wav2vec-2-0-on)`
	`[![PWC](https://img.shields.io/endpoint.svg?url=https://paperswithcode.com/badge/evaluating-variants-of-wav2vec-2-0-on/high-on-hume-vb)](https://paperswithcode.com/sota/high-on-hume-vb?p=evaluating-variants-of-wav2vec-2-0-on)`

Evaluating Variants of wav2vec 2.0 on Affective Vocal Burst Tasks

ICASSP 2023 · Bagus Tris Atmaja, Akira Sasou ·

The search for emotional biomarkers within the human voice is a challenging research area. Previous studies focused on predicting affective state from speech; this study explores various tasks on affective vocal bursts. Borrowing the success of self-supervised learning in automatic speech recognition, we extracted acoustic embedding using variants of wav2vec 2.0 for four affective vocal bursts tasks: High, Two, Culture, and Type. Using a similar architecture for all tasks, the evaluation of acoustic embeddings reveals the potential use of wav2vec 2.0 variants over the conventional acoustic features in affective vocal bursts tasks. We evaluated both conventional acoustic features and these acoustic embeddings on the different number of twenty seeds evaluation and reported the maximum and average scores with their standard deviations in the validation set. Three high scores from these validations for all tasks assist the generation of predictions for the test set. We compared the test scores with previous studies and obtained remarkable improvements.

PDF

Code

Add Remove Mark official

bagustris/A-VB2022_CCC

Tasks

Add Remove

Automatic Speech Recognition

Cultural Vocal Bursts Intensity Prediction

Self-Supervised Learning

speech-recognition

Speech Recognition

Vocal Bursts Intensity Prediction

Vocal Bursts Type Prediction

Vocal Bursts Valence Prediction

Datasets

HUME-VB

Results from the Paper

Add Remove

Ranked #1 on Vocal Bursts Type Prediction on HUME-VB

Get a GitHub badge

Task	Dataset	Model	Metric Name	Metric Value	Global Rank	Benchmark
Vocal Bursts Type Prediction	HUME-VB	w2v2-r-er	Average Recall	0.4902	# 1	Compare
Cultural Vocal Bursts Intensity Prediction	HUME-VB	w2v2-r-er	Concordance correlation coefficient (CCC)	0.5199	# 2	Compare
Vocal Bursts Valence Prediction	HUME-VB	w2v2-r-er	Concordance correlation coefficient (CCC)	0.6290	# 1	Compare
Vocal Bursts Intensity Prediction	HUME-VB	w2v2-r-vad	Concordance correlation coefficient (CCC)	0.6554	# 2	Compare

Methods

Add Remove

Test

Edit Social Preview

Evaluating Variants of wav2vec 2.0 on Affective Vocal Burst Tasks

Code Edit Add Remove Mark official

Tasks Edit Add Remove

Datasets Edit

Results from the Paper Edit Add Remove

Methods Edit Add Remove

Code

Add Remove Mark official

Tasks

Add Remove

Datasets

Results from the Paper

Add Remove

Methods

Add Remove