2 dataset results for Multimodal Emotion Recognition AND Audio AND English

IEMOCAP (The Interactive Emotional Dyadic Motion Capture (IEMOCAP) Database)

Multimodal Emotion Recognition IEMOCAP The IEMOCAP dataset consists of 151 videos of recorded dialogues, with 2 speakers per session for a total of 302 videos across the dataset. Each segment is annotated for the presence of 9 emotions (angry, excited, fear, sad, surprised, frustrated, happy, disappointed and neutral) as well as valence, arousal and dominance. The dataset is recorded across 5 sessions with 5 pairs of speakers.

640 PAPERS • 3 BENCHMARKS

CMU-MOSEI

CMU Multimodal Opinion Sentiment and Emotion Intensity (CMU-MOSEI) is the largest dataset of sentence level sentiment analysis and emotion recognition in online videos. CMU-MOSEI contains more than 65 hours of annotated video from more than 1000 speakers and 250 topics.

154 PAPERS • 2 BENCHMARKS

Datasets

2 dataset results for Multimodal Emotion Recognition AND Audio AND English