Meld
Progress Over Time
Interactive timeline showing model performance evolution on Meld
Meld Leaderboard
| Context | Cost | License | ||||
|---|---|---|---|---|---|---|
| 1 | Alibaba Cloud / Qwen Team | 7B | — | — |
What is Meld?
MELD (Multimodal EmotionLines Dataset) is a multimodal multi-party dataset for emotion recognition in conversations. Contains approximately 13,000 utterances from 1,433 dialogues extracted from the TV series Friends. Each utterance is annotated with emotion (Anger, Disgust, Sadness, Joy, Neutral, Surprise, Fear) and sentiment labels across audio, visual, and textual modalities.
Meld is a multimodal benchmark evaluating models on multimodal, psychology, and creativity tasks. LLM Stats tracks 1 models on this benchmark, scored on a 0–1 scale. The current average is 0.6, with the leader at 0.6.
Compare leaders on the best AI for multimodal, best AI for psychology and best AI for creativity leaderboards.
Current leaders
Qwen2.5-Omni-7B from Alibaba Cloud / Qwen Team currently leads the Meld leaderboard with a score of 0.570 across 1 evaluated AI models.
FAQ
Common questions about the Meld benchmark and leaderboard.