MER 2024: Semi-Supervised Learning, Noise Robustness, and Open-Vocabulary Multimodal Emotion Recognition

Lian, Zheng; Sun, Haiyang; Sun, Licai; Wen, Zhuofan; Zhang, Siyuan; Chen, Shun; Gu, Hao; Zhao, Jinming; Ma, Ziyang; Chen, Xie; Yi, Jiangyan; Liu, Rui; Xu, Kele; Liu, Bin; Cambria, Erik; Zhao, Guoying; Schuller, Björn W.; Tao, Jianhua

Computer Science > Machine Learning

arXiv:2404.17113v3 (cs)

[Submitted on 26 Apr 2024 (v1), revised 23 May 2024 (this version, v3), latest version 18 Jul 2024 (v4)]

Title:MER 2024: Semi-Supervised Learning, Noise Robustness, and Open-Vocabulary Multimodal Emotion Recognition

Authors:Zheng Lian, Haiyang Sun, Licai Sun, Zhuofan Wen, Siyuan Zhang, Shun Chen, Hao Gu, Jinming Zhao, Ziyang Ma, Xie Chen, Jiangyan Yi, Rui Liu, Kele Xu, Bin Liu, Erik Cambria, Guoying Zhao, Björn W. Schuller, Jianhua Tao

View PDF HTML (experimental)

Abstract:Multimodal emotion recognition is an important research topic in artificial intelligence. Over the past few decades, researchers have made remarkable progress by increasing dataset size and building more effective architectures. However, due to various reasons (such as complex environments and inaccurate annotations), current systems are hard to meet the demands of practical applications. Therefore, we organize a series of challenges around emotion recognition to further promote the development of this area. Last year, we launched MER2023, focusing on three topics: multi-label learning, noise robustness, and semi-supervised learning. This year, we continue to organize MER2024. In addition to expanding the dataset size, we introduce a new track around open-vocabulary emotion recognition. The main consideration for this track is that existing datasets often fix the label space and use majority voting to enhance annotator consistency, but this process may limit the model's ability to describe subtle emotions. In this track, we encourage participants to generate any number of labels in any category, aiming to describe the emotional state as accurately as possible. Our baseline is based on MERTools and the code is available at: this https URL.

Subjects:	Machine Learning (cs.LG); Human-Computer Interaction (cs.HC)
Cite as:	arXiv:2404.17113 [cs.LG]
	(or arXiv:2404.17113v3 [cs.LG] for this version)
	https://doi.org/10.48550/arXiv.2404.17113

Submission history

From: Zheng Lian [view email]
[v1] Fri, 26 Apr 2024 02:05:20 UTC (243 KB)
[v2] Mon, 29 Apr 2024 02:11:25 UTC (243 KB)
[v3] Thu, 23 May 2024 12:43:15 UTC (244 KB)
[v4] Thu, 18 Jul 2024 11:23:25 UTC (290 KB)

Computer Science > Machine Learning

Title:MER 2024: Semi-Supervised Learning, Noise Robustness, and Open-Vocabulary Multimodal Emotion Recognition

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Machine Learning

Title:MER 2024: Semi-Supervised Learning, Noise Robustness, and Open-Vocabulary Multimodal Emotion Recognition

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators