MMAFFBen
收藏资源简介:
MMAFFBen是一个多语言和多模态情感分析基准数据集,旨在评估大型语言模型(LLMs)和视觉语言模型(VLMs)的情感分析能力。该数据集涵盖了35种语言的文本、图像和视频模态,并支持对四种情感分析任务的评估:情感强度、情感分类、情感极性和情感强度。MMAFFBen通过整合现有开源数据集,经过精心筛选和重构,为情感分析任务提供了全面的评估框架。该数据集的创建旨在解决当前情感分析评估基准的局限性,如覆盖范围有限、模态单一或语言单一等问题,以促进情感分析研究和应用的发展。
MMAFFBen is a multilingual and multimodal sentiment analysis benchmark dataset developed to evaluate the sentiment analysis capabilities of Large Language Models (LLMs) and Vision-Language Models (VLMs). This dataset covers text, image, and video modalities across 35 languages, and supports evaluation of four sentiment analysis tasks: emotion intensity, sentiment classification, sentiment polarity, and emotion intensity. MMAFFBen provides a comprehensive evaluation framework for sentiment analysis tasks by integrating existing open-source datasets and undergoing rigorous screening and restructuring. The creation of this dataset aims to address the limitations of current sentiment analysis evaluation benchmarks, such as limited coverage, single-modality or single-language constraints, so as to facilitate the advancement of sentiment analysis research and applications.
MMAFFBen 数据集概述
数据集简介
- 名称:MMAFFBen (Multilingual and Multimodal Affective Analysis Benchmark)
- 类型:多语言多模态情感分析基准数据集
- 目的:用于评估大型语言模型(LLMs)和视觉语言模型(VLMs)的性能
- 特点:开源、多语言、多模态
数据集组成
相关模型
- 3B参数模型:MMAFFLM-3b
- 7B参数模型:MMAFFLM-7b
使用说明
模型微调
- 下载MMAFFIn训练数据集至data文件夹
- 运行命令: bash bash run_sft_stream.sh
模型评估
-
下载MMAFFBen数据至data文件夹
-
运行命令: bash bash run_inference.sh
-
预测结果将保存在predicts文件夹
-
使用evaluation.ipynb计算各子数据集得分
技术基础
- 基于LLaMA-Factory框架
- 当前版本支持Qwen-VL系列模型

- 1MMAFFBen: A Multilingual and Multimodal Affective Analysis Benchmark for Evaluating LLMs and VLMsThe University of Manchester United Kingdom · 2025年



