EVADE
收藏资源简介:
EVADE是一个专家精心策划的中文多模态基准数据集,旨在评估基础模型在电子商务中检测规避内容的性能。该数据集包含2,833个注释文本样本和13,961张图像,涵盖六个具有挑战性的产品类别,包括身体塑形、增高和健康补充品。数据集由电子商务平台收集,每个样本都由领域专家进行迭代注释,以确保其准确性和可靠性。EVADE包括两个互补的任务:Single-Violation和All-in-One,分别评估模型在短提示下的细粒度推理能力和在合并重叠政策规则为统一指令下的长上下文推理能力。数据集旨在解决电子商务中规避内容检测的问题,帮助开发更安全、更透明的内容审查系统。
EVADE is a Chinese multimodal benchmark dataset meticulously curated by domain experts, designed to evaluate the performance of foundation models in detecting evasive content within e-commerce scenarios. This dataset comprises 2,833 annotated text samples and 13,961 images, covering six challenging product categories including body shaping, height-increasing, and health supplement products. Collected from e-commerce platforms, each sample in the dataset undergoes iterative annotation by domain experts to ensure its accuracy and reliability. EVADE includes two complementary tasks: Single-Violation and All-in-One, which respectively evaluate the model's fine-grained reasoning capability under short prompts and long-context reasoning capability when merging overlapping policy rules into a unified instruction. This dataset aims to address the problem of evasive content detection in e-commerce, and assist in developing safer and more transparent content moderation systems.
EVADE-Bench 数据集概述
基本信息
- 许可证: CC-BY-NC-4.0
- 任务类别:
- 文本分类
- 问答
- 零样本分类
- 语言: 中文
- 标签:
- 规避内容检测
- 基准测试
- 大语言模型 (LLMs)
- 视觉语言模型 (VLMs)
- 规模: 10K < n < 100K
数据集内容
字段说明
- id: 样本唯一标识符
- content_type: 内容类型
- single_risk_question: 单一风险任务提示
- single_risk_options: 单一风险任务选项
- all_in_one_detail_question: 包含示例的allinone任务提示
- all_in_one_simple_question: 无示例的allinone任务提示
- all_in_one_options: allinone任务选项
- content_image: 图像分割中的图像信息
- content_text: 文本分割中的文本信息
- extra: 附加信息
数据规模
- 图像: 13,961张
- 文本: 2,833条
使用条款
-
学术限定原则
- 仅限非营利性学术研究使用
- 禁止用于侵犯隐私或违反法律法规的活动
-
内容中立声明
- 数据呈现形式不代表提供方立场
- 使用者需独立判断并承担相关风险
-
最终解释权
- 归数据提供方所有
相关资源




