anonymous-neurips-2026/memebench
收藏资源简介:
MemeBench是一个双语诊断性基准,用于开放式模因解释,评估大型视觉语言模型的文化语义理解能力。它包含1,253个专家标注的模因,覆盖7个文化领域(中文和英文)。每个模因都标注了结构化的VIKR模式,涵盖四个诊断层:视觉线索(描述所见内容)、身份链接(识别描绘的对象)、知识单元(了解文化背景)和推理机制(解释模因的幽默之处)。数据集还包括详细的元数据、统计信息、评估协议和创建过程。
MemeBench is a bilingual diagnostic benchmark for open-ended meme interpretation, evaluating large vision-language models cultural-semantic understanding. It contains 1,253 expert-annotated memes spanning 7 cultural domains in Chinese and English. Each meme is annotated with a structured VIKR schema covering four diagnostic layers: Visual clues (describing what is seen), Identity links (identifying depicted entities), Knowledge units (understanding cultural background), and Reasoning mechanisms (explaining why the meme is funny). The dataset also includes detailed metadata, statistics, evaluation protocols, and creation processes.




