ViSAGe
收藏资源简介:
ViSAGe数据集用于评估文本到图像生成模型中的已知民族基于的刻板印象,涵盖135个民族。该数据集通过区分更有可能具有视觉描述的刻板印象和那些不太具体的刻板印象来丰富现有的文本刻板印象资源。
The ViSAGe dataset is designed to evaluate known ethnicity-based stereotypes in text-to-image generation models, encompassing 135 ethnic groups. This dataset enriches existing textual stereotype resources by distinguishing between stereotypes that are more likely to have visual descriptions and those that are less specific.
数据集概述
名称:ViSAGe (Visual Stereotypes Around the Globe)
目的:评估Text-to-Image (T2I)模型中的国家基于的视觉刻板印象,涵盖135个国籍。
特点:
- 通过区分更可能具有视觉描绘的刻板印象(如
sombrero)与较少视觉具体的刻板印象(如attractive),丰富了现有的文本刻板印象资源。 - 展示了刻板印象属性在生成图像中的显著性和冒犯性,特别是在非洲、南美洲和东南亚的身份群体中。
- 评估了身份群体视觉描绘的刻板印象吸引力,揭示了所有身份群体的默认表示向刻板印象描绘的倾向,尤其是全球南方的身份群体。
数据集内容
- 数据卡片:包含数据集的详细信息,如预期用途、字段名称和含义、标注者招募和支付。
- 文件:
visual_attributes:包含基于Likert量表的属性视觉性质的标注。Image_Annotations:包含图像中属性的存在与否及其坐标的标注。
引用信息
@inproceedings{jha-2024-beyond, title={ViSAGe: A Global-Scale Analysis of Visual Stereotypes in Text-to-Image Generation}, author={Jha, Akshita and Prabhakaran, Vinodkumar and Denton, Remi and Laszlo, Sarah and Dave, Shachi and Qadri, Rida and Reddy, Chandan K and Dev, Sunipa}, journal={arXiv preprint arXiv:2401.06310}, year={2024} }




