遇见数据集

aieng-lab/gradiend_religion_data

收藏
Hugging Face2025-10-14 更新2025-10-25 收录
官方服务:

资源简介:

这是一个包含敏感宗教词汇的模板句子的数据集,例如“犹太人”。该数据集是Wikipedia-10的过滤版本,仅包含包含源宗教的偏见敏感词汇的句子。基于掩码术语(源),从相应的偏见属性对中派生出一个关联的目标,匹配源的大小写(例如,“JEWISH”对应“CHRISTIAN”而不是“christian”)。

This dataset consists of templated sentences with the masked word being sensitive to religion, e.g., *Jewish*. It is a filtered version of Wikipedia-10 containing only sentences that include a religion bias sensitive word of the `source_id` religion. Based on the masked term (`source`), an associated `target` is derived from a corresponding bias attribute pair, matching the casing of `source` (e.g., `JEWISH` gets to `CHRISTIAN` and not `christian`).

提供机构:
aieng-lab
二维码
社区交流群
二维码
科研交流群
商业服务