数据链接:
官方服务:
资源简介:
Discriminatory Keywords Dictionary (DKD).
应用场景:
创建时间:
2020-06-04
相关数据集
LanguageShades/BiasShadesBaseEval_bigscience_bloom-7b1
该数据集包含多语言偏见句子及其相关信息,涵盖了多种语言的偏见类型、偏见来源语言、偏见有效语言、偏见有效地区、被刻板印象的群体、偏见句子、偏见模板、是否为表达、评论、以及使用bigscience_bloom-7b1模型生成的token和logprob等信息。数据集包含阿拉伯语、孟加拉语、葡萄牙语、中文、荷兰语、英语、法语、德语、印地语、意大利语、马拉地语、波兰语、罗马尼亚语、俄语和西班牙语等多种语言
Hugging Face2024-06-15 更新160
SpoofLab: a framework for audio deepfake and spoofing detection evaluation
This dataset corresponds to the SpoofLab framework, an open-source research framework developed for the study of audio deepfake and spoofing detection systems. The repository provides implementations
DataCite Commons2026-03-26 更新100
iproskurina/bias-in-bios-qwen-hf-r30-contamination-np-iter-iter5
--- dataset_info: features: - name: text dtype: string splits: - name: train num_bytes: 10530002 num_examples: 27752 download_size: 6733478 dataset_size: 10530002 configs: - co
Hugging Face2026-04-27 更新50
Older Workers Need Not Apply? Ageist Language in Job Ads and Age Discrimination in Hiring
We study the relationships between ageist stereotypes as reflected in the language used in job ads and age discrimination in hiring, exploiting the text of job ads and differences in callbacks to ol
NBER2019-12-01 更新140
Nachiket-S/train_dataset
该数据集包含多个特征字段,涵盖了文本数据、性别描述、偏好描述、名词及其复数形式、名词短语、输入ID、注意力掩码、文本内容、模板、是否仅第一轮、是否必须是名词、未命名字段、更多句子、更少句子、刻板印象与反刻板印象、偏见类型、注释、匿名作者、匿名注释者、偏见文本、偏见粗俗词汇、去偏见文本、上下文、句子等。数据集分为训练集,包含463,843个样本,总大小为221,609,906字节,下载大小为42,6
Hugging Face2024-11-30 更新80



