GenderAlign
收藏资源简介:
GenderAlign是一个包含8,362个单轮对话的数据集,旨在减轻语言模型中的广泛性别偏见。每个条目包含一对文本,一个是“选定的”,另一个是“拒绝的”。数据集包含可能令人不快或冒犯的内容,包括性别偏见、性别刻板印象、基于性别的暴力和其他可能令人不安的主题。
GenderAlign is a dataset containing 8,362 single-turn dialogues, designed to mitigate widespread gender bias in language models. Each entry consists of a pair of texts, one labeled as "selected" and the other as "rejected". The dataset contains potentially unpleasant or offensive content, including gender bias, gender stereotypes, gender-based violence, and other potentially disturbing topics.
GenderAlign: 用于减轻大型语言模型中性别偏见的对齐数据集
数据集描述
该数据集在论文 "GenderAlign: An Alignment Dataset for Mitigating Gender Bias in Large Language Models" 中进行了描述。如果您发现该数据集有用,请引用该论文。数据集格式非常简单——每个条目包含一对文本,一个是“选定的”,另一个是“拒绝的”。
免责声明
数据集包含可能具有冒犯性或令人不安的内容。主题包括但不限于性别偏见、性别刻板印象、基于性别的暴力和其他可能令人不安的主题。请根据您的个人风险承受能力与数据集互动。该数据集旨在用于研究目的,特别是旨在减少模型中性别偏见的研究。数据中表达的观点并不反映作者的观点。




