遇见数据集

UkCat/PyQwen-7B-Over-Compliance

收藏
Hugging Face2026-05-25 更新2026-05-31 收录
官方服务:

资源简介:

该数据集是一个用于自然语言处理任务的数据集,包含50,000个训练样本,每个样本由三个字符串特征组成:prompt(提示)、chosen(选择的回答)和rejected(拒绝的回答)。数据集可能用于训练或评估模型在比较不同回答时的性能,例如在对话生成或强化学习场景中。数据以训练集形式组织,总大小约190MB,下载大小约90MB。

This dataset is designed for natural language processing tasks, consisting of 50,000 training examples. Each example includes three string features: prompt, chosen, and rejected. It is likely used for training or evaluating models in scenarios that involve comparing different responses, such as in dialogue generation or reinforcement learning. The dataset is organized into a single train split, with a total size of approximately 190MB and a download size of about 90MB.

提供机构:
UkCat
二维码
社区交流群
二维码
科研交流群
商业服务