遇见数据集

scottgeng00/olmo-3-preference-mix-deltas_reasoning-yolo_scottmix

收藏
Hugging Face2025-09-15 更新2025-10-25 收录
官方服务:

资源简介:

这是一个包含文本数据的训练集,其中每个样本包括一个提示(prompt)、一个被选中的文本内容(chosen)及其角色(role)、一个被拒绝的文本内容(rejected)及其角色、选择的模型(chosen_model)、拒绝的模型(rejected_model)、数据集名称(dataset)、提示ID(prompt_id)和类别(category)。数据集被划分为训练集(train),包含大约294909个示例,总大小约为6.63GB。

This is a training dataset containing text data, where each sample includes a prompt, a chosen text content with its role, a rejected text content with its role, the chosen model, the rejected model, the dataset name, prompt ID, and category. The dataset is split into a training set (train) with approximately 294909 examples and a total size of about 6.63GB.

提供机构:
scottgeng00
二维码
社区交流群
二维码
科研交流群
商业服务