遇见数据集

saurabh5/olmo-3-preference-mix-deltas_reasoning-yolo_even_split-DECON-no-chinese

收藏
Hugging Face2025-09-24 更新2025-10-25 收录
官方服务:

资源简介:

该数据集包含了提示文本(prompt)、选中内容(chosen)和未选中内容(rejected)等字段。选中内容和未选中内容都包含内容和角色两个子字段。此外,数据集还记录了选中模型(chosen_model)和拒绝模型(rejected_model)的信息,以及数据来源(dataset)和提示文本的ID(prompt_id)。数据集的训练集大小为11818146822.161951字节,共有526135个示例。

The dataset includes fields such as prompt, chosen, and rejected. The chosen and rejected fields contain sub-fields for content and role. Additionally, the dataset records information about the chosen model (chosen_model) and the rejected model (rejected_model), the source of the data (dataset), and the ID of the prompt (prompt_id). The training set of the dataset is 11818146822.161951 bytes in size and contains 526135 examples.

提供机构:
saurabh5
二维码
社区交流群
二维码
科研交流群
商业服务