遇见数据集

scottgeng00/olmo-3-preference-mix-deltas-100k_base_complement-only_chat

收藏
Hugging Face2025-09-04 更新2025-10-25 收录
官方服务:

资源简介:

该数据集包含多个字段,如提示文本(prompt)、选择(chosen)的内容、国家(country)、语言(language)、是否涉黄(toxic)等。数据集分为训练集(train),包含大约1270959个样本,总大小约为11GB。数据集的具体内容和用途未在README中描述。

The dataset includes multiple fields such as prompt text, chosen content, country, language, toxicity flag, etc. The dataset is split into a training set (train) containing approximately 1,270,959 samples, with a total size of about 11GB. The specific content and purpose of the dataset are not described in the README.

提供机构:
scottgeng00
二维码
社区交流群
二维码
科研交流群
商业服务