遇见数据集

mlfoundations-dev/seed_code_multiple_samples_all_scale_up_2K

收藏
Hugging Face2025-02-19 更新2025-04-12 收录
官方服务:

资源简介:

该数据集包含了多个字段,如问题(problem)、来源(source)、领域(domain)、模型响应(r1_distill_70b_response)、原始行索引(__original_row_idx)、多数响应(_majority_responses)、验证过的模型响应(verified_r1_distill_70b_response)和对话(conversations)。对话字段包含了对话的发送者(from)和消息内容(value)。数据集分为训练集,共有6000个示例。数据集的总大小为约885MB。

The dataset includes fields such as problem, source, domain, model response (r1_distill_70b_response), original row index (__original_row_idx), majority response (_majority_responses), verified model response (verified_r1_distill_70b_response), and conversation. The conversation field contains the sender (from) and message content (value) of the conversation. The dataset is split into a training set with a total of 6000 examples. The total size of the dataset is approximately 885MB.

提供机构:
mlfoundations-dev
二维码
社区交流群
二维码
科研交流群
商业服务