dev_set_v2_a3_rl_DCAgent_exp_rpt_unitsyn_python_large_10_8B_20260825_204607
收藏资源简介:
该数据集包含多轮对话记录,每个样本由以下字段组成:conversations(对话历史,每条消息包含角色 role 和内容 content)、agent(代理名称)、model(模型名称)、model_provider(模型提供商)、date(日期)、task(任务描述)、episode(轮次)、run_id(运行ID)、trial_name(试验名称)、result(结果)、verifier_output(验证器输出)和 trace_source(跟踪来源)。训练集共有 3447 个样本,总数据量约为 270MB。该数据集适用于研究多轮对话、任务导向型代理行为、模型输出验证等场景。
This dataset contains multi-turn conversation records. Each sample consists of the following fields: conversations (conversation history, each message includes role and content), agent (agent name), model (model name), model_provider (model provider), date (date), task (task description), episode (episode number), run_id (run ID), trial_name (trial name), result (result), verifier_output (verifier output), and trace_source (trace source). The training set has 3,447 samples with a total data size of approximately 270MB. This dataset is suitable for research on multi-turn dialogue, task-oriented agent behavior, model output verification, etc.
数据集概述
该数据集名为 laion/dev_set_v2_a3_rl_DCAgent_exp_rpt_unitsyn_python_large_10_8B_20260825_204607,托管于 Hugging Face 平台,由 LAION 组织发布。数据集规模为 3447 条样本,总大小约 270.7 MB,下载大小约 234.9 MB,仅包含一个 train 分割。
数据字段
该数据集共包含 13 个字段:
conversations:对话列表,每个对话由role(角色)和content(内容)两部分组成。agent:智能体标识。model:使用的模型名称。model_provider:模型提供方。date:日期信息。task:任务描述。episode:回合编号。run_id:运行标识。trial_name:试验名称。result:结果记录。verifier_output:验证器输出。trace_source:追踪来源。
数据用途
从数据集命名(包含 rl、DCAgent、exp、unitsyn、python、8B 等关键词)和字段结构来看,该数据集主要用于强化学习(RL)场景下的智能体(Agent)训练与评估,重点关注单元同步、Python 编程任务,并以 8B 参数模型为基础。数据集中包含对话历史、任务结果、验证器输出等关键信息,适用于训练和评估具备代码生成与执行能力的智能体系统。




