dev_set_v2_a3_rl_DCAgent2_nl2bash_tasks_cleaned_oracle_40_8B_20260827_004039
收藏资源简介:
该数据集包含多轮对话记录,每条数据包括对话内容(conversations,由角色和内容组成)、代理名称(agent)、模型名称(model)、模型提供商(model_provider)、日期(date)、任务(task)、轮次(episode)、运行ID(run_id)、试验名称(trial_name)、结果(result)、验证器输出(verifier_output)以及轨迹来源(trace_source)等字段。训练集共有4270个样本,总大小约309MB。该数据集可用于对话系统训练、模型行为分析、推理过程追踪等场景。
This dataset contains multi-turn conversation records, with each data entry including fields such as conversation content (conversations, composed of roles and content), agent name (agent), model name (model), model provider (model_provider), date (date), task (task), episode (episode), run ID (run_id), trial name (trial_name), result (result), verifier output (verifier_output), and trace source (trace_source). The training set has 4,270 samples, totaling approximately 309 MB. This dataset can be used for dialogue system training, model behavior analysis, reasoning process tracing, and other scenarios.
数据集概述
该数据集名为 dev_set_v2_a3_rl_DCAgent2_nl2bash_tasks_cleaned_oracle_40_8B,由 LAION 组织发布,是一个用于训练和评估对话代理(特别是 DCAgent2)在 nl2bash 任务上的开发集(dev set)。数据集规模约 309 MB,包含 4,270 条训练样本。
数据特征
每条样本包含以下字段:
- conversations:对话内容,是一个列表,每条消息包含角色(role)和内容(content)两个字段。
- agent:执行任务的代理名称。
- model:使用的模型名称。
- model_provider:模型提供方。
- date:数据生成日期。
- task:任务名称。
- episode:会话轮次。
- run_id:运行标识。
- trial_name:试验名称。
- result:任务执行结果。
- verifier_output:验证器输出。
- trace_source:轨迹来源。
数据划分
数据集仅包含一个 train 划分,共 4,270 条样本,占用约 309 MB 存储空间(压缩后约 271 MB)。该数据集未提供验证集或测试集。
用途说明
该数据集专注于自然语言到 bash 脚本转换(nl2bash)任务,结合了强化学习(RL)生成的轨迹,适合用于训练或评估基于对话的代码生成代理模型。




