swebench_verified_random_100_folders_a3_rl_DCAgent_exp_rpt_e2egit_large_15_8B_2026ddfac428
收藏资源简介:
该数据集包含5806个训练样本,总大小约947MB。每条样本记录了一个多轮对话(conversations),对话由不同角色(role)和内容(content)构成。此外,每条样本还提供了代理(agent)、模型(model)、模型提供商(model_provider)、日期(date)、任务(task)、轮次(episode)、运行ID(run_id)、试验名称(trial_name)、结果(result)、验证器输出(verifier_output)和追踪来源(trace_source)等元信息。该数据集适用于训练或评估基于多轮对话的AI代理,尤其关注不同模型在不同任务下的表现与验证结果。
This dataset contains 5,806 training samples with a total size of approximately 947 MB. Each sample records a multi-turn conversation (conversations) consisting of different roles and content. Additionally, each sample provides metadata such as agent, model, model provider, date, task, episode, run ID, trial name, result, verifier output, and trace source. The dataset is suitable for training or evaluating AI agents based on multi-turn conversations, with a particular focus on the performance and verification results of different models across various tasks.
数据集概述:swebench_verified_random_100_folders_a3_rl_DCAgent_exp_rpt_e2egit_large_15_8B_2026ddfac428
基本信息
该数据集来源于 Hugging Face 平台,路径为 laion/swebench_verified_random_100_folders_a3_rl_DCAgent_exp_rpt_e2egit_large_15_8B_2026ddfac428,是一个用于记录多轮对话及智能体(Agent)执行轨迹的实验数据集合。
数据规模
- 数据集总大小:947,082,877 字节(约 903 MB)
- 下载大小:497,707,537 字节(约 475 MB)
- 数据划分:仅包含
train划分- 样本数量:5,806 条
- 该划分大小:947,082,877 字节
数据特征(Features)
数据集中每条样本包含以下字段:
-
conversations(列表类型)
- 每项包含两个子字段:
role(字符串):对话角色content(字符串):对话内容
- 每项包含两个子字段:
-
agent(字符串):智能体名称或标识
-
model(字符串):使用的模型名称
-
model_provider(字符串):模型提供商
-
date(字符串):日期信息
-
task(字符串):任务描述
-
episode(字符串):回合或情景编号
-
run_id(字符串):运行标识符
-
trial_name(字符串):试验名称
-
result(字符串):执行结果
-
verifier_output(字符串):验证器输出
-
trace_source(字符串):轨迹来源
配置信息
- 配置名称:
default - 数据文件路径:
data/train-*(对应 train 划分)
内容特点
该数据集主要记录多轮对话(conversations)以及智能体(agent)在执行任务时的完整过程,包含模型信息、运行参数、验证结果等元数据,适合用于研究强化学习(RL)环境下的智能体行为分析、对话系统评估或代码生成与验证等场景。




