swebench_verified_random_100_folders_a1_stackexchange_overflow_20260818_162729
收藏资源简介:
该数据集包含多轮对话交互记录,每条样本包括对话内容(conversations,包含角色和消息文本)、代理名称(agent)、使用的模型(model)及模型提供商(model_provider)、日期(date)、任务描述(task)、轮次(episode)、运行ID(run_id)、试验名称(trial_name)、结果(result)、验证器输出(verifier_output)以及追踪来源(trace_source)。数据集共包含8556个训练样本,总大小约1GB,适用于对话代理评估、模型性能分析、多轮交互研究等任务。
This dataset contains multi-turn dialogue interaction records. Each sample includes conversation content (conversations, containing roles and message texts), agent name, model used and model provider, date, task description, episode, run ID, trial name, result, verifier output, and trace source. The dataset contains 8556 training samples with a total size of approximately 1GB, suitable for tasks such as dialogue agent evaluation, model performance analysis, and multi-turn interaction research.
数据集概述:swebench_verified_random_100_folders_a1_stackexchange_overflow_20260818_162729
基本信息
- 数据集地址:https://huggingface.co/datasets/laion/swebench_verified_random_100_folders_a1_stackexchange_overflow_20260818_162729
- 数据集规模:训练集包含 8,556 条样本,总大小约 1,040,079,246 字节(约 992 MB),下载大小约 768,691,428 字节(约 733 MB)。
数据特征
该数据集包含以下字段:
| 字段名 | 类型 | 说明 |
|---|---|---|
conversations |
列表(包含 role 和 content,均为字符串) |
对话内容,角色与内容成对出现 |
agent |
字符串 | 智能体标识 |
model |
字符串 | 使用的模型名称 |
model_provider |
字符串 | 模型提供方 |
date |
字符串 | 日期信息 |
task |
字符串 | 任务标识 |
episode |
字符串 | 回合/场景编号 |
run_id |
字符串 | 运行编号 |
trial_name |
字符串 | 试验名称 |
result |
字符串 | 结果信息 |
verifier_output |
字符串 | 验证器输出 |
trace_source |
字符串 | 轨迹来源 |
数据划分
- 仅包含 train 划分,共 8,556 条样本。
- 数据文件路径为
data/train-*,采用通配符匹配多个文件。
适用场景
该数据集包含对话轨迹及多种元数据(如智能体、模型、任务、验证结果等),适用于以下研究和应用场景:
- 智能体行为分析与评估
- 多轮对话建模与生成
- 强化学习或模仿学习中的轨迹数据处理
- 模型性能验证与结果分析




