swebench_verified_random_100_folders_a1_taco_20260819_162248
收藏资源简介:
该数据集包含11580个训练样本,每个样本主要由一个多轮对话(conversations)组成,对话中每条消息包含角色(role)和内容(content)。此外,每个样本还附带了丰富的元数据字段:agent(代理名称)、model(使用的模型)、model_provider(模型提供商)、date(日期)、task(任务)、episode(回合编号)、run_id(运行ID)、trial_name(试验名称)、result(结果)、verifier_output(验证器输出)以及trace_source(追踪来源)。数据规模约为1.5GB。该数据集适用于多轮对话系统的训练、agent行为分析、模型评估等任务。
This dataset contains 11,580 training samples, each primarily consisting of a multi-turn conversation, where each message in the conversation includes a role and content. Additionally, each sample is accompanied by rich metadata fields: agent, model, model_provider, date, task, episode, run_id, trial_name, result, verifier_output, and trace_source. The data size is approximately 1.5GB. The dataset is suitable for tasks such as training multi-turn dialogue systems, agent behavior analysis, and model evaluation.
数据集详情总结:swebench_verified_random_100_folders_a1_taco_20260819_162248
基本信息
- 数据集名称:swebench_verified_random_100_folders_a1_taco_20260819_162248
- 所属机构:LAION
- 数据集页面:https://huggingface.co/datasets/laion/swebench_verified_random_100_folders_a1_taco_20260819_162248
- 数据集大小:约1.58 GB(1578847669 字节)
- 下载大小:约1.13 GB(1125506359 字节)
- 数据分割:仅包含训练集(train),共11,580条样本
数据特征(Features)
该数据集包含以下字段:
| 字段名 | 类型 | 说明 |
|---|---|---|
| conversations | 列表(包含role和content两字段) | 对话记录,role为字符串类型,content为字符串类型 |
| agent | 字符串 | 智能体名称 |
| model | 字符串 | 使用的模型 |
| model_provider | 字符串 | 模型提供方 |
| date | 字符串 | 日期 |
| task | 字符串 | 任务描述 |
| episode | 字符串 | 回合/轮次标识 |
| run_id | 字符串 | 运行ID |
| trial_name | 字符串 | 试验名称 |
| result | 字符串 | 结果 |
| verifier_output | 字符串 | 验证器输出 |
| trace_source | 字符串 | 追踪来源 |
数据格式
- 配置文件:default
- 数据文件路径:data/train-*
- 数据分割类型:仅提供train分割
数据用途概述
该数据集来自LAION,涉及SWE-bench验证任务,包含随机选取的100个文件夹,其命名中的“a1_taco”和“20260819_162248”暗示了特定的实验配置和时间戳。数据集记录了智能体(agent)在软件开发任务中的完整对话轨迹、运行信息、模型信息及验证结果,适用于研究智能体在编程任务中的行为分析、模型评估和对话追踪。




