terminal_bench_2_a1_stackexchange_overflow_20260810_001300
收藏资源简介:
该数据集包含多轮对话记录,每条样本包括对话历史(conversations,由 role 和 content 字段组成)、agent 标识、模型名称、模型提供商、日期、任务、episode 编号、run_id、trial_name、结果、验证输出以及追踪来源。数据集共包含 5277 条训练样本,总大小约 782MB。适用于对话系统研究、agent 行为分析、模型评估等任务。
This dataset contains multi-turn conversation records, with each sample including conversation history (composed of role and content fields), agent identifier, model name, model provider, date, task, episode number, run_id, trial_name, result, validation output, and trace source. The dataset consists of 5277 training samples with a total size of approximately 782MB. It is suitable for tasks such as dialogue system research, agent behavior analysis, and model evaluation.
数据集概述:laion/terminal_bench_2_a1_stackexchange_overflow_20260810_001300
基本信息
- 数据集名称:terminal_bench_2_a1_stackexchange_overflow_20260810_001300
- 所属组织:LAION
- 数据集地址:https://huggingface.co/datasets/laion/terminal_bench_2_a1_stackexchange_overflow_20260810_001300
数据规模
- 数据集总大小:782,274,601 字节(约 782 MB)
- 下载大小:406,686,737 字节(约 407 MB)
- 数据划分:仅包含训练集(train)
- 训练集样本数:5,277 条
数据特征(字段结构)
每条数据包含以下字段:
| 字段名 | 类型 | 说明 |
|---|---|---|
| conversations | 列表 | 对话记录,包含两个子字段 |
| - role | 字符串 | 角色(如用户或助手) |
| - content | 字符串 | 对话内容 |
| agent | 字符串 | 智能体名称 |
| model | 字符串 | 使用的模型名称 |
| model_provider | 字符串 | 模型提供方 |
| date | 字符串 | 日期信息 |
| task | 字符串 | 任务类型 |
| episode | 字符串 | 回合编号 |
| run_id | 字符串 | 运行标识 |
| trial_name | 字符串 | 试验名称 |
| result | 字符串 | 结果信息 |
| verifier_output | 字符串 | 验证器输出 |
| trace_source | 字符串 | 追踪来源 |
数据用途
该数据集属于 terminal_bench 系列,聚焦于终端命令行基准测试场景,数据来源于 StackExchange Overflow 平台,用于评估和训练智能体在终端环境下的任务执行能力。




