terminal_bench_2_a3_rl_laion_exp_rpt_ghactions_v3_20_8B_20260829_084542
收藏资源简介:
该数据集包含多轮对话交互记录,每条数据包括一个完整的对话序列(conversations),其中每轮对话包含角色(role)和内容(content)。此外,还记录了参与对话的智能体(agent)、使用的模型(model)和模型提供商(model_provider)、日期(date)、任务(task)、回合(episode)、运行ID(run_id)、试验名称(trial_name)、结果(result)、验证器输出(verifier_output)以及跟踪来源(trace_source)。数据集共包含3727个训练样本,总大小约288MB。该数据集适用于对话系统评估、多智能体交互分析、模型行为追踪等任务。
The dataset contains multi-turn dialogue interaction records, each data including a complete conversation sequence (conversations) with role and content for each turn. Additionally, it records the agent, model, model provider, date, task, episode, run ID, trial name, result, verifier output, and trace source. The dataset has 3,727 training samples with a total size of approximately 288MB. It is suitable for tasks such as dialogue system evaluation, multi-agent interaction analysis, and model behavior tracking.
数据集概述
该数据集名为 laion/terminal_bench_2_a3_rl_laion_exp_rpt_ghactions_v3_20_8B_20260829_084542,托管于 Hugging Face。
数据规模
- 总大小:约 288.34 MB(下载大小约 233.20 MB)
- 样本数量:3,727 条(训练集)
数据特征
每个样本包含以下字段:
| 字段名 | 类型 | 说明 |
|---|---|---|
conversations |
列表(含 role 和 content 字符串) |
对话记录,包含角色与内容 |
agent |
字符串 | 智能体标识 |
model |
字符串 | 使用的模型 |
model_provider |
字符串 | 模型提供商 |
date |
字符串 | 日期 |
task |
字符串 | 任务名称 |
episode |
字符串 | 回合编号 |
run_id |
字符串 | 运行 ID |
trial_name |
字符串 | 试验名称 |
result |
字符串 | 结果 |
verifier_output |
字符串 | 验证器输出 |
trace_source |
字符串 | 追踪来源 |
数据划分
- 该数据集仅包含一个划分:
train,包含全部 3,727 条样本。
数据内容特点
- 数据涉及终端环境下的基准测试任务(
terminal_bench),可能包含强化学习(RL)相关的实验数据。 - 文件名中提及
20_8B,暗示可能涉及 80 亿参数规模的模型实验。 - 包含对话记录、智能体行为、任务执行结果等结构化信息,适用于分析智能体在终端任务中的表现。




