terminal_bench_2_a3_rl_laion_nemotron_gym_agent_calendar_80_8B_20260829_094705
收藏资源简介:
该数据集是一个结构化的多轮对话数据集,包含3191个训练样本。每条样本包含以下字段:多轮对话(角色和内容)、agent 名称、模型名称、模型提供者、日期、任务、episode、run_id、trial_name、结果、验证器输出以及跟踪来源。数据可用于训练和评估语言模型在多轮对话、agent 交互与任务完成等方面的能力,尤其适用于基于对话的智能体行为分析、模型性能对比和验证器输出分析。
This dataset is a structured multi-turn dialogue dataset containing 3,191 training samples. Each sample includes the following fields: multi-turn dialogue (role and content), agent name, model name, model provider, date, task, episode, run_id, trial_name, result, validator output, and trace source. The data can be used to train and evaluate language models in multi-turn dialogue, agent interaction, and task completion, especially suitable for dialogue-based agent behavior analysis, model performance comparison, and validator output analysis.
数据集概述
该数据集名为 laion/terminal_bench_2_a3_rl_laion_nemotron_gym_agent_calendar_80_8B_20260829_094705,由 LAION 组织发布,托管于 Hugging Face 平台。
基本信息
- 数据集大小:约 245 MB(下载大小约 198 MB,解压后约 245 MB)
- 数据分割:仅包含训练集(train),共 3,191 个样本
- 特色字段:总共包含 12 个字段
数据结构
每个样本包含以下字段:
| 字段名 | 数据类型 | 说明 |
|---|---|---|
conversations |
列表(含 role 和 content 两个子字段) |
对话记录,记录角色与内容 |
agent |
字符串 | 智能体标识 |
model |
字符串 | 使用的模型名称 |
model_provider |
字符串 | 模型提供方 |
date |
字符串 | 日期信息 |
task |
字符串 | 任务类型 |
episode |
字符串 | 回合编号 |
run_id |
字符串 | 运行标识符 |
trial_name |
字符串 | 试验名称 |
result |
字符串 | 执行结果 |
verifier_output |
字符串 | 验证器输出 |
trace_source |
字符串 | 追踪来源 |
数据集配置
- 配置名称:
default - 数据文件路径:
data/train-*(采用通配符匹配多个文件)
可能用途
根据字段命名(如 agent、verifier_output、trial_name 等),该数据集可能用于评估或训练智能体(agent)在终端环境中的任务执行能力,特别是针对日历管理类任务,属于强化学习或智能体行为分析领域的数据资源。




