terminal_bench_2_a3_rl_laion_nemotron_gym_math_advanced_calculations_v3_60_8B_20261ba9adc8
收藏资源简介:
该数据集包含AI agent与用户之间的对话交互记录,每条数据包括对话内容(conversations,由角色role和内容content组成)、agent名称、模型名称、模型提供商、日期、任务类型、轮次(episode)、运行ID、试验名称、结果、验证输出以及追踪来源。数据集共3147个训练样本,总大小约232MB。适用于对话系统评估、agent行为分析、模型输出质量研究、多轮对话理解等任务。
This dataset contains conversation interaction records between AI agents and users. Each data entry includes conversation content (conversations, consisting of role and content), agent name, model name, model provider, date, task type, episode, run ID, experiment name, result, validation output, and trace source. The dataset has a total of 3147 training samples, with a total size of approximately 232MB. It is suitable for tasks such as dialogue system evaluation, agent behavior analysis, model output quality research, and multi-turn dialogue understanding.
数据集概述
该数据集为 laion/terminal_bench_2_a3_rl_laion_nemotron_gym_math_advanced_calculations_v3_60_8B_20261ba9adc8,来源于 Hugging Face 数据集平台。
核心内容
这是一个包含 3,147 条训练样本 的对话数据集,主要用于强化学习或数学高级计算相关的智能体(Agent)训练与评估任务。数据集总大小约 222 MB(下载大小约 173 MB)。
数据字段
数据集每条样本包含以下字段:
| 字段名 | 类型 | 说明 |
|---|---|---|
conversations |
列表 | 对话内容,包含 role(角色)和 content(内容)两个子字段 |
agent |
字符串 | 智能体标识 |
model |
字符串 | 使用的模型名称 |
model_provider |
字符串 | 模型提供方 |
date |
字符串 | 日期信息 |
task |
字符串 | 任务类型 |
episode |
字符串 | 回合编号 |
run_id |
字符串 | 运行标识 |
trial_name |
字符串 | 试验名称 |
result |
字符串 | 结果输出 |
verifier_output |
字符串 | 验证器输出 |
trace_source |
字符串 | 追踪来源 |
数据划分
- 唯一划分:
train训练集,包含 3,147 条样本,数据格式为data/train-*。
适用方向
该数据集适用于强化学习、数学高级计算任务、智能体训练与评估、对话生成等领域的研究与开发。




