TMax-SFT-16.5K
收藏资源简介:
TMax SFT是一个用于监督式微调(SFT)的数据集,源自Tmax研究项目。该数据集包含了Qwen 3.6 27B大型语言模型在大约2000个独立生成的环境中所产生的交互轨迹(traces),旨在为研究和教育目的提供训练数据,特别适用于开发或微调终端智能体(terminal agents)相关的任务。数据遵循ODC-BY许可证,并需在遵守Ai2负责任使用指南及谷歌服务条款(因部分数据由Gemini 3.1 Pro生成)的前提下使用。
TMax SFT is a dataset for supervised fine-tuning (SFT), derived from the Tmax research project. It contains interaction traces generated by the Qwen 3.6 27B large language model in approximately 2000 independently generated environments. The dataset is designed to provide training data for research and educational purposes, particularly suitable for tasks related to developing or fine-tuning terminal agents. The data follows the ODC-BY license and must be used in compliance with the Ai2 Responsible Use Guidelines and Googles Terms of Service (as some data was generated by Gemini 3.1 Pro).
数据集概述:TMax-SFT-16.5K
该数据集由 Allen AI 发布,用于训练终端的简单代理(terminal agents)。
- 许可证:ODC-BY (Open Data Commons Attribution License)
- 语言:英语 (en)
数据集内容
- 包含来自 Qwen 3.6 27B 模型在 Tmax 生成的约 2000 个环境中的轨迹数据。
- 构成 SFT(监督式微调) 数据集,名称为 TMax SFT。
- 另有适用于 open-instruct 工具的数据集版本,位于
https://huggingface.co/datasets/allenai/tmax-sft。
许可与用途
- 依据 Ai2 的 负责任使用指南,仅供研究和教育用途。
- 数据中包含使用 Gemini 3.1 Pro 生成的输出,受 Google 服务条款约束。
引用
如需使用该模型或数据,请引用论文:https://arxiv.org/abs/2606.23321 (Hamish Ivison 等人,2026年)。




