swebench_verified_random_100_folders_a1_stackexchange_tor_20260819_162242
收藏资源简介:
该数据集包含6925个样本,每个样本记录了AI代理或模型在特定任务中的完整交互轨迹。数据字段包括多轮对话(conversations,包含角色和内容)、代理名称(agent)、模型名称(model)、模型提供商(model_provider)、日期(date)、任务类型(task)、轮次(episode)、运行ID(run_id)、试验名称(trial_name)、结果(result)、验证器输出(verifier_output)以及轨迹来源(trace_source)。数据集仅包含训练集,总大小约1.63GB,适用于对话系统、模型评估、推理验证或强化学习策略分析等场景。
This dataset contains 6,925 samples, each recording the complete interaction trajectory of an AI agent or model in a specific task. Data fields include multi-turn conversations (conversations, containing role and content), agent name (agent), model name (model), model provider (model_provider), date (date), task type (task), episode (episode), run ID (run_id), trial name (trial_name), result (result), verifier output (verifier_output), and trace source (trace_source). The dataset only contains a training set, with a total size of approximately 1.63 GB, suitable for scenarios such as dialogue systems, model evaluation, reasoning verification, or reinforcement learning strategy analysis.
数据集概述
基本信息
- 名称:
laion/swebench_verified_random_100_folders_a1_stackexchange_tor_20260819_162242 - 地址:https://huggingface.co/datasets/laion/swebench_verified_random_100_folders_a1_stackexchange_tor_20260819_162242
- 数据集大小:1,632,886,597 字节(约 1.63 GB)
- 下载大小:724,653,344 字节(约 724.65 MB)
数据划分
- 训练集(train):包含 6,925 个样本,大小为 1,632,886,597 字节
数据特征
数据集包含以下字段:
| 字段名 | 类型 | 说明 |
|---|---|---|
conversations |
列表(含 role 和 content 两个子字段,均为字符串) |
对话记录 |
agent |
字符串 | 智能体标识 |
model |
字符串 | 使用的模型 |
model_provider |
字符串 | 模型提供方 |
date |
字符串 | 日期 |
task |
字符串 | 任务描述 |
episode |
字符串 | 回合编号 |
run_id |
字符串 | 运行标识 |
trial_name |
字符串 | 试验名称 |
result |
字符串 | 结果 |
verifier_output |
字符串 | 验证器输出 |
trace_source |
字符串 | 轨迹来源 |
数据集配置
- 配置名称:
default - 数据文件:训练集位于
data/train-*
数据集特点
该数据集属于 SWE-bench Verified 相关项目的高质量子集,包含随机的 100 个文件夹样本,附加了 StackExchange 相关的数据,并带有 Tor 网络标识,推测用于特定场景下的智能体评估与训练。




