dev_set_v2_a1_stack_rspec_20260811_201503
收藏资源简介:
该数据集包含 4606 个训练样本,总大小约 493 MB。每个样本记录了一次完整的对话交互过程,包括对话历史(conversations,由角色和内容组成)、使用的代理(agent)、模型(model)及模型提供商(model_provider)、日期(date)、任务(task)、剧集(episode)、运行 ID(run_id)、试验名称(trial_name)、结果(result)、验证器输出(verifier_output)以及跟踪来源(trace_source)。数据集可用于研究多轮对话中代理的行为、模型性能评估、任务完成情况分析等场景。
This dataset contains 4606 training samples with a total size of approximately 493 MB. Each sample records a complete dialogue interaction process, including conversation history (conversations, composed of role and content), the agent used (agent), the model (model) and model provider (model_provider), date (date), task (task), episode (episode), run ID (run_id), trial name (trial_name), result (result), verifier output (verifier_output), and trace source (trace_source). The dataset can be used for studying agent behavior in multi-turn dialogues, model performance evaluation, task completion analysis, and other scenarios.
数据集详情
数据集名称: laion/dev_set_v2_a1_stack_rspec_20260811_201503
数据集地址: https://huggingface.co/datasets/laion/dev_set_v2_a1_stack_rspec_20260811_201503
数据规模
- 数据集总大小: 493,174,948 字节(约 493 MB)
- 下载大小: 392,794,591 字节(约 393 MB)
- 分割 (Split): 仅包含
train分割 - 训练集样本数: 4,606 条
数据特征 (Features)
该数据集包含以下字段:
| 字段名 | 类型 | 说明 |
|---|---|---|
conversations |
列表 | 对话记录,每个元素包含 role(角色,字符串)和 content(内容,字符串) |
agent |
string | 智能体标识 |
model |
string | 使用的模型 |
model_provider |
string | 模型提供方 |
date |
string | 日期 |
task |
string | 任务类型 |
episode |
string | 回合编号 |
run_id |
string | 运行标识 |
trial_name |
string | 试验名称 |
result |
string | 结果 |
verifier_output |
string | 验证器输出 |
trace_source |
string | 追踪来源 |
配置信息
- 配置名称:
default - 数据文件路径:
data/train-*(通配符匹配多个文件)
数据用途
该数据集为开发验证集(dev set),版本标识为 v2_a1_stack_rspec_20260811_201503,包含对话式交互数据及多维度元信息,适用于训练、评估或分析基于智能体的对话系统、模型输出验证及相关研究任务。




