eval-fsr-a1-stack-rspec-swe-r442-rf0710-traces
收藏资源简介:
该数据集是一个用于记录和评估对话系统交互轨迹的数据集。它包含多轮对话数据,每个样本由对话内容、参与角色(如用户或助手)、使用的代理、模型及其提供商、日期、任务类型、回合编号、运行ID、试验名称、结果输出、验证器输出和追踪来源等字段构成。数据集旨在支持对话系统的性能分析、任务完成度评估或交互行为研究,适用于自然语言处理中的对话建模、代理评估和任务导向对话等场景。数据规模为5423个训练样本,总大小约842MB。
This dataset is designed for recording and evaluating interaction trajectories in dialogue systems. It contains multi-turn dialogue data, with each sample comprising fields such as dialogue content, participant roles (e.g., user or assistant), used agents, models and their providers, date, task type, turn number, run ID, trial name, result output, validator output, and trace source. The dataset aims to support performance analysis, task completion evaluation, or interaction behavior research for dialogue systems, and is applicable to scenarios like dialogue modeling, agent evaluation, and task-oriented dialogue in natural language processing. The data scale includes 5423 training samples, with a total size of approximately 842MB.
数据集概述
- 数据集名称:eval-fsr-a1-stack-rspec-swe-r442-rf0710-traces
- 数据集来源:LAION
- 数据集地址:https://huggingface.co/datasets/laion/eval-fsr-a1-stack-rspec-swe-r442-rf0710-traces
数据集特征
数据集包含以下字段:
| 字段名 | 类型 | 说明 |
|---|---|---|
| conversations | list of objects | 对话内容列表,每个对象包含 content(字符串)和 role(字符串) |
| agent | string | 代理标识 |
| model | string | 模型名称 |
| model_provider | string | 模型提供者 |
| date | string | 日期 |
| task | string | 任务描述 |
| episode | string | 轮次标识 |
| run_id | string | 运行ID |
| trial_name | string | 试验名称 |
| result | string | 结果 |
| verifier_output | string | 验证器输出 |
| trace_source | string | 追踪来源 |
数据集划分
- 训练集(train):
- 样本数量:5423
- 数据大小:841684718 字节
数据集大小
- 下载大小:664973359 字节
- 数据集总大小:841684718 字节
配置文件
- 配置名称:default
- 数据文件路径:data/train-*




