eval-fsr-a1-repo-scaffold-swe-r463-rf0710-traces
收藏资源简介:
该数据集是一个用于记录和评估AI代理交互与性能的结构化数据集。数据以多轮对话形式组织,每个样本包含对话内容(conversations,其中每条消息有内容(content)和角色(role)字段),以及相关的元数据,如代理类型(agent)、使用的模型(model)、模型提供方(model_provider)、日期(date)、任务类型(task)、剧集标识(episode)、运行ID(run_id)、试验名称(trial_name)、任务结果(result)、验证器输出(verifier_output)和追踪来源(trace_source)。数据集规模为4982个训练样本,总大小约939MB,适用于AI代理评估、对话系统分析、任务性能追踪等应用场景。
This dataset is a structured dataset for recording and evaluating AI agent interactions and performance. The data is organized in multi-turn dialogue format, with each sample containing dialogue content (conversations, where each message has content and role fields) and related metadata, such as agent type, model used, model provider, date, task type, episode identifier, run ID, trial name, task result, verifier output, and trace source. The dataset consists of 4982 training samples, with a total size of approximately 939MB, and is suitable for applications such as AI agent evaluation, dialogue system analysis, and task performance tracking.
数据集概述:laion/eval-fsr-a1-repo-scaffold-swe-r463-rf0710-traces
- 数据集地址:https://huggingface.co/datasets/laion/eval-fsr-a1-repo-scaffold-swe-r463-rf0710-traces
数据集特征
该数据集包含以下字段:
- conversations:对话列表,每个对话包含两个子字段:
content:字符串类型,对话内容。role:字符串类型,角色(如用户或助手)。
- agent:字符串类型,代理(agent)名称。
- model:字符串类型,使用的模型。
- model_provider:字符串类型,模型提供者。
- date:字符串类型,日期。
- task:字符串类型,任务描述。
- episode:字符串类型,迭代轮次(episode)。
- run_id:字符串类型,运行ID。
- trial_name:字符串类型,试验名称。
- result:字符串类型,结果。
- verifier_output:字符串类型,验证器输出。
- trace_source:字符串类型,追踪来源。
数据集划分与规模
- 训练集 (train):
- 样本数:4982条
- 文件大小:939,123,946 字节(约939 MB)
总下载大小
- 总下载大小:626,943,380 字节(约627 MB)
数据集配置
- 配置名称:default
- 数据文件:
- 训练集数据文件路径:
data/train-*(通配符匹配多个文件)
- 训练集数据文件路径:




