dev_set_v2_a1_stack_phpunit_20260811_174202
收藏资源简介:
该数据集包含多轮对话记录,共4725条训练样本,总大小约445MB。每条样本包含一个对话列表(conversations),其中每条消息由角色(role)和内容(content)组成。此外,还记录了执行对话的智能体(agent)、使用的模型(model)及模型提供方(model_provider)、日期(date)、任务类型(task)、回合编号(episode)、运行ID(run_id)、试验名称(trial_name)、结果(result)、验证器输出(verifier_output)以及追踪来源(trace_source)。该数据集可用于训练或评估对话智能体、多轮对话系统、模型行为分析等任务。
This dataset contains multi-turn dialogue records, with a total of 4725 training samples and a size of approximately 445MB. Each sample includes a conversation list (conversations), where each message consists of a role and content. Additionally, it records the agent executing the dialogue, the model used and its provider, date, task type, episode number, run ID, trial name, result, verifier output, and trace source. The dataset can be used for training or evaluating dialogue agents, multi-turn dialogue systems, model behavior analysis, and other tasks.
数据集概述:dev_set_v2_a1_stack_phpunit_20260811_174202
- 数据集名称:dev_set_v2_a1_stack_phpunit_20260811_174202
- 所属机构/作者:LAION
- 数据集地址:https://huggingface.co/datasets/laion/dev_set_v2_a1_stack_phpunit_20260811_174202
数据集内容
该数据集主要服务于特定开发任务(PHPUnit 相关),包含多轮对话记录及运行结果信息。每条样本包含:
- 对话记录(conversations):由 role(角色)和 content(内容)组成的多轮对话列表;
- 代理信息(agent):记录执行任务的具体代理;
- 模型信息(model / model_provider):记录所使用的模型及其提供方;
- 日期与任务信息:包括日期(date)、任务(task)、轮次(episode);
- 运行标识:run_id、trial_name(试验名称);
- 执行结果:result(最终结果)与 verifier_output(验证输出);
- 追踪来源:trace_source(追踪来源)。
数据规模
- 总大小:约 445.5 MB(数据集整体大小)
- 下载大小:约 365.9 MB
- 训练集样本数:4,725 条样本
- 训练集大小:445,495,820 字节
数据划分
- 划分名称:train(训练集)
- 文件路径:data/train-*(支持通配符读取)
数据格式
- 格式:基于 Hugging Face Datasets 的标准格式,采用 Parquet / Arrow 存储
- 字段结构:包含对话列表(conversations)、代理(agent)、模型(model)、模型提供方(model_provider)、日期(date)、任务(task)、episode(episode)、run_id、trial_name、result、verifier_output、trace_source 等字段
用途建议
该数据集适用于以下场景:
- 多轮对话系统的训练与评测
- 代码生成或单元测试(PHPUnit)相关任务的模型微调
- 代理(Agent)行为分析与结果验证
- 模型执行轨迹(trace)的记录与复现研究




