terminal_bench_2_a3_rl_laion_exp_rpt_crosscodeeval_csharp_v4_50_8B_20260828_184838
收藏资源简介:
该数据集包含多轮对话记录,每条记录由以下字段构成:conversations(对话列表,每条对话包含角色role和内容content)、agent(使用的智能体)、model(模型名称)、model_provider(模型提供商)、date(日期)、task(任务标识)、episode(轮次)、run_id(运行ID)、trial_name(试验名称)、result(结果)、verifier_output(验证器输出)、trace_source(追踪来源)。数据集仅包含训练集,共2781个样本,总大小约260MB。适用于对话系统评估、多轮对话分析、智能体行为追踪等任务。
This dataset contains multi-turn dialogue records. Each record consists of the following fields: conversations (list of dialogues, each containing role and content), agent (the agent used), model (model name), model_provider (model provider), date, task (task identifier), episode, run_id, trial_name, result, verifier_output, and trace_source. The dataset only includes the training set, with 2781 samples and a total size of approximately 260MB. It is suitable for tasks such as dialogue system evaluation, multi-turn dialogue analysis, and agent behavior tracking.
数据集概述
该数据集名为 laion/terminal_bench_2_a3_rl_laion_exp_rpt_crosscodeeval_csharp_v4_50_8B_20260828_184838,托管于 Hugging Face,属于 LAION 组织。该数据集包含用于训练或评估智能体(agent)的交互对话记录,具体聚焦于 C# 编程任务的强化学习(RL)实验。
数据集结构
- 格式:数据集由多个特征字段组成,每条记录包含一个
conversations列表(其中每个对话项有role和content两个字符串字段),以及一系列元数据字段。 - 主要字段:
conversations:对话历史,包含角色(如系统、用户、助手)和内容。agent:执行任务的智能体名称。model:使用的模型名称。model_provider:模型提供方。date:数据生成日期。task:具体任务标识。episode:执行轮次。run_id:运行编号。trial_name:试验名称。result:任务结果。verifier_output:验证器输出。trace_source:轨迹来源。
数据划分与规模
- 划分:仅包含一个
train划分。 - 样本数量:训练集共有 2,781 个样本。
- 总大小:数据集总大小约 260.8 MB(
dataset_size),下载大小约 191.9 MB。
数据内容
该数据集来源于一个强化学习(RL)实验,使用 8B 参数的模型,针对 CrossCodeEval C# 编程任务进行训练或评估。数据包含完整的对话轨迹、智能体行为、模型输出及结果验证信息,适用于研究代码生成、智能体交互和强化学习场景。




