gsd-smith-Kikuyu
收藏资源简介:
该数据集是一个包含多轮对话交互和代理行为轨迹的数据集。每个数据样本包含以下字段:唯一标识符(id)、初始种子提示(seed_prompt)、语言类型(language)、使用的模型(model)、消息序列(messages,其中每条消息包含角色和内容)、代理执行轨迹(agent_trace,以JSON列表格式记录代理的行为过程)、来源标识(source_id)以及研究早期停止标志(research_early_stopping)。数据集规模为357个训练样本,总数据量约10.5MB。该数据集适用于对话系统研究、多轮对话生成、代理行为分析、人机交互建模等任务,能够为基于大语言模型的对话代理开发和评估提供支持。
This dataset is a collection of multi-turn dialogue interactions and agent behavior trajectories. Each data sample includes the following fields: unique identifier (id), initial seed prompt (seed_prompt), language type (language), model used (model), message sequence (messages, where each message contains a role and content), agent execution trace (agent_trace, recorded in JSON list format detailing the agents behavior process), source identifier (source_id), and research early stopping flag (research_early_stopping). The dataset consists of 357 training samples with a total size of approximately 10.5 MB. It is suitable for tasks such as dialogue system research, multi-turn dialogue generation, agent behavior analysis, and human-computer interaction modeling, providing support for the development and evaluation of large language model-based dialogue agents.





