Viorra-Reasoning-Baseline
收藏资源简介:
Viorra Reasoning Baseline是一个用于文本生成任务的数据集,专门设计用于Viorra项目的第一阶段微调。该数据集的核心内容是从Claude Sonnet和Claude Opus模型生成的输出中,经过筛选后得到的、以人文学科为主题的推理轨迹(reasoning traces)。其主要目的是为Gemma 4 E2B模型提供训练数据,以增强其在第一阶段的人格稳定性(Persona Stability)。数据集规模为13,117条样本,以JSONL聊天格式组织。数据语言为英语,采用Apache-2.0开源许可证。
Viorra Reasoning Baseline is a dataset for text generation tasks, specifically designed for the first-stage fine-tuning of the Viorra project. Its core content consists of reasoning traces on humanities topics, which are filtered from outputs generated by Claude Sonnet and Claude Opus models. The primary goal is to provide training data for the Gemma 4 E2B model to enhance its Persona Stability in the first stage. The dataset contains 13,117 samples, organized in JSONL chat format. The data language is English, and it is released under the Apache-2.0 open-source license.
数据集概述:Viorra Reasoning Baseline
该数据集主要用于 Viorra 项目第一阶段微调,专注于人文领域的推理轨迹。
- 语言: 英语 (
en) - 许可协议: Apache-2.0
- 规模: 包含 13,117 行数据 (10K < n < 100K)
- 数据格式: JSONL Chat 格式
- 任务类别: 文本生成 (
text-generation) - 标签:
gemma,viorra,reasoning - 数据来源: 经过筛选的 Claude Sonnet/Opus 生成的人文领域推理轨迹 (humanities-focused reasoning traces)
- 主要用途: 用于 Gemma 4 E2B 模型的第一阶段人格稳定性微调 (Stage 1 Persona Stability)





