SerFabio89/italian-logic-repair-sft-dataset
收藏资源简介:
意大利逻辑修复SFT数据集是一个用于监督微调修复的教师支持的合成意大利语优先数据集。该数据集专门针对精确算术运算、简洁直接问答、可执行Python函数、仅JSON输出、约束遵循、停止行为以及推理最终答案标记行为进行优化。数据集使用教师输出作为候选答案,然后通过确定性检查进行验证、纠正或拒绝。数据集包含149,506行数据,全部属于逻辑修复家族。经过严格的质量控制,包括代码执行测试、JSON解析测试、算术验证等,确保数据可靠性。该数据集主要用于修复简洁意大利语行为、精确答案、结构化输出和小型Python实用程序实现,而非广泛的聊天、安全、多轮或长上下文应用。
Italian Logic Repair SFT Dataset is a teacher-backed synthetic Italian-first dataset designed for supervised fine-tuning repair. It targets exact arithmetic, concise direct QA, executable Python functions, JSON-only output, constraint following, stop behavior, and reasoning final-answer-marker behavior. Teacher outputs are used as candidates, then validated, corrected, or rejected by deterministic checks. The dataset contains 149,506 rows, all belonging to the logic_repair family. It undergoes rigorous quality controls including execution testing for code examples, parse testing for JSON examples, and deterministic validation for arithmetic and critical evaluator-style examples. The dataset is intended for SFT repair of concise Italian behavior, exact answers, structured output, and small Python utility implementation, not for broad chat, safety, multi-turn, or long-context applications.



