llm-jp/llm-jp-4-thinking-sft-data
收藏资源简介:
该数据集是一个用于训练`llm-jp-4-*-thinking`模型的监督微调(SFT)数据集。它通过从多个数据源提取提示并使用`gpt-oss-120b`生成推理过程和最终响应来构建。数据集包含`reasoning_low`、`reasoning_medium`和`reasoning_high`等分割,对应于生成过程中不同的推理努力设置。数据集来源列有各自的许可证,部分子集因重新分发限制而未包含。该数据集支持LLM-jp的持续发展,并鼓励用户通过调查表分享他们的使用成果。
This dataset is a supervised fine-tuning (SFT) dataset used to train `llm-jp-4-*-thinking` models. It is constructed by extracting prompts from multiple data sources and generating reasoning processes and final responses using `gpt-oss-120b`. The dataset includes splits such as `reasoning_low`, `reasoning_medium`, and `reasoning_high`, corresponding to different reasoning effort settings during generation. The data sources are listed with their respective licenses, and some subsets are not included due to redistribution restrictions. The dataset supports the continued development of LLM-jp and encourages users to share how they utilize the outcomes via a survey form.




