bionotes-storage
收藏资源简介:
该数据集旨在帮助训练模型遵循人类指令,其核心任务是“指令跟随”。数据由GPT-4生成,并经过人工筛选,以确保质量。数据集规模庞大,包含约120万个示例。每个数据样本通常由两个主要字段构成:“input”(指令或输入)和“output”(期望的模型输出或回复)。数据集适用于训练和评估大语言模型在理解和执行多样化指令方面的能力。
This dataset is designed to assist in training models to follow human instructions, with its core task being "instruction following". The data is generated by GPT-4 and manually filtered to ensure quality. It has a large scale, containing approximately 1.2 million examples. Each data sample typically consists of two main fields: "input" (the instruction or input) and "output" (the expected model output or response). This dataset is suitable for training and evaluating Large Language Models (LLMs) on their ability to understand and execute diverse instructions.
数据集概述
- 数据集名称:bionotes-storage
- 来源:Hugging Face 数据集中心
- 许可证:MIT 许可证(开放使用、复制、修改和分发)
许可信息
该数据集采用 MIT 许可证,允许用户自由使用、复制、修改、合并、出版、分发、再许可和/或出售副本,仅需在软件和软件的所有副本中包含版权声明和许可声明。
注意
- 当前 README 文件中未提供数据集的具体描述、用途、规模、格式或样本内容。
- 如需进一步了解该数据集的详细信息(如任务类型、数据构成、引用方式等),建议直接访问数据集页面或查看其相关文档。




