mlfoundations-dev/pos_neg_ablation_instruction_filtering_seed_code_best_codegolf
收藏数据链接:
官方服务:
资源简介:
这是一个包含多个文本字段的数据集,主要用于训练模型理解指令、来源、推理过程等。数据集由训练集组成,共有10000个示例,每个示例包含如指令种子、来源、推理过程、解决方案、原始行索引、推理轨迹和对话等字段。数据集的总大小为567874823字节,下载大小为240248256字节。
This dataset consists of multiple text fields designed for training models to understand instructions, sources, reasoning processes, etc. The dataset is composed of a training set with 10,000 examples, each containing fields such as instruction seed, source, reasoning, solution, original row index, final reasoning trace, and conversations. The total size of the dataset is 567874823 bytes, with a download size of 240248256 bytes.
提供机构:
mlfoundations-dev


