autophagycode_D_mercury_Qwen3-4B_lr0.0001_c142_trust_t0.2_g4
收藏资源简介:
该数据集包含142个训练样本,每个样本包含6个结构化字段:任务ID(task_id)、入口点(entry_point)、提示文本(prompt)、补全内容(completion)、top_k进度(top_k_progression)和测试内容(test)。数据集总大小为5.8MB,下载文件体积为572KB。从字段命名推测可能用于代码生成或文本补全类任务,但README未提供明确的任务定义或数据来源说明。
This dataset contains 142 training samples, each with 6 structured fields: task_id, entry_point, prompt, completion, top_k_progression, and test. The total size of the dataset is 5.8 MB, and the compressed download file size is 572 KB. Based on the field naming, it can be inferred that the dataset may be used for code generation or text completion tasks, but no explicit task definition or data source description is provided in the README.
根据您提供的数据集详情页面信息,以下是对该数据集的关键内容总结:
数据集概述
数据集名称:autophagycode_D_mercury_Qwen3-4B_lr0.0001_c142_trust_t0.2_g4
来源地址:https://huggingface.co/datasets/stefanocarrera/autophagycode_D_mercury_Qwen3-4B_lr0.0001_c142_trust_t0.2_g4
数据特征
该数据集包含以下6个字段:
- task_id:字符串类型,表示任务ID
- entry_point:字符串类型,表示入口点
- prompt:字符串类型,表示提示内容
- completion:字符串类型,表示完成内容
- top_k_progression:字符串类型,表示Top-K进度
- test:字符串类型,表示测试数据
数据集划分
数据集仅包含一个划分:
- 训练集(train):包含142个样本,占用约5.82 MB(5,823,887字节)
数据集大小
- 下载大小:约572,789字节(约0.55 MB)
- 总数据集大小:约5,823,887字节(约5.82 MB)
配置文件
- 默认配置(default):包含训练集数据文件,路径为
data/train-*




