items_prompts_lite
收藏资源简介:
该数据集包含20,000个训练样本、1,000个验证样本和1,000个测试样本,每个样本由'prompt'(输入提示)和'completion'(完成文本)两个字符串字段组成。数据集总大小为8,914,097字节,下载大小为4,415,566字节。数据分为训练集、验证集和测试集三个部分,分别存储在指定的文件路径中。该结构适合用于文本生成、对话系统或其他需要输入-输出对的自然语言处理任务。
This dataset consists of 20,000 training samples, 1,000 validation samples, and 1,000 test samples. Each sample contains two string fields: 'prompt' (input prompt) and 'completion' (completion text). The total size of the dataset is 8,914,097 bytes, and its download size is 4,415,566 bytes. The data is split into three subsets: training, validation, and test sets, which are stored in their respective designated file paths. This structure is suitable for natural language processing tasks such as text generation, dialogue systems, or other tasks requiring input-output pairs.




