遇见数据集

prli/wiki-full_draft_chopped-rare-qwen-alpha0-60

收藏
Hugging Face2026-04-27 更新2026-05-03 收录
官方服务:

资源简介:

该数据集包含文本数据,每个样本由文本内容和一个元数据结构组成,元数据中包括pile_set_name字段,可能表示数据来源的子集名称。数据集仅提供验证集,共4000个样本,总大小约为9.29MB,适用于自然语言处理任务,如文本分析或模型评估。

This dataset contains text data, where each sample consists of text content and a metadata structure, including a pile_set_name field that may indicate the subset source of the data. The dataset only provides a validation split with 4000 samples, totaling approximately 9.29MB, and is suitable for natural language processing tasks such as text analysis or model evaluation.

提供机构:
prli
二维码
社区交流群
二维码
科研交流群
商业服务