遇见数据集

CrowdMind/finewebedu-100K-samples-256t

收藏
Hugging Face2025-09-19 更新2025-10-25 收录
官方服务:

资源简介:

该数据集包含文本内容、ID、语言等多种信息,划分为训练集,可用于语言相关的模型训练等任务。

The dataset includes text content, ID, language and other information, divided into training set, which can be used for language-related model training tasks.

提供机构:
CrowdMind
二维码
社区交流群
二维码
科研交流群
商业服务