遇见数据集

bobby-nakamoto/zangiyev_az

收藏
Hugging Face2024-07-12 更新2024-07-13 收录
官方服务:

资源简介:

该数据集包含两个主要字段:id和text,其中id是数据的唯一标识符,text是文本内容。数据集分为训练集,包含41,379,968个示例,总大小为7,641,793,441字节。数据集的下载大小为4,294,063,013字节。配置信息指定了默认配置下的数据文件路径。

The dataset contains two main fields: id and text, where id is the unique identifier for the data and text is the textual content. The dataset is divided into a training set, which includes 41,379,968 examples with a total size of 7,641,793,441 bytes. The download size of the dataset is 4,294,063,013 bytes. The configuration information specifies the data file path under the default configuration.

提供机构:
bobby-nakamoto
二维码
社区交流群
二维码
科研交流群
商业服务