遇见数据集

tensorkelechi/arxiv-cstext

收藏
Hugging Face2024-06-24 更新2024-06-29 收录
官方服务:

资源简介:

该数据集主要用于文本生成任务,包含大量的文本数据。数据集的特征包括一个名为text的字符串类型字段。数据集分为一个训练集,包含4897900个样本,总大小为273792457字节。数据集的下载大小为157503259字节。

This dataset is primarily used for text generation tasks and contains a large amount of text data. The features of the dataset include a field named text of string type. The dataset is divided into a training set containing 4,897,900 samples with a total size of 273,792,457 bytes. The download size of the dataset is 157,503,259 bytes.

提供机构:
tensorkelechi
二维码
社区交流群
二维码
科研交流群
商业服务