mosama/eng_10k_arb_30k_urd_100k_v2_prefixed
收藏官方服务:
资源简介:
该数据集包含文本特征,适用于文本相关的任务。训练集共有140,000个样本,数据集总大小为722,473,567字节。不过,README文件中并没有提供数据集的具体描述,因此无法给出详细的数据集中文描述。
The dataset includes a text feature, which is suitable for text-related tasks. The training set contains 140,000 samples, and the total size of the dataset is 722,473,567 bytes. However, the README file does not provide a specific description of the dataset, so no detailed Chinese description of the dataset can be given.
提供机构:
mosama


