遇见数据集

Vipplav/pocphase3

收藏
Hugging Face2025-01-06 更新2025-02-15 收录
官方服务:

资源简介:

该数据集包含了输入和目标字符串,以及对应的序列编码和注意力掩码。数据集被拆分为训练集,共有4797548个示例,总大小约为1.5GB。数据集适用于序列到序列的任务,如文本生成或机器翻译。

The dataset consists of input and target strings, along with their corresponding sequence encodings and attention masks. The dataset is split into a training set with a total of 4,797,548 examples, approximately 1.5GB in size. It is suitable for sequence-to-sequence tasks such as text generation or machine translation.

提供机构:
Vipplav
二维码
社区交流群
二维码
科研交流群
商业服务