遇见数据集

NeutrinoPit/OpenSubtitles2024-en-ar-batch42

收藏
Hugging Face2025-03-04 更新2025-04-12 收录
官方服务:

资源简介:

该数据集是一个包含英文和阿拉伯文字符串的数据集,用于训练模型。它包含一个训练集,共有100万条示例,数据集总大小为104,563,267字节。

This dataset contains strings in both English (en) and Arabic (ar) for training models. It includes a training set with a total of 1,000,000 examples, with a dataset size of 104,563,267 bytes.

提供机构:
NeutrinoPit
二维码
社区交流群
二维码
科研交流群
商业服务