遇见数据集

NeutrinoPit/OpenSubtitles2024-en-ar-batch25

收藏
Hugging Face2025-03-04 更新2025-04-12 收录
官方服务:

资源简介:

这是一个包含两个语言字段(英文和阿拉伯文)的数据集,适用于训练自然语言处理模型。数据集仅包含训练集分割,共有100万条样本,总大小约为103MB。

This dataset includes two language fields (English and Arabic) and is suitable for training natural language processing models. The dataset contains only a training split with 1 million samples, totaling approximately 103MB in size.

提供机构:
NeutrinoPit
二维码
社区交流群
二维码
科研交流群
商业服务