遇见数据集

bismarck91/enA-frA-xc-tokenized-combined-speaker-transfer-en

收藏
Hugging Face2025-10-09 更新2025-10-25 收录
官方服务:

资源简介:

该数据集包含输入ID序列、注意力掩码序列和标签序列,适用于训练机器学习模型。数据集仅包含一个训练集split,共有5098543个示例,数据大小为33181879593字节。

The dataset includes input ID sequences, attention mask sequences, and label sequences, suitable for training machine learning models. The dataset contains only one training set split with a total of 5,098,543 examples, and the data size is 33,181,879,593 bytes.

提供机构:
bismarck91
二维码
社区交流群
二维码
科研交流群
商业服务