遇见数据集

lexi-ml/wmt_22_processed_data

收藏
Hugging Face2025-10-23 更新2025-10-25 收录
官方服务:

资源简介:

该数据集包含文本对及其相关信息,如唯一标识符pair_id,源文本source_text,源语言source_lang,数据来源source_dataset,数据集分割方式split,音频文本audio_text,目标文本target_text和目标语言target_lang。数据集分为训练集train,包含106507个示例和324492210字节大小。

The dataset contains text pairs and related information such as unique identifier pair_id, source text source_text, source language source_lang, dataset source source_dataset, dataset split method split, audio text audio_text, target text target_text, and target language target_lang. The dataset is split into training set train, containing 106507 examples and 324492210 bytes in size.

提供机构:
lexi-ml
二维码
社区交流群
二维码
科研交流群
商业服务