官方服务:
资源简介:
Source: 1 Corinthians (Thai)
来源:《哥林多前书》(泰语)
应用场景:
相关数据集
NLU-Metaphor
# SEA Metaphor SEA Metaphor evaluates a model's ability to interpret paired figurative phrases with divergent meanings. It is sampled from [Multilingual-Fig-QA](https://aclanthology.org/2023.finding
魔搭社区2026-04-28 更新270
fleurs_yo_en
这是一个从Google FLEURS数据集中提取的Yoruba-to-English翻译数据集。数据集包含训练集、验证集和测试集中的音频记录,分别对应13小时48分钟32秒、44分钟32秒和45分钟27秒的音频数据。所有音频都以16kHz采样。数据集的特征包括音频、Yoruba语言的转录和相应的英语翻译。
Hugging Face2024-12-10 更新90
africa-intelligence/aya101-benchmarking
--- pretty_name: Evaluation run of CohereForAI/aya-101 dataset_summary: "Dataset automatically created during the evaluation run of model\ \ [CohereForAI/aya-101](https://huggingface.co/CohereForAI/
Hugging Face2024-10-01 更新70
britllm/TransWebEdu
TransWebEdu是一个预训练规模的多语言平行语料库,支持十种语言:阿拉伯语、威尔士语、德语、英语、西班牙语、法语、印度尼西亚语、意大利语、俄语和斯瓦希里语。它专门用于从零开始预训练TransWebLLM模型,聚焦于多语言网络教育内容。
Hugging Face2025-04-22 更新210
multi_para_crawl
MultiParaCrawl主要用于机器翻译任务,支持包括保加利亚语、加泰罗尼亚语、捷克语等多种语言。数据集规模在10万到100万条样本之间,采用CC0 1.0通用许可。该数据集允许用户指定语言代码对进行加载,并提供标准化数据操作。
Opencsg2024-07-19 更新80



