官方服务:
资源简介:
NADI, is an Arabic dialect corpus
应用场景:
创建时间:
2024-12-01
相关数据集
CHILDES Mandarin BJCMC Corpus
(1) The BJCMC Corpus) The Beijing Child Mandarin Corpus (BJCMC) was constructed to address the absence of systematic documentation of child Mandarin speech in naturalistic contexts at preschool age (3
DataCite Commons2026-05-04 更新130
TSM Corpus - D Version
The Taiwanese Southern Min (TSM) Corpus project has been funded by a research grant from the China Institute, University of Alberta and funding from the Department of Linguistics, University of Albert
DataONE2015-06-25 更新90
QCRI/AraDiCE-TruthfulQA
--- license: cc-by-nc-sa-4.0 pretty_name: 'AraDiCE -- TruthfulQA' dataset_info: - config_name: TruthfulQA-eng splits: - name: test num_examples: 780 - config_name: TruthfulQA-msa
Hugging Face2024-11-04 更新90
Sibirientyska
Korpusen har c:a 34 000 ord. Ryska ord och alla verbformer har annoterats (ryska ord och hybrider står i parentes; böjda verbformer får attribut FINIT eller INFINIT). Sibirientyska ingår i ett samarbe
DataCite Commons2025-12-12 更新70



