相关数据集
TSM Corpus - D Version
The Taiwanese Southern Min (TSM) Corpus project has been funded by a research grant from the China Institute, University of Alberta and funding from the Department of Linguistics, University of Albert
DataONE2015-06-25 更新90
CorpusCanarioWA
This dataset is a subcorpus of WhatsApp voice-message recordings and transcriptions in Canarian Spanish (Tenerife), extracted and annotated for studies in dialectology, sociolinguistics, and spoken la
Zenodo2026-05-19 更新20
CiteCalc 1.2
The main purpose of the program CiteCalc is to find corresponding passages in text editions. The program calculates the volume and page near which the passage can be found in another edition.
DataCite Commons2026-04-01 更新00
Sibirientyska
Korpusen har c:a 34 000 ord. Ryska ord och alla verbformer har annoterats (ryska ord och hybrider står i parentes; böjda verbformer får attribut FINIT eller INFINIT). Sibirientyska ingår i ett samarbe
DataCite Commons2025-12-12 更新70
Sherdukpen Lexicon - Rupa
This dataset contains all the sound files and transcribes files of the Rupa variety of Sherdukpen. Bodt, Timotheus Adrianus. 2024. Proto-Western Kho-Bwa: Reconstructing a communities' past through l
Zenodo2024-08-29 更新10



