遇见数据集

BanglaAccent: Annotated Bangla Dialects with Demographic Metadata

收藏
NIAID Data Ecosystem2026-05-10 收录
官方服务:

资源简介:

Desher Vasha is a curated speech dataset capturing dialectal variation across more than 20 distinct regions of Bangladesh. Designed to support linguistic research, educational tools, and low-resource NLP development, the corpus includes voice recordings annotated with rich metadata such as speaker location, age, gender, and dialectal features. Use Cases • Linguistic research: Dialect classification, phoneme variation, sociolinguistic mapping • Educational tools: Role-based apps for dialect awareness and pronunciation training • NLP & ASR: Benchmarking for Bangla speech recognition, especially in low-resource and dialect-sensitive contexts • Cultural preservation: Documenting endangered or underrepresented dialects

创建时间:
2025-11-13
二维码
社区交流群
二维码
科研交流群
商业服务