HiACC: Hinglish Adult & Children Code-switched corpus
收藏官方服务:
资源简介:
The HiACC corpus is a novel Hinglish code-switched speech dataset featuring both adult and child speakers. It captures naturalistic code-switching through spontaneous responses to everyday questions, story reading, and image-based prompts. The dataset comprises 5.24 hours of segmented audio, including 3,318 utterances from adults and 1,858 from children, all of which have been manually transcribed and annotated for code-switching.
提供机构:
Zenodo创建时间:
2025-05-27



