遇见数据集

Synthesized audio of 300 core words of 42 Indo-European languages

收藏
Zenodo2024-06-10 更新2026-04-07 收录
官方服务:

资源简介:

The speech sounds of 300 core words in this repository are synthesized using the text-to-speech engine in Microsoft Azure AI Speech Studio, which encompasses 42 Indo-European languages. Each word is synthesized in both male and female voices, resulting in a total of 25,200 audio clips (300 words × 42 languages × 2 genders). All audio clips are in 16-bit, 16 kHz, mono WAV format, with leading and trailing silences trimmed.

提供机构:
Nankai University
创建时间:
2024-06-10
二维码
社区交流群
二维码
科研交流群
商业服务