遇见数据集

Neyshekar: A Large-Scale Open Persian Speech Dataset

收藏
Zenodo2025-12-29 更新2026-05-26 收录
官方服务:

资源简介:

Neyshekar is an open, community-driven Persian speech dataset collected via a web-based crowdsourcing platform at https://ney.shekar.io. It is designed to support research and development in text-to-speech (TTS), automatic speech recognition (ASR), speech representation learning, and other downstream Persian speech applications. The recordings are provided by a combination of volunteer contributors and paid voice actors, all of whom are native Persian speakers. Each release represents a stable snapshot of the dataset, enabling reproducible research and consistent benchmarking.

提供机构:
Zenodo
创建时间:
2025-12-29
二维码
社区交流群
二维码
科研交流群
商业服务