遇见数据集

voices365/102_Hours_High_Quality_Chinese_Audio_Dataset_For_Speech_Synthesis_Female_Samples

收藏
Hugging Face2024-09-06 更新2025-04-26 收录
官方服务:

资源简介:

--- license: cc-by-nc-2.0 task_categories: - text-to-speech - text-to-audio - text-to-video language: - zh size_categories: - 10B<n<100B --- ### Dataset Description 102-Hour Chinese Mandarin Audio Dataset for Speech Synthesis. It's recorded by 102 professional Chinese voice artists (Male and Female). Professional Chinese phoneticians participated in the annotation process. For more details, please refer to the link: www.vodataset.com or email info@voices365.com. ### Audio Format 48,000Hz, 24bit, wav, mono. ### Recording Environment Professional Recording Studio. ### Recording Content 6 different novel books with different themes. (We own the copyright of these books). ### Speakers 102 Professional Chinese voice artists, each recorded one hour. ### Language Chinese Mandarin. ### Annotation Chinese Characters and Pinyin (carefully reviewed by phoneticians). ### Useage ASR and Speech Synthesis. ### Licensing Information Commercial License

许可证:CC BY-NC 2.0(知识共享署名-非商业性使用2.0协议) 任务类别: - 文本转语音 - 文本转音频 - 文本转视频 语言:中文(zh) 样本规模区间:100亿 < n < 1000亿 ### 数据集描述 本数据集为面向语音合成的102小时普通话音频数据集,由102名涵盖男女的专业华语配音演员录制,专业华语语音学家参与了标注流程。如需获取更多详细信息,请访问链接www.vodataset.com或发送邮件至info@voices365.com。 ### 音频格式 48000Hz、24位、单声道WAV格式。 ### 录制环境 专业录音棚。 ### 录制内容 6部不同主题的小说,我方拥有该些书籍的版权。 ### 配音员 102名专业华语配音演员,每人录制时长为1小时。 ### 语言 普通话。 ### 标注信息 标注内容为汉字与拼音,且经语音学家严格审核。 ### 应用场景 自动语音识别(ASR)与语音合成。 ### 授权信息 商业授权。

提供机构:
voices365
二维码
社区交流群
二维码
科研交流群
商业服务