遇见数据集

MTC-VC: A Multi-Task Contrastive Learning Method for Efficient and Controllable Voice Cloning

收藏
IEEE2026-04-17 收录
官方服务:

资源简介:

The LibriSpeech corpus, a publicly available English speech dataset derived from audiobook recordings. The corpus contains approximately 1,000 hours of 16 kHz read speech from over 2,400 speakers, encompassing diverse speaking styles, rates, and regional accents. For the purpose of contrastive learning, a subset of 100 speakers was sampled, with 20 utterances per speaker ranging from 3 to 10 seconds. The dataset provides clean, labeled speech suitable for tasks involving speaker representation, acoustic modeling, and multi-style synthesis.

提供机构:
Zhou, Rui
二维码
社区交流群
二维码
科研交流群
商业服务