官方服务:
资源简介:
Procssed LJSpeech dataset
经过处理的LJSpeech数据集
应用场景:
创建时间:
2025-02-08
相关数据集
Localizing Fake Segments in Speech
Partial Synthetic Detection (Psynd) dataset is a multi-speaker English corpus of 2294 utterances, approximately 13 hours English speech at 24kHz sampling rate. It is derived from LibriTTS , a read Eng
DataCite Commons2022-06-24 更新210
Codec-SUPERB/fluent_speech_commands_test_subset_synth
--- configs: - config_name: default data_files: - split: original path: data/original-* - split: academicodec_hifi_16k_320d path: data/academicodec_hifi_16k_320d-* - split: academicode
Hugging Face2024-02-10 更新110
AdoCleanCode/hiffi_train_edit_v1
--- dataset_info: features: - name: speaker_id dtype: string - name: flac_filename dtype: string - name: transcription_full dtype: string - name: removed_words dtype: string
Hugging Face2025-12-05 更新60
nb-librivox
这是一个由挪威国家图书馆制作的高质量挪威语文本到语音(TTS)数据集,包含了从LibriVox公共领域有声书衍生的音频片段和伪对齐的转录文本及标点符号。适用于语音合成和自动语音识别研究。
Hugging Face2025-06-17 更新120
UGLabs/Twilight-Sparkle_tagged
--- dataset_info: features: - name: speaker dtype: string - name: source dtype: string - name: start dtype: float64 - name: end dtype: float64 - name: style dtype: stri
Hugging Face2024-05-04 更新110



