Contains ~22k real audio from the FMA medium dataset and 18k generated audio from the FakeMusicCaps dataset. All audios are preprocessed in .mp3 format. Sampling rate is 16k Hz and bit rate is 215 kb
Audio sensors, essential for automatic speaker verification (ASV) systems, face growing threats from spoofed audio generated by advanced speech synthesis techniques. Traditional spoof detection method
--- license: apache-2.0 --- # SpeechFake Dataset Please use the download scripts from https://github.com/XIAOYixuan/AUDDT/tree/yixuan-dev to download and process the dataset. ``` chmod +x download/