Ultraspeechdataset2017
收藏资源简介:
Ultrasound/Audio databases related to (Fabre et al., Speech Communication, 2017) —————————————————————— HOW-to-CITE: Fabre, D., Hueber, T., Girin, L., Alameda-Pineda, X., Badin, P., (2017) "Automatic animation of an articulatory tongue model from ultrasound images of the vocal tract", Speech Communication, vol. 93, pp. 63-75 These databases are distributed in the hope that they will be useful, but WITHOUT ANY WARRANTY; without even the implied warranty of MERCHANTABILITY or FITNESS FOR A PARTICULAR PURPOSE. See the GNU General Public License for more details. CONTENT: - (unzip first) Ultrasound/Audio data (synchronized) for three native French speakers (2 male speakers (PB and TH), one female speaker (DF)), recorded using Ultraspeech software (www.ultraspeech.com). Data acquisition protocol is described in (Fabre et al., Speech Communication, 2017). More information on the Ultraspeech acquisition system can be found at www.ultraspeech.com and in (Hueber et al., ISSP 2008). - audio_capture/, ultrasound_capture/ directories contain audio/ultrasound data (one folder per sentence). Ultrasound data are 640x480 jpeg images (with 100% quality during compression) recorded at 60 fps. Audio files are in WAV format, 44.1 kHz, 32 bits. Text files in audio directory contain the text prompt for each recorded item. *** WARNING *** : Item 418 of speaker PB is missing. Thomas Hueber, Ph. D., CNRS researcher, GIPSA-lab (Grenoble, France), thomas.hueber@gipsa-lab.fr



