DCASE 2024 Challenge Task 7 Development Dataset : Environmental Sound Scene Synthesis
收藏NIAID Data Ecosystem2026-05-01 收录
下载链接:
https://zenodo.org/record/10869643
下载链接
链接失效反馈官方服务:
资源简介:
Description
This dataset comprises embeddings and captions utilized as the development dataset for DCASE 2024 Challenge Task 7, focusing on 'Environmental Sound Scene Synthesis.' The embeddings are derived from 60 different 4-second audio files formatted as mono 32-bit 32kHz, and are contained in the 'embeddings.tar.xz' file. Captions corresponding to each audio file can be found in 'caption.csv'. This dataset does not comprise the audio files, only the embeddings. Three different types of embeddings are provided: VGGish (vggish), MS-CLAP (clap-2023), and PANNs CNN14 Wavegram-Logmel (panns-wavegram-logmel). Only PANNs CNN14 Wavegram-Logmel (panns-wavegram-logmel) embeddings are used for evaluation in the challenge. For further details, please refer to the challenge website.
Contact
Modan Tailleur, modan.tailleur@ls2n.fr
Mathieu Lagrange, mathieu.lagrange@ls2n.fr
创建时间:
2024-03-28



