DCASE Task 6 - CASTELLA Extracted Features: M2D-CLAP, LAION-CLAP, BEATs, T5
收藏资源简介:
Extracted features for the CASTELLA dataset: for DCASE Task 6: Audio Moment Retrieval from Long Audio (2026) Summary: Models (Audio features): M2D-CLAP, LAION-CLAP, BEATs Models (Text features): M2D-CLAP, LAION-CLAP, T5 Background: The shared features are primarily useful for DCASE Task 6: Audio Moment Retrieval from Long Audio (2026), and all the features of the audio and text files necessary to train the model for this task are present in this repository. Features npz files (respective model) Audio 1,862 Text 3,881 File structure: BEATs |___BEATs_audio |___ __rPhqHS1a0.npz |___ : LAION |___LAION_audio |___ __rPhqHS1a0.npz |___ : |___LAION_text |___ qid__rPhqHS1a0_1.npz |___ : M2D |___M2D_audio |___ __rPhqHS1a0.npz |___ : |___M2D_text |___ qid__rPhqHS1a0_1.npz |___ : T5 |___T5_text |___ qid__rPhqHS1a0_1.npz |___ : Related Code To train or evaluate our model's performance, please use these features along with the instructions and code shared on my GitHub repository. GitHub Repository: AMR-encoder-exploration



