遇见数据集

Text to audio grounding (TAG) dataset: AudioGrounding

收藏
Zenodo2023-10-23 更新2026-05-26 收录
官方服务:

资源简介:

AudioGrounding dataset, including audio files and timestamp annotations. Changes in version 2: The train/validation/test sets are re-split. The validation and test annotations are refined. ---------------------------------------------------------- References [1] Xuenan Xu, Heinrich Dinkel, Mengyue Wu and Kai Yu. "Text-to-audio grounding: Building correspondence between captions and sound events." In Proceedings of IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP). IEEE, 2021, pp. 606-610. [2] Xuenan Xu, Mengyue Wu, and Kai Yu. "Investigating Pooling Strategies and Loss Functions for Weakly-Supervised Text-to-Audio Grounding via Contrastive Learning." In Proceedings of IEEE International Conference on Acoustics, Speech, and Signal Processing Workshops (ICASSPW). IEEE, 2023, pp. 1-5.

提供机构:
Zenodo
创建时间:
2023-10-23
二维码
社区交流群
二维码
科研交流群
商业服务