遇见数据集

Vocal Imitation Set v1.1.3 : Thousands of vocal imitations of hundreds of sounds from the AudioSet ontology

收藏
Zenodo2020-09-20 更新2026-05-25 收录
数据链接:
官方服务:

资源简介:

The VocalImitationSet is a collection of crowd-sourced vocal imitations of a large set of diverse sounds collected from Freesound (https://freesound.org/), which were curated based on Google's AudioSet ontology (https://research.google.com/audioset/). We expect that this dataset will help research communities obtain a better understanding of human's vocal imitation and build a machine understand the imitations as humans do. See https://github.com/interactiveaudiolab/VocalImitationSet for more information about this dataset and its latest updates. For citations, please use this reference: Bongjun Kim, Madhav Ghei, Bryan Pardo, and Zhiyao Duan, "Vocal Imitation Set: a dataset of vocally imitated sound events using the AudioSet ontology," <em>Proceedings of the Detection and Classification of Acoustic Scenes and Events 2018 Workshop (DCASE2018)</em>, Nov. 2018. Contact Info: - Interactive Audio Lab: http://music.eecs.northwestern.edu - Bongjun Kim bongjun@u.northwestern.edu | http://www.bongjunkim.com - Bryan Pardo pardo@northwestern.edu | http://www.bryanpardo.com

语音模仿数据集(VocalImitationSet)是一个众包构建的语音模仿样本集合,其样本源自Freesound平台(https://freesound.org/)收录的大量多样化声音素材,并基于谷歌AudioSet本体(https://research.google.com/audioset/)完成精选整理。本数据集旨在助力研究学界更深入地认知人类语音模仿行为,并推动机器实现类人的语音模仿理解能力。 如需了解该数据集的更多详情与最新更新,请访问:https://github.com/interactiveaudiolab/VocalImitationSet。 引用该数据集时,请使用以下参考文献: 金奉俊(Bongjun Kim)、马达夫·盖(Madhav Ghei)、布莱恩·帕多(Bryan Pardo)、段志尧(Zhiyao Duan),《语音模仿数据集:基于AudioSet本体的语音模仿声学事件数据集》,载于《2018年声学场景与事件检测与分类研讨会(DCASE2018)论文集》,2018年11月。 联系信息: - 交互式音频实验室(Interactive Audio Lab):http://music.eecs.northwestern.edu - 金奉俊(Bongjun Kim):bongjun@u.northwestern.edu | http://www.bongjunkim.com - 布莱恩·帕多(Bryan Pardo):pardo@northwestern.edu | http://www.bryanpardo.com

提供机构:
Zenodo
创建时间:
2018-08-07
二维码
社区交流群
二维码
科研交流群
商业服务