遇见数据集

N20EM dataset for multimodal lyric transcription

收藏
Zenodo2024-03-15 更新2026-05-25 收录
数据链接:
官方服务:

资源简介:

<strong>Note: To access the dataset, please submit the request at Zenodo page of the newer version of this dataset.</strong> N20EM dataset for multimodal lyric transcription, proposed in our ACM MM 2022 paper, MM-ALT: A Multimodal Automatic Lyric Transcription System. It contains recordings of three modalities: audio, video, and IMU motion signal. Please cite our work as: <pre><code>@inproceedings{gu2022mm, title={MM-ALT: A multimodal automatic lyric transcription system}, author={Gu, Xiangming and Ou, Longshen and Ong, Danielle and Wang, Ye}, booktitle={Proceedings of the 30th ACM International Conference on Multimedia}, pages={3328--3337}, year={2022} }</code></pre> Our paper's camera ready version: https://arxiv.org/abs/2207.06127 Project website: https://n20em.github.io/

提供机构:
Zenodo
创建时间:
2022-10-05
二维码
社区交流群
二维码
科研交流群
商业服务