ActionArt
收藏资源简介:
The dataset is introduced in our paper titled "ActionArt: Advancing Multimodal Large Models for Fine-Grained Human-Centric Video Understanding" (https://arxiv.org/abs/2504.18152). It comprises a training set and a test set. The training set includes video names along with manually annotated video captions. The test set is further divided into multiple subsets, each documented in a JSON file that contains the video names and their corresponding manually annotated QA. This dataset utilizes videos from the Movid dataset, and we encourage you to refer to their work to access the original videos. 可以通过如下GIT Clone命令,或者ModelScope SDK来下载数据集 #### 下载方法 :modelscope-code[]{type="sdk"} :modelscope-code[]{type="git"}
ActionArt数据集的相关介绍刊载于论文《ActionArt:推进面向细粒度以人为中心的视频理解的多模态大模型(Multimodal Large Model)》(即将发表)




