MEMECAP
收藏资源简介:
MEMECAP数据集由不列颠哥伦比亚大学和向量人工智能研究所共同创建,专注于网络用户使用视觉隐喻表达思想的模因。该数据集包含6384个模因,每个模因都附有帖子标题、模因标题、字面图像标题和视觉隐喻。数据集的创建过程涉及从Reddit收集模因,并通过人工筛选确保质量和排除攻击性内容。MEMECAP的应用领域在于提升视觉和语言模型对视觉隐喻的理解能力,特别是在模因生成和解释方面,旨在解决现有模型在处理视觉隐喻时的不足。
The MEMECAP dataset was co-created by the University of British Columbia and the Vector Institute for Artificial Intelligence, focusing on memes through which internet users express their ideas via visual metaphors. This dataset contains 6,384 memes, with each meme paired with a post title, meme title, literal image caption, and visual metaphor annotation. The dataset construction process involves collecting memes from Reddit and conducting manual filtering to ensure quality and exclude offensive content. The application scenarios of MEMECAP lie in enhancing the ability of vision-language models to comprehend visual metaphors, particularly in meme generation and interpretation, aiming to address the shortcomings of existing models when handling visual metaphors.




