FineCops-Ref: A new Dataset and Task for Fine-Grained Compositional Referring Expression Comprehension
收藏数据链接:
官方服务:
资源简介:
FineCops-Ref is a dataset for Compositional Referring Expression Comprehension (REC) that rigorously evaluates Vision-Language Models (VLMs) on compositional reasoning and their ability to identify inconsistencies between images and text. Beyond standard REC tasks, it challenges models with fine-grained correspondences involving objects, attributes, and relationships. The dataset comprises both training and testing sets, designed to thoroughly assess model performance across various difficulty level. The Paper can be found in [2409.14750] FineCops-Ref: A new Dataset and Task for Fine-Grained Compositional Referring Expression Comprehension (arxiv.org)
提供机构:
figshare创建时间:
2025-05-28



