遇见数据集

COOCO: Common Objects Out-of-Context

收藏
Zenodo2025-09-19 更新2026-05-26 收录
官方服务:

资源简介:

COOCO is a large-scale dataset designed to investigate how Vision-Language Models (VLMs) leverage scene context during referring expression generation. It focuses on semantic violations, evaluating model performance when target objects have low, medium, or high semantic relatedness to their surrounding scenes. Each sample in COOCO contains: A scene from COCO-Search18 with a labeled target object, called Original version A version of the scene where the target object is just removed, called Clean version Multiple manipulated versions with inpainted objects of varying semantic relatedness Each original image is augmented by replacing the target object with semantically related or unrelated alternatives. Relatedness is computed using ConceptNet embeddings and THINGSplus norms: Low Medium High Same-target: original object category generated (control) Inpainting is guided by LLaVA-generated prompts and verified via an automated visual QA procedure. For complete information, please see our Huggingface Dataset card: https://huggingface.co/datasets/fmerlo/COOCO

提供机构:
Zenodo
创建时间:
2025-09-19
二维码
社区交流群
二维码
科研交流群
商业服务