数据链接:
官方服务:
资源简介:
Coco train captions produced by FuseCap.
应用场景:
创建时间:
2023-07-10
相关数据集
5CD-AI/Vietnamese-BAAI-SVIT-llava-v1.5-format-gg-translated
该数据集专注于视觉问答和问答任务,支持英语和越南语,数据规模在10万到100万之间,主要文件为SVIT_core_150K_vi.parquet。
Hugging Face2024-06-14 更新120
"MVVLP"
"Autonomous Valet Parking (AVP) represents a critical application of autonomous driving; however, existing approaches remain constrained by limited scene understanding, insufficient instruction compre
DataCite Commons2025-10-03 更新70
lego_minifigure_captions
LEGO Minifigure Captions数据集包含12966张LEGO迷你人偶的图像及其描述。数据集包含以下列:'fig_num'表示迷你人偶的编号,'image'为JPEG格式的图像,'short_caption'为图像中迷你人偶的简短描述。数据来源于Rebrickable网站,图像从原始'minifigs.csv'文件的'img_url'列下载。未来计划添加使用Gemini-1.5-f
Hugging Face2024-11-29 更新60
VerboVision/VerboVision-Detail-Tags
--- dataset_info: features: - name: image dtype: image: decode: false - name: instruction dtype: string - name: description dtype: string splits: - name: train
Hugging Face2025-12-08 更新80
Form and sequence: comic books and visual literacy in narrative tenses
Visuals are often used in presenting language concepts with the pedagogical rationale that visual, non-linguistic support can serve to make such concepts comprehensible across linguistic boundaries. C
Figshare2026-02-07 更新60



