遇见数据集

kamruzzaman-asif/image_captions_x

收藏
Hugging Face2025-05-09 更新2025-08-30 收录
官方服务:

资源简介:

这是一个包含图像-文本对的数据集,由LAION-400M、COYO-700M和Conceptual Captions三个子集合并而成。每个子集包含数百万到数千万的图像和对应的文本描述。数据集适用于训练和评估视觉语言模型,以及图像-文本检索任务。

This dataset consists of image-caption pairs merged from LAION-400M, COYO-700M, and Conceptual Captions subsets. Each subset contains millions to tens of millions of images with corresponding textual descriptions. The dataset is suitable for training and evaluating vision-language models, as well as image-text retrieval tasks.

提供机构:
kamruzzaman-asif
二维码
社区交流群
二维码
科研交流群
商业服务