遇见数据集

BoweiLiu/Lvis_No_OmniVerifier

收藏
Hugging Face2025-12-15 更新2025-12-20 收录
官方服务:

资源简介:

该数据集是一个包含图像和文本的多模态数据集,主要用于视觉问答任务。每个样本包含一张图像(img)、一个相关问题(question)和一个答案(answer)。答案部分包括边界框坐标(bbox)和一个布尔值(bool)。数据集仅包含训练集,共有46,110个样本,总大小约为22.44 GB。

This dataset is a multimodal dataset containing images and text, primarily used for visual question answering tasks. Each sample includes an image (img), a related question (question), and an answer (answer). The answer part consists of bounding box coordinates (bbox) and a boolean value (bool). The dataset only includes a training set with a total of 46,110 samples and a size of approximately 22.44 GB.

提供机构:
BoweiLiu
二维码
社区交流群
二维码
科研交流群
商业服务