遇见数据集

tsystems/sharegpt4v_vqa_200k_batch2

收藏
Hugging Face2025-01-26 更新2025-04-12 收录
官方服务:

资源简介:

这是一个基于图像到文本任务的英文数据集,包含图像和对应的查询字符串。数据集分为训练集,大小约为197.95GB,共有20万个样本。数据集遵循CC BY NC 4.0许可,仅限于非商业用途和研究目的。数据集是基于ShareGPT4V团队的工作构建的。

This is an English dataset for image-to-text tasks, containing images and corresponding query strings. The dataset is split into a training set, which is approximately 197.95GB in size with 200,000 samples. The dataset is licensed under CC BY NC 4.0, allowing only for non-commercial use and research purposes. The dataset is built based on the work of the ShareGPT4V team.

提供机构:
tsystems
二维码
社区交流群
二维码
科研交流群
商业服务