Lazysoldier1838/laparoscopy-distortions-v3-fixed
收藏资源简介:
该数据集是一个包含图像和文本对的多模态数据集,主要用于训练任务。数据集包含7600个训练示例,总大小为195,732,323字节(约186 MB),下载大小为444,325,711字节(约424 MB)。特征包括image(图像类型)和text(字符串类型),支持图像与文本的关联分析。数据仅提供train分割,文件路径为data/train-*,适用于计算机视觉和自然语言处理的交叉领域应用,如图像标注或文本生成任务。
This dataset is a multimodal dataset containing image and text pairs, primarily designed for training tasks. It includes 7,600 training examples, with a total size of 195,732,323 bytes (approximately 186 MB) and a download size of 444,325,711 bytes (approximately 424 MB). The features consist of image (image type) and text (string type), supporting the analysis of associations between images and text. The data is provided only in a train split, with file paths as data/train-*, and is suitable for cross-domain applications in computer vision and natural language processing, such as image captioning or text generation tasks.




