科技创新多模态大模型图像-文本数据集
收藏资源简介:
本数据集是专为科技创新大模型训练而构建的图片文本数据集,主要是从专利领域的图像与相应的文本描述配对而成,旨在为模型在专利相关领域提供丰富的视觉和语言信息。本数据集主要用于人工智能领域多模态大模型的图文场景训练和验证。作为训练集,可提升大模型对专利领域的图像理解能力;作为测试集,可以对专利领域的检索和识别能力做出评测。该数据在集专利检索、图像识别、自然语言处理、多模态学习、辅助设计等具体场景下有重要提升作用。
This image-text dataset is specifically constructed for training large-scale models targeting technological innovation. It is mainly composed of paired images and their corresponding text descriptions from the patent field, aiming to provide rich visual and linguistic information for models in patent-related domains. This dataset is primarily used for the training and validation of multimodal large models in the field of artificial intelligence for image-text scenarios. When used as a training set, it can enhance the image understanding capability of large models in the patent field; when used as a test set, it can evaluate the retrieval and recognition capabilities in the patent domain. This dataset plays a significant role in improving performance in specific scenarios including patent retrieval, image recognition, natural language processing, multimodal learning, and auxiliary design.




