Mini-dataset for VL-Models fine-tuning (VL-Tune-dataset-mini)

Name: Mini-dataset for VL-Models fine-tuning (VL-Tune-dataset-mini)
Creator: Universität Hamburg
Published: 2023-06-30 20:14:23
License: 暂无描述

DataCite Commons2023-06-30 更新2025-04-16 收录

下载链接：

https://www.fdr.uni-hamburg.de/record/12670

下载链接

链接失效反馈

官方服务：

资源简介：

A minimal dataset of 125 image-text pairs and 10 text queries for fine-tuning vision-language models on manuscript images. It is dedicated to the task of text-based image retrieval, and splited into "train" and "test" sets. The train set consists of 100 image-text pairs, while the test set consists of 25 image-text pairs. This dataset is constructed from the following sources: - images from the DocExplore dataset of medieval manuscripts. - images from two manuscripts from Al-Ḥarīrī, <em>Maqāmāt</em>, © Paris, Bibliothèque nationale de France. Département des manuscrits, namely MS arabe 3929 and MS arabe 5847. - the descriptions in the text files are prepared by Martina Dinelli The research for this work was funded by the Deutsche Forschungsgemeinschaft (DFG, German Research Foundation) under Germany's Excellence Strategy – EXC 2176 ‘Understanding Written Artefacts: Material, Interaction and Transmission in Manuscript Cultures', project no. 390893796. The research was conducted within the scope of the Centre for the Study of Manuscript Cultures (CSMC) at Universität Hamburg.

提供机构：

Universität Hamburg

创建时间：

2023-06-30

5,000+

优质数据集

54 个

任务类型

进入经典数据集