Dataset Open Access
{"conceptdoi":"10.25592/uhhfdm.12670","conceptrecid":"12670","created":"2023-06-30T12:07:03.457297+00:00","doi":"10.25592/uhhfdm.12671","id":12671,"links":{"badge":"https://www.fdr.uni-hamburg.de/badge/doi/10.25592/uhhfdm.12671.svg","conceptbadge":"https://www.fdr.uni-hamburg.de/badge/doi/10.25592/uhhfdm.12670.svg","conceptdoi":"http://doi.org/10.25592/uhhfdm.12670","doi":"http://doi.org/10.25592/uhhfdm.12671"},"metadata":{"access_right":"open","access_right_category":"success","communities":[{"id":"csmc"},{"id":"uhh"}],"creators":[{"affiliation":"Universit\u00e4t Hamburg","name":"Hussein Mohammed","orcid":"0000-0001-5020-3592"}],"description":"<p>A minimal dataset of 125 image-text pairs and 10 text queries for fine-tuning vision-language models on manuscript images. It is dedicated to the task of text-based image retrieval, and splited into "train" and "test" sets. The train set consists of 100 image-text pairs, while the test set consists of 25 image-text pairs. This dataset is constructed from the following sources:</p>\n\n<p>- images from the <a href=\"http://spotting.univ-rouen.fr/\">DocExplore</a> dataset of medieval manuscripts.</p>\n\n<p>- images from two manuscripts from Al-\u1e24ar\u012br\u012b, <em>Maq\u0101m\u0101t</em>, © Paris, Bibliothèque nationale de France. Département des manuscrits, namely MS arabe 3929 and MS arabe 5847.</p>\n\n<p>- the descriptions in the text files are prepared by Martina Dinelli</p>\n\n<p>The research for this work was funded by the Deutsche Forschungsgemeinschaft (DFG, German Research Foundation) under Germany's Excellence Strategy – EXC 2176 ‘Understanding Written Artefacts: Material, Interaction and Transmission in Manuscript Cultures', project no. 390893796. The research was conducted within the scope of the Centre for the Study of Manuscript Cultures (CSMC) at Universität Hamburg.</p>","doi":"10.25592/uhhfdm.12671","keywords":["Vision-Language Models, Dataset"],"language":"eng","license":{"id":"CC-BY-4.0"},"publication_date":"2023-06-30","related_identifiers":[{"identifier":"10.25592/uhhfdm.12670","relation":"isVersionOf","scheme":"doi"}],"relations":{"version":[{"count":1,"index":0,"is_last":true,"last_child":{"pid_type":"recid","pid_value":"12671"},"parent":{"pid_type":"recid","pid_value":"12670"}}]},"resource_type":{"title":"Dataset","type":"dataset"},"title":"Mini-dataset for VL-Models fine-tuning (VL-Tune-dataset-mini)","version":"1.0"},"owners":[96],"revision":3,"updated":"2023-06-30T20:14:20.401443+00:00"}