Mini-dataset for VL-Models fine-tuning (VL-Tune-dataset-mini)

Hussein Mohammed

doi:10.25592/uhhfdm.12671

June 30, 2023 Dataset Open Access

Mini-dataset for VL-Models fine-tuning (VL-Tune-dataset-mini)

Hussein Mohammed

Citation Style Language JSON Export

{"DOI":"10.25592/uhhfdm.12671","abstract":"<p>A minimal dataset of 125 image-text pairs&nbsp;and 10&nbsp;text queries&nbsp;for fine-tuning vision-language models on manuscript images.&nbsp;It is dedicated to the task of text-based image retrieval, and splited into &quot;train&quot; and &quot;test&quot; sets. The train set&nbsp;consists of 100 image-text pairs, while the test set consists of 25 image-text pairs. This dataset is constructed from the following sources:</p>\n\n<p>- images from&nbsp;the <a href=\"http://spotting.univ-rouen.fr/\">DocExplore</a> dataset&nbsp;of medieval manuscripts.</p>\n\n<p>- images from&nbsp;two manuscripts from Al-\u1e24ar\u012br\u012b,&nbsp;<em>Maq\u0101m\u0101t</em>,&nbsp;&copy; Paris,&nbsp;Biblioth&egrave;que nationale de France. D&eacute;partement des manuscrits,&nbsp;namely MS arabe 3929 and MS arabe 5847.</p>\n\n<p>- the descriptions in the text files are prepared by&nbsp;Martina Dinelli</p>\n\n<p>The research for this work was funded by the Deutsche Forschungsgemeinschaft (DFG, German Research Foundation) under Germany&#39;s Excellence Strategy &ndash; EXC 2176 &lsquo;Understanding Written Artefacts: Material, Interaction and Transmission in Manuscript Cultures&#39;, project no. 390893796. The research was conducted within the scope of the Centre for the Study of Manuscript Cultures (CSMC) at Universit&auml;t Hamburg.</p>","author":[{"family":"Hussein Mohammed"}],"id":"12671","issued":{"date-parts":[[2023,6,30]]},"language":"eng","title":"Mini-dataset for VL-Models fine-tuning (VL-Tune-dataset-mini)","type":"dataset","version":"1.0"}

Publication date:

June 30, 2023

DOI:

Keyword(s):

Vision-Language Models, Dataset

Communities:

License (for files):

Creative Commons Attribution 4.0 International

Versions

Version 1.0 10.25592/uhhfdm.12671

Jun 30, 2023

Cite all versions? You can cite all versions by using the DOI 10.25592/uhhfdm.12670. This DOI represents all versions, and will always resolve to the latest one.

Zentrumfür Nachhaltiges Forschungsdatenmanagement

Suche

Mini-dataset for VL-Models fine-tuning (VL-Tune-dataset-mini)

Citation Style Language JSON Export

Versions

Cite record as

Export

Mini-dataset for VL-Models fine-tuning (VL-Tune-dataset-mini)

Citation Style Language JSON Export

DOI Badge

Markdown

[![DOI](https://www.fdr.uni-hamburg.de/badge/DOI/10.25592/uhhfdm.12671.svg)](https://doi.org/10.25592/uhhfdm.12671)

reStructedText

.. image:: https://www.fdr.uni-hamburg.de/badge/DOI/10.25592/uhhfdm.12671.svg :target: https://doi.org/10.25592/uhhfdm.12671

HTML

<a href="https://doi.org/10.25592/uhhfdm.12671"><img src="https://www.fdr.uni-hamburg.de/badge/DOI/10.25592/uhhfdm.12671.svg" alt="DOI"></a>

Image URL

https://www.fdr.uni-hamburg.de/badge/DOI/10.25592/uhhfdm.12671.svg

Target URL

https://doi.org/10.25592/uhhfdm.12671

Versions

Cite record as

Export