86
stars
7
commits
1
linked in READMEs
Apr 18, 2024
updated
Accelerating the development of large-scale multi-modality models (LMMs) with
lmms-eval
π Homepage | π Documentation | π€ Huggingface Datasets
This is a formatted version of DocVQA. It is used in our lmms-eval pipeline to allow for one-click evaluations of large multi-modality models.
@article{mathew2020docvqa,
title={DocVQA: A Dataset for VQA on Document Images. CoRR abs/2007.00398 (2020)},
author={Mathew, Minesh and Karatzas, Dimosthenis and Manmatha, R and Jawahar, CV},
journal={arXiv preprint arXiv:2007.00398},
year={2020}
}
86
stars
7
commits
1
linked in READMEs
Apr 18, 2024
updated
Accelerating the development of large-scale multi-modality models (LMMs) with
lmms-eval
π Homepage | π Documentation | π€ Huggingface Datasets
This is a formatted version of DocVQA. It is used in our lmms-eval pipeline to allow for one-click evaluations of large multi-modality models.
@article{mathew2020docvqa,
title={DocVQA: A Dataset for VQA on Document Images. CoRR abs/2007.00398 (2020)},
author={Mathew, Minesh and Karatzas, Dimosthenis and Manmatha, R and Jawahar, CV},
journal={arXiv preprint arXiv:2007.00398},
year={2020}
}