lmms-lab-encoder/DocVQA

Dataset

86

stars

7

commits

1

linked in READMEs

Apr 18, 2024

updated

Browse cluster: Multimodal Document and Visual Understanding

README

Large-scale Multi-modality Models Evaluation Suite

Accelerating the development of large-scale multi-modality models (LMMs) with lmms-eval

🏠 Homepage | 📚 Documentation | 🤗 Huggingface Datasets

This Dataset

This is a formatted version of DocVQA. It is used in our lmms-eval pipeline to allow for one-click evaluations of large multi-modality models.

@article{mathew2020docvqa,
  title={DocVQA: A Dataset for VQA on Document Images. CoRR abs/2007.00398 (2020)},
  author={Mathew, Minesh and Karatzas, Dimosthenis and Manmatha, R and Jawahar, CV},
  journal={arXiv preprint arXiv:2007.00398},
  year={2020}
}

Contributors

pufanyi

6 commits

luodian

1 commits

lmms-lab-encoder/DocVQA

Dataset

86

stars

7

commits

1

linked in READMEs

Apr 18, 2024

updated

Browse cluster: Multimodal Document and Visual Understanding

README

Large-scale Multi-modality Models Evaluation Suite

Accelerating the development of large-scale multi-modality models (LMMs) with lmms-eval

🏠 Homepage | 📚 Documentation | 🤗 Huggingface Datasets

This Dataset

This is a formatted version of DocVQA. It is used in our lmms-eval pipeline to allow for one-click evaluations of large multi-modality models.

@article{mathew2020docvqa,
  title={DocVQA: A Dataset for VQA on Document Images. CoRR abs/2007.00398 (2020)},
  author={Mathew, Minesh and Karatzas, Dimosthenis and Manmatha, R and Jawahar, CV},
  journal={arXiv preprint arXiv:2007.00398},
  year={2020}
}

Contributors

pufanyi

6 commits

luodian

1 commits