4
stars
10
commits
2
linked in READMEs
Mar 8, 2024
updated
Accelerating the development of large-scale multi-modality models (LMMs) with
lmms-eval
π Homepage | π Documentation | π€ Huggingface Datasets
This is a formatted version of SEED-Bench. It is used in our lmms-eval pipeline to allow for one-click evaluations of large multi-modality models.
@article{li2023seed,
title={Seed-bench: Benchmarking multimodal llms with generative comprehension},
author={Li, Bohao and Wang, Rui and Wang, Guangzhi and Ge, Yuying and Ge, Yixiao and Shan, Ying},
journal={arXiv preprint arXiv:2307.16125},
year={2023}
}
10 commits
4
stars
10
commits
2
linked in READMEs
Mar 8, 2024
updated
Accelerating the development of large-scale multi-modality models (LMMs) with
lmms-eval
π Homepage | π Documentation | π€ Huggingface Datasets
This is a formatted version of SEED-Bench. It is used in our lmms-eval pipeline to allow for one-click evaluations of large multi-modality models.
@article{li2023seed,
title={Seed-bench: Benchmarking multimodal llms with generative comprehension},
author={Li, Bohao and Wang, Rui and Wang, Guangzhi and Ge, Yuying and Ge, Yixiao and Shan, Ying},
journal={arXiv preprint arXiv:2307.16125},
year={2023}
}
10 commits