kaiyuyue/llava-1.5-665k-instructions

Dataset

10

stars

14

commits

3

linked in READMEs

Aug 5, 2025

updated

vision-language-model
vlm
zero-shot

README

This dataset repository, LLaVA-1.5-665K-Instructions, is notably utilized in the paper Zero-Shot Vision Encoder Grafting via LLM Surrogates.

The official code repository for the paper can be found here: https://github.com/kaiyuyue/zero

LLaVA-1.5-665K-Instructions

This dataset repo contains the entire LLaVA-1.5-665K-Instructions in one place, including images and text sequences. The images are in train_split/*.tars and the text sequences are in jsons:

Details on LLaVA-1.5 dataset construction: For more information on the original LLaVA-1.5 dataset construction, refer to: https://arxiv.org/abs/2310.03744

License: Creative Commons Attribution 4.0 International, from LLaVA HF dataset repo.

Contributors

kaiyuyue

13 commits

nielsr

1 commits

kaiyuyue/llava-1.5-665k-instructions

Dataset

10

stars

14

commits

3

linked in READMEs

Aug 5, 2025

updated

vision-language-model
vlm
zero-shot

README

This dataset repository, LLaVA-1.5-665K-Instructions, is notably utilized in the paper Zero-Shot Vision Encoder Grafting via LLM Surrogates.

The official code repository for the paper can be found here: https://github.com/kaiyuyue/zero

LLaVA-1.5-665K-Instructions

This dataset repo contains the entire LLaVA-1.5-665K-Instructions in one place, including images and text sequences. The images are in train_split/*.tars and the text sequences are in jsons:

Details on LLaVA-1.5 dataset construction: For more information on the original LLaVA-1.5 dataset construction, refer to: https://arxiv.org/abs/2310.03744

License: Creative Commons Attribution 4.0 International, from LLaVA HF dataset repo.

Contributors

kaiyuyue

13 commits

nielsr

1 commits