This dataset is used for the training of the LLaVA-Video model. We only allow the use of this dataset for academic research and education purpose. For OpenAI GPT-4 generated data, we recommend the users to check the OpenAI Usage Policy.
For the training of LLaVA-Video, we utilized video-language data from five primary sources:
The LLaVA-Video-178K dataset is the only contribution from this repository; we provide additional datasets for reproducing LLaVA-Video.
The following directories are provided for generating captions and QA data:
LLaVA-Video-178K/gpt4o_caption_promptLLaVA-Video-178K/gpt4o_qa_promptWe have included captions and open-ended questions in the 0_30_s_academic_v0_1 split, along with 240,000 open-ended QA items and 15,000 caption entries, as part of the video data in LLaVA-Hound for LLaVA-OneVision.
@misc{zhang2024videoinstructiontuningsynthetic,
title={Video Instruction Tuning With Synthetic Data},
author={Yuanhan Zhang and Jinming Wu and Wei Li and Bo Li and Zejun Ma and Ziwei Liu and Chunyuan Li},
year={2024},
eprint={2410.02713},
archivePrefix={arXiv},
primaryClass={cs.CV},
url={https://arxiv.org/abs/2410.02713},
}
499 commits
This dataset is used for the training of the LLaVA-Video model. We only allow the use of this dataset for academic research and education purpose. For OpenAI GPT-4 generated data, we recommend the users to check the OpenAI Usage Policy.
For the training of LLaVA-Video, we utilized video-language data from five primary sources:
The LLaVA-Video-178K dataset is the only contribution from this repository; we provide additional datasets for reproducing LLaVA-Video.
The following directories are provided for generating captions and QA data:
LLaVA-Video-178K/gpt4o_caption_promptLLaVA-Video-178K/gpt4o_qa_promptWe have included captions and open-ended questions in the 0_30_s_academic_v0_1 split, along with 240,000 open-ended QA items and 15,000 caption entries, as part of the video data in LLaVA-Hound for LLaVA-OneVision.
@misc{zhang2024videoinstructiontuningsynthetic,
title={Video Instruction Tuning With Synthetic Data},
author={Yuanhan Zhang and Jinming Wu and Wei Li and Bo Li and Zejun Ma and Ziwei Liu and Chunyuan Li},
year={2024},
eprint={2410.02713},
archivePrefix={arXiv},
primaryClass={cs.CV},
url={https://arxiv.org/abs/2410.02713},
}
499 commits