yanziang/InternVideo3_Dataset

Dataset

2

stars

3

commits

1

linked in READMEs

Jun 11, 2026

updated

README

Dataset

The InternVideo3 long-video supervised fine-tuning data is available on Hugging Face:

DatasetRowsFormat
yanziang/InternVideo3_Dataset380KJSON / Parquet

Each sample contains a YouTube video id and QA annotations for detailed long-video description and reasoning. If you use this dataset, please cite:

@misc{yan2026internvideo3agentifyfoundationmodels,
      title={InternVideo3: Agentify Foundation Models with Multimodal Contextual Reasoning}, 
      author={Ziang Yan and Sheng Xia and Jiashuo Yu and Yue Wu and Tianxiang Jiang and Songze Li and Kanghui Tian and Yicheng Xu and Yinan He and Kai Chen and Limin Wang and Yu Qiao and Yi Wang},
      year={2026},
      eprint={2606.12195},
      archivePrefix={arXiv},
      primaryClass={cs.CV},
      url={https://arxiv.org/abs/2606.12195}, 
}

Contributors

yanziang

3 commits

yanziang/InternVideo3_Dataset

Dataset

2

stars

3

commits

1

linked in READMEs

Jun 11, 2026

updated

README

Dataset

The InternVideo3 long-video supervised fine-tuning data is available on Hugging Face:

DatasetRowsFormat
yanziang/InternVideo3_Dataset380KJSON / Parquet

Each sample contains a YouTube video id and QA annotations for detailed long-video description and reasoning. If you use this dataset, please cite:

@misc{yan2026internvideo3agentifyfoundationmodels,
      title={InternVideo3: Agentify Foundation Models with Multimodal Contextual Reasoning}, 
      author={Ziang Yan and Sheng Xia and Jiashuo Yu and Yue Wu and Tianxiang Jiang and Songze Li and Kanghui Tian and Yicheng Xu and Yinan He and Kai Chen and Limin Wang and Yu Qiao and Yi Wang},
      year={2026},
      eprint={2606.12195},
      archivePrefix={arXiv},
      primaryClass={cs.CV},
      url={https://arxiv.org/abs/2606.12195}, 
}

Contributors

yanziang

3 commits