CaiYuanhao/OmniVCus-Test

Dataset

1

stars

8

commits

1

linked in READMEs

Dec 28, 2025

updated

README

[NeurIPS 2025] OmniVCus: Feedforward Subject-driven Video Customization with Multimodal Control Conditions

Dataset Description

This is a testing dataset for multi-modal control video generation. It contains 648 manually collected and annotated data samples to support reference-to-video, reference-mask-to-video, reference-depth-to-video, and reference-instruction-to-video customization.

Here is the data overview:

Task#SubjectNumberPath
Reference-to-Video1113reference/single_subject
Reference-to-Video276reference/double_subject
Reference-to-Video374reference/three_subject
Reference-to-Video456reference/four_subject
Reference-mask-to-video168mask/single_subject
Reference-depth-to-video1108depth/single_subject
Reference-depth-to-video340depth/three_subject
Reference-instruction-to-video1113instruct_edit/single_subject

We also release a training dataset on HuggingFace at

https://huggingface.co/datasets/CaiYuanhao/OmniVCus-Train

This dataset is intended to be used together with our code. Please refer to the GitHub repository below for more detailed instructions.

https://github.com/caiyuanhao1998/Open-OmniVCus

We also release three models based on Wan2.1-1.3B, Wan2.1-14B, and Wan2.2-14B in the following link:

https://huggingface.co/CaiYuanhao/OmniVCus

For more video customization results, please refer to our project page:

https://caiyuanhao1998.github.io/project/OmniVCus/

For more technical details, please refer to our NeurIPS 2025 paper:

https://arxiv.org/abs/2506.23361

Citation

If you find our code, data, and models useful, please consider citing our paper:

@inproceedings{omnivcus,
  title={OmniVCus: Feedforward Subject-driven Video Customization with Multimodal Control Conditions},
  author={Yuanhao Cai and He Zhang and Xi Chen and Jinbo Xing and Kai Zhang and Yiwei Hu and Yuqian Zhou and Zhifei Zhang and Soo Ye Kim and Tianyu Wang and Yulun Zhang and Xiaokang Yang and Zhe Lin and Alan Yuille},
  booktitle={NeurIPS},
  year={2025}
}

Contributors

CaiYuanhao

8 commits

CaiYuanhao/OmniVCus-Test

Dataset

1

stars

8

commits

1

linked in READMEs

Dec 28, 2025

updated

README

[NeurIPS 2025] OmniVCus: Feedforward Subject-driven Video Customization with Multimodal Control Conditions

Dataset Description

This is a testing dataset for multi-modal control video generation. It contains 648 manually collected and annotated data samples to support reference-to-video, reference-mask-to-video, reference-depth-to-video, and reference-instruction-to-video customization.

Here is the data overview:

Task#SubjectNumberPath
Reference-to-Video1113reference/single_subject
Reference-to-Video276reference/double_subject
Reference-to-Video374reference/three_subject
Reference-to-Video456reference/four_subject
Reference-mask-to-video168mask/single_subject
Reference-depth-to-video1108depth/single_subject
Reference-depth-to-video340depth/three_subject
Reference-instruction-to-video1113instruct_edit/single_subject

We also release a training dataset on HuggingFace at

https://huggingface.co/datasets/CaiYuanhao/OmniVCus-Train

This dataset is intended to be used together with our code. Please refer to the GitHub repository below for more detailed instructions.

https://github.com/caiyuanhao1998/Open-OmniVCus

We also release three models based on Wan2.1-1.3B, Wan2.1-14B, and Wan2.2-14B in the following link:

https://huggingface.co/CaiYuanhao/OmniVCus

For more video customization results, please refer to our project page:

https://caiyuanhao1998.github.io/project/OmniVCus/

For more technical details, please refer to our NeurIPS 2025 paper:

https://arxiv.org/abs/2506.23361

Citation

If you find our code, data, and models useful, please consider citing our paper:

@inproceedings{omnivcus,
  title={OmniVCus: Feedforward Subject-driven Video Customization with Multimodal Control Conditions},
  author={Yuanhao Cai and He Zhang and Xi Chen and Jinbo Xing and Kai Zhang and Yiwei Hu and Yuqian Zhou and Zhifei Zhang and Soo Ye Kim and Tianyu Wang and Yulun Zhang and Xiaokang Yang and Zhe Lin and Alan Yuille},
  booktitle={NeurIPS},
  year={2025}
}

Contributors

CaiYuanhao

8 commits