Text-to-Video and Image-to-Video Generation

24 repos

Tools and models for generating video content from text prompts and images, with a focus on diffusion-based approaches. The cluster centers on CogVideoX variants and related implementations that enable video synthesis at various resolutions and efficiency levels. Researchers and practitioners exploring this area will find model implementations, inference frameworks, and techniques for scaling video generation across different hardware constraints.

Python · 2
text-to-video ·22,803
image-to-video ·19,402
video-generation ·16,769
cogvideox ·15,100
llm ·13,024
sora ·13,024
safetensors ·9,169
diffusers ·9,144
multimodal ·6,534
audio-video-generation ·5,442