11 repos
Open-source models and implementations for generating videos from text descriptions and images. The cluster centers on diffusion-based approaches like CogVideoX and Flux, with a focus on making text-to-video synthesis accessible through fine-tuning, inference optimization, and custom dataset adaptation. Developers working in this area will find model weights, training pipelines, and inference frameworks primarily in Python.