1 repo
Generative models and frameworks for creating videos from textual descriptions, still images, or multimodal prompts using diffusion-based and neural approaches. This cluster encompasses training pipelines, inference optimizations, and end-to-end systems for converting natural language or visual inputs into coherent video sequences, with repositories like VideoCrafter, t2v-turbo, and TIP-I2V representing different stages of the technology stack from core model development to practical deployment.