Video Generation from Text and Images

1 repo

Generative models and frameworks for creating videos from textual descriptions, still images, or multimodal prompts using diffusion-based and neural approaches. This cluster encompasses training pipelines, inference optimizations, and end-to-end systems for converting natural language or visual inputs into coherent video sequences, with repositories like VideoCrafter, t2v-turbo, and TIP-I2V representing different stages of the technology stack from core model development to practical deployment.

Python · 1
fvd ·586
metrics ·586
psnr ·586
pytorch ·586
ssim ·586
video ·586
video-generation ·586
video-prediction ·586