Cluster 436915

109 repos across 6 sub-areas

Video Generation and Diffusion Models

37 repos

Libraries and model implementations for generating, processing, and controlling video content using diffusion-based approaches. The cluster centers on video generation frameworks built with popular diffusion libraries like Hugging Face Diffusers, leveraging safe tensor formats for model distribution. Repositories include both general-purpose video generation tools and specialized model variants (such as the Wan2.1-Fun series at different model scales) that enable video synthesis with varying levels of control and computational requirements.

Cluster 462208

35 repos

Cluster 462209

16 repos

AI Video Generation

11 repos

Open-source models and implementations for generating videos from text descriptions and images. The cluster centers on diffusion-based approaches like CogVideoX and Flux, with a focus on making text-to-video synthesis accessible through fine-tuning, inference optimization, and custom dataset adaptation. Developers working in this area will find model weights, training pipelines, and inference frameworks primarily in Python.

Cluster 462210

9 repos

Video Generation from Text and Images

1 repos

Generative models and frameworks for creating videos from textual descriptions, still images, or multimodal prompts using diffusion-based and neural approaches. This cluster encompasses training pipelines, inference optimizations, and end-to-end systems for converting natural language or visual inputs into coherent video sequences, with repositories like VideoCrafter, t2v-turbo, and TIP-I2V representing different stages of the technology stack from core model development to practical deployment.