3 repos
MiZhenxing/ThinkDiff
ICML2025, I Think, Therefore I Diffuse: Enabling Multimodal In-Context Reasoning in Diffusion Models
192
26 commits
OPPO-Mente-Lab/X2I
Official code for ICCV 2025 paper, X2I: Seamless Integration of Multimodal Understanding into…
89
ShihaoZhaoZSH/LaVi-Bridge
[ECCV 2024] Bridging Different Language Models and Generative Vision Models for Text-to-Image…
299
14 commits