[!IMPORTANT] We only hold the core publicly available ones pre-trained on ImageNet.
A selected collection of CORE visual generative foundation model including code, paper, checkpoint etc.
[!NOTE] β DiT-MoE uses additional synthetic training data generated by FLUX and SD3. FID and IS are evaluated on 50k samples, reported with CFG if applicable. Γ2 in NFEs indicates that CFG doubles NFEs at inference time.
Python
100.0%
[!IMPORTANT] We only hold the core publicly available ones pre-trained on ImageNet.
A selected collection of CORE visual generative foundation model including code, paper, checkpoint etc.
[!NOTE] β DiT-MoE uses additional synthetic training data generated by FLUX and SD3. FID and IS are evaluated on 50k samples, reported with CFG if applicable. Γ2 in NFEs indicates that CFG doubles NFEs at inference time.
Python
100.0%