🔥🔥🔥A curated list of papers on recent diffusion-based high-resolution image and video synthesis works.
170
70 commits
updated Dec 26, 2024
Collection of recent diffusion-based high-resolution (e.g., $>1024^2$) image and video synthesis works. Questions and discussions are most welcome! Upcoming works will be updated on a regular basis, feel free to contact me to add... :thumbsup:
FreeScale: Unleashing the Resolution of Diffusion Models via Tuning-Free Scale Fusion (12 Dec 2024)
I-Max: Maximize the Resolution Potential of Pre-trained Rectified Flow Transformers with Projected Flow (10 Oct 2024)
HiPrompt: Tuning-free Higher-Resolution Generation with Hierarchical MLLM Prompts (9 Sep 2024)
MegaFusion: Extend Diffusion Models towards Higher-resolution Image Generation without Further Tuning (20 Aug 2024)
ResMaster: Mastering High-Resolution Image Generation via Structural and Fine-Grained Guidance (24 June 2024)
DiffuseHigh: Training-free Progressive High-Resolution Image Synthesis through Structure Guidance (26 June 2024)
ECCV'24 FouriScale: A Frequency Perspective on Training-Free High-Resolution Image Synthesis (19 Mar 2024)
ECCV'24 ZIGMA: A DiT-style Zigzag Mamba Diffusion Model (30 July 2024)
ECCV'24 AccDiffusion: An Accurate Method for Higher-Resolution Image Generation (18 July 2024)
ECCV'24 HiDiffusion: Unlocking High-Resolution Creativity and Efficiency in Low-Resolution Trained Diffusion Models (29 Nov 2023)
ECCV'24 BeyondScene: Higher-Resolution Human-Centric Scene Generation With Pretrained Diffusion (6 Apr 2024)
ECCV'24 Make a Cheap Scaling: A Self-Cascade Diffusion Model for Higher-Resolution Adaptation (16 Feb 2024)
ICML'24 Mastering Text-to-Image Diffusion: Recaptioning, Planning, and Generating with Multimodal LLMs (22 Jan 2024)
CVPR'24 DemoFusion: Democratising High-Resolution Image Generation With No $$$ (24 Nov 2023)
CVPR'24 Generative Powers of Ten (4 Dec 2024)
CVPR'24 FreeU: Free Lunch in Diffusion U-Net (20 Sep 2023)
ICLR'24 ScaleCrafter: Tuning-Free Higher-Resolution Visual Generation with Diffusion Models (11 Oct 2023)
NeurIPS'23 Training-free Diffusion Model Adaptation for Variable-Sized Text-to-Image Synthesis (26 Oct 2023)
NeurIPS'23 SyncDiffusion: Coherent Montage via Synchronized Joint Diffusions (8 Jun 2023)
ICML'23 MultiDiffusion: Fusing Diffusion Paths for Controlled Image Generation (16 Feb 2023)
LINFUSION: 1 GPU, 1 MINUTE, 16K IMAGE (3 Sep 2024)
UltraPixel: Advancing Ultra-High-Resolution Image Synthesis to New Peaks (2 July 2024)
ECCV'24 Make a Cheap Scaling: A Self-Cascade Diffusion Model for Higher-Resolution Adaptation (16 Feb 2024)
CVPR'24 Image Neural Field Diffusion Models (11 Jun 2024)
ICLR'24 PixArt-α: Fast Training of Diffusion Transformer for Photorealistic Text-to-Image Synthesis (13 Apr 2023)
ICCV'23 DiffFit: Unlocking Transferability of Large Diffusion Models via Simple Parameter-Efficient Fine-Tuning (13 Apr 2023)
AAAI'24 Any-Size-Diffusion: Toward Efficient Text-Driven Synthesis for Any-Size HD Images (31 Aug 2023)
Cascaded Model
ICLR'24 Relay Diffusion: Unifying diffusion process across resolutions for image synthesis (4 Sep 2023)
Show-1: Marrying Pixel and Latent Diffusion Models for Text-to-Video Generation (27 Sep 2023)
LaVie: High-Quality Video Generation with Cascaded Latent Diffusion Models (26 Sep 2023)
JMLR'22 [CDM] Cascaded Diffusion Models for High Fidelity Image Generation (30 May 2021)
End-to-End Model
SANA: Efficient High-Resolution Image Synthesis with Linear Diffusion Transformers (14 Oct 2024)
ICML'24 Scalable High-Resolution Pixel-Space Image Synthesis with Hourglass Diffusion Transformers (21 Jan 2024)
ICML'24 FiT: Flexible Vision Transformer for Diffusion Model (19 Feb 2024)
ICLR'24 SDXL: Improving Latent Diffusion Models for High-Resolution Image Synthesis (4 Jul 2023)
ICLR'24 [Patch-DM] Patched Denoising Diffusion Models For High-Resolution Image Synthesis (2 Aug 2023)
ICLR'24 Matryoshka Diffusion Models (23 Oct 2023)
ICLR'24 ∞-Diff: Infinite Resolution Diffusion with Subsampled Mollified States (31 Mar 2023)
On the Importance of Noise Scheduling for Diffusion Models (26 Jan 2023)
ICML'23 Simple diffusion: End-to-end diffusion for high resolution images (26 Jan 2023)
[PromptSR] Image Super-Resolution with Text Prompt Diffusion (24 Nov 2023)
TIP: Text-Driven Image Processing with Semantic and Restoration Instructions (18 Dec 2023)
CVPR'24 [SUPIR] Scaling Up to Excellence: Practicing Model Scaling for Photo-Realistic Image Restoration In the Wild (24 Jan 2024)
CVPR'24 SinSR: Diffusion-Based Image Super-Resolution in a Single Step (23 Nov 2023)
IJCV'24 [StableSR] Exploiting Diffusion Prior for Real-World Image Super-Resolution (11 May 2023)
🔥🔥🔥A curated list of papers on recent diffusion-based high-resolution image and video synthesis works.
170
70 commits
updated Dec 26, 2024
Collection of recent diffusion-based high-resolution (e.g., $>1024^2$) image and video synthesis works. Questions and discussions are most welcome! Upcoming works will be updated on a regular basis, feel free to contact me to add... :thumbsup:
FreeScale: Unleashing the Resolution of Diffusion Models via Tuning-Free Scale Fusion (12 Dec 2024)
I-Max: Maximize the Resolution Potential of Pre-trained Rectified Flow Transformers with Projected Flow (10 Oct 2024)
HiPrompt: Tuning-free Higher-Resolution Generation with Hierarchical MLLM Prompts (9 Sep 2024)
MegaFusion: Extend Diffusion Models towards Higher-resolution Image Generation without Further Tuning (20 Aug 2024)
ResMaster: Mastering High-Resolution Image Generation via Structural and Fine-Grained Guidance (24 June 2024)
DiffuseHigh: Training-free Progressive High-Resolution Image Synthesis through Structure Guidance (26 June 2024)
ECCV'24 FouriScale: A Frequency Perspective on Training-Free High-Resolution Image Synthesis (19 Mar 2024)
ECCV'24 ZIGMA: A DiT-style Zigzag Mamba Diffusion Model (30 July 2024)
ECCV'24 AccDiffusion: An Accurate Method for Higher-Resolution Image Generation (18 July 2024)
ECCV'24 HiDiffusion: Unlocking High-Resolution Creativity and Efficiency in Low-Resolution Trained Diffusion Models (29 Nov 2023)
ECCV'24 BeyondScene: Higher-Resolution Human-Centric Scene Generation With Pretrained Diffusion (6 Apr 2024)
ECCV'24 Make a Cheap Scaling: A Self-Cascade Diffusion Model for Higher-Resolution Adaptation (16 Feb 2024)
ICML'24 Mastering Text-to-Image Diffusion: Recaptioning, Planning, and Generating with Multimodal LLMs (22 Jan 2024)
CVPR'24 DemoFusion: Democratising High-Resolution Image Generation With No $$$ (24 Nov 2023)
CVPR'24 Generative Powers of Ten (4 Dec 2024)
CVPR'24 FreeU: Free Lunch in Diffusion U-Net (20 Sep 2023)
ICLR'24 ScaleCrafter: Tuning-Free Higher-Resolution Visual Generation with Diffusion Models (11 Oct 2023)
NeurIPS'23 Training-free Diffusion Model Adaptation for Variable-Sized Text-to-Image Synthesis (26 Oct 2023)
NeurIPS'23 SyncDiffusion: Coherent Montage via Synchronized Joint Diffusions (8 Jun 2023)
ICML'23 MultiDiffusion: Fusing Diffusion Paths for Controlled Image Generation (16 Feb 2023)
LINFUSION: 1 GPU, 1 MINUTE, 16K IMAGE (3 Sep 2024)
UltraPixel: Advancing Ultra-High-Resolution Image Synthesis to New Peaks (2 July 2024)
ECCV'24 Make a Cheap Scaling: A Self-Cascade Diffusion Model for Higher-Resolution Adaptation (16 Feb 2024)
CVPR'24 Image Neural Field Diffusion Models (11 Jun 2024)
ICLR'24 PixArt-α: Fast Training of Diffusion Transformer for Photorealistic Text-to-Image Synthesis (13 Apr 2023)
ICCV'23 DiffFit: Unlocking Transferability of Large Diffusion Models via Simple Parameter-Efficient Fine-Tuning (13 Apr 2023)
AAAI'24 Any-Size-Diffusion: Toward Efficient Text-Driven Synthesis for Any-Size HD Images (31 Aug 2023)
Cascaded Model
ICLR'24 Relay Diffusion: Unifying diffusion process across resolutions for image synthesis (4 Sep 2023)
Show-1: Marrying Pixel and Latent Diffusion Models for Text-to-Video Generation (27 Sep 2023)
LaVie: High-Quality Video Generation with Cascaded Latent Diffusion Models (26 Sep 2023)
JMLR'22 [CDM] Cascaded Diffusion Models for High Fidelity Image Generation (30 May 2021)
End-to-End Model
SANA: Efficient High-Resolution Image Synthesis with Linear Diffusion Transformers (14 Oct 2024)
ICML'24 Scalable High-Resolution Pixel-Space Image Synthesis with Hourglass Diffusion Transformers (21 Jan 2024)
ICML'24 FiT: Flexible Vision Transformer for Diffusion Model (19 Feb 2024)
ICLR'24 SDXL: Improving Latent Diffusion Models for High-Resolution Image Synthesis (4 Jul 2023)
ICLR'24 [Patch-DM] Patched Denoising Diffusion Models For High-Resolution Image Synthesis (2 Aug 2023)
ICLR'24 Matryoshka Diffusion Models (23 Oct 2023)
ICLR'24 ∞-Diff: Infinite Resolution Diffusion with Subsampled Mollified States (31 Mar 2023)
On the Importance of Noise Scheduling for Diffusion Models (26 Jan 2023)
ICML'23 Simple diffusion: End-to-end diffusion for high resolution images (26 Jan 2023)
[PromptSR] Image Super-Resolution with Text Prompt Diffusion (24 Nov 2023)
TIP: Text-Driven Image Processing with Semantic and Restoration Instructions (18 Dec 2023)
CVPR'24 [SUPIR] Scaling Up to Excellence: Practicing Model Scaling for Photo-Realistic Image Restoration In the Wild (24 Jan 2024)
CVPR'24 SinSR: Diffusion-Based Image Super-Resolution in a Single Step (23 Nov 2023)
IJCV'24 [StableSR] Exploiting Diffusion Prior for Real-World Image Super-Resolution (11 May 2023)