2hiTee/awesome-3D-Generation

This is a collective repository for all 3D and 4D Object Generation papers

20

15 commits

updated May 22, 2026

See the code

README

🚀 Awesome 3D/4D Generation 🚀

3D Generation

A curated collection of cutting-edge research papers on 3D and 4D Content Generation.

Awesome PRs Welcome If you find this repo useful, please consider giving it a ⭐!


🔗 Explore Our Other Curated Lists

ListDescription
✏️Awesome 3D/4D EditingPapers on 3D and 4D scene/object editing
🔄Awesome FeedForward 3D/4D ReconstructionPapers on feed-forward 3D/4D reconstruction

📋 Table of Contents


🔥 SDS Based

Score Distillation Sampling approaches for text/image-to-3D generation.

#PaperLink
1DreamFusion: Text-to-3D Using 2D Diffusion📄
2Latent-NeRF for Shape-Guided Generation of 3D Shapes and Textures📄
3Magic3D: High-Resolution Text-to-3D Content Creation📄
4NerfDiff: Single-Image View Synthesis with NeRF-Guided Distillation from 3D-Aware Diffusion📄
5DreamBooth3D: Subject-Driven Text-to-3D Generation📄
6Fantasia3D: Disentangling Geometry and Appearance for High-Quality Text-to-3D Content Creation📄
7Make-It-3D: High-Fidelity 3D Creation from a Single Image with Diffusion Prior📄
8TextMesh: Generation of Realistic 3D Meshes From Text Prompts📄
9IT3D: Improved Text-to-3D Generation with Explicit View Synthesis📄
10Text-to-3D Using Gaussian Splatting📄
11DreamGaussian: Generative Gaussian Splatting for Efficient 3D Content Creation📄
12GaussianDreamer: Fast Generation from Text to 3D Gaussians by Bridging 2D and 3D Diffusion Models📄
13DreamCraft3D: Hierarchical 3D Generation with Bootstrapped Diffusion Prior📄
14Learn to Optimize Denoising Scores for 3D Generation📄
15REPARO: Compositional 3D Assets Generation with Differentiable 3D Layout Alignment📄
16Recon3D: High Quality 3D Reconstruction from a Single Image Using Generated Back-View Explicit Priors📄
17COMOGen: A Controllable Text-to-3D Multi-Object Generation Framework📄
18Enhancing Single Image to 3D Generation Using Gaussian Splatting and Hybrid Diffusion Priors📄
19ModeDreamer: Mode Guiding Score Distillation for Text-to-3D Generation📄
20Enhanced 3D Generation by 2D Editing📄
21Diverse Score Distillation📄
22Chirpy3D: Continuous Part Latents for Creative 3D Bird Generation📄
23Probability-Flow Distillation: Exact Wasserstein Gradient Flow for High-Fidelity 3D Generation📄

(⬆️ back to top)

🔥 4D (SDS Based)

SDS-based dynamic 4D content generation.

#PaperLink
1Text-to-4D Dynamic Scene Generation📄
2Consistent4D: Consistent 360° Dynamic Object Generation from Monocular Video📄
3Animate124: Animating One Image to 4D Dynamic Scene📄
44D-fy: Text-to-4D Generation Using Hybrid Score Distillation Sampling📄
5DreamGaussian4D: Generative 4D Gaussian Splatting📄
6SC4D: Sparse-Controlled Video-to-4D Generation and Motion Transfer📄

(⬆️ back to top)


👁️ Multi-View Generation Based

Multi-view diffusion models for 3D-consistent image generation and reconstruction.

#PaperLink
1Zero-1-to-3: Zero-Shot One Image to 3D Object📄
2One-2-3-45: Any Single Image to 3D Mesh in 45 Seconds Without Per-Shape Optimization📄
3MVDiffusion: Enabling Holistic Multi-View Image Generation with Correspondence-Aware Diffusion📄
4MVDream: Multi-View Diffusion for 3D Generation📄
5SyncDreamer: Generating Multiview-Consistent Images from a Single-View Image📄
6Wonder3D: Single Image to 3D Using Cross-Domain Diffusion📄
7Zero123++: A Single Image to Consistent Multi-View Diffusion Base Model📄
8DMV3D: Denoising Multi-View Diffusion Using 3D Large Reconstruction Model📄
9ViVid-1-to-3: Novel View Synthesis with Video Diffusion Models📄
10ImageDream: Image-Prompt Multi-View Diffusion for 3D Generation📄
114DGen: Grounded 4D Content Generation with Spatial-Temporal Consistency📄
12EscherNet: A Generative Model for Scalable View Synthesis📄
13LGM: Large Multi-View Gaussian Model for High-Resolution 3D Content Creation📄
14IM-3D: Iterative Multiview Diffusion and Reconstruction for High-Quality 3D Generation📄
15MVDiffusion++: A Dense High-Resolution Multi-View Diffusion Model for 3D Object Reconstruction📄
16V3D: Video Diffusion Models Are Effective 3D Generators📄
17Make-Your-3D: Fast and Consistent Subject-Driven 3D Content Generation📄
18SV3D: Novel Multi-View Synthesis and 3D Generation from a Single Image Using Latent Video Diffusion📄
19VFusion3D: Learning Scalable 3D Generative Models from Video Diffusion Models📄
20CAT3D: Create Anything in 3D with Multi-View Diffusion Models📄
21Era3D: High-Resolution Multiview Diffusion Using Efficient Row-Wise Attention📄
22Hi3D: Pursuing High-Resolution Image-to-3D Generation with Video Diffusion Models📄
23Enhancing Single Image to 3D Generation Using Gaussian Splatting and Hybrid Diffusion Priors📄
24DreamCraft3D++: Efficient Hierarchical 3D Generation with Multi-Plane Reconstruction Model📄
25MVPaint: Synchronized Multi-View Diffusion for Painting Anything 3D📄
26Edify 3D: Scalable High-Quality 3D Asset Generation📄
27Direct and Explicit 3D Generation from a Single Image📄
28Fancy123: One Image to High-Quality 3D Mesh Generation via Plug-and-Play Deformation📄
29ModeDreamer: Mode Guiding Score Distillation for Text-to-3D Generation📄
30RIGI: Rectifying Image-to-3D Generation Inconsistency via Uncertainty-Aware Learning📄
31Turbo3D: Ultra-Fast Text-to-3D Generation📄
32Gen-3Diffusion: Realistic Image-to-3D Generation via 2D & 3D Diffusion Synergy📄
33PartGen: Part-Level 3D Generation and Reconstruction with Multi-View Diffusion Models📄

(⬆️ back to top)

👁️ 4D (Multi-View Based)

Multi-view generation approaches for dynamic 4D content.

#PaperLink
14DGen: Grounded 4D Content Generation with Spatial-Temporal Consistency📄
2STAG4D: Spatial-Temporal Anchored Generative 4D Gaussians📄
3Diffusion4D: Fast Spatial-Temporal Consistent 4D Generation via Video Diffusion Models📄
4L4GM: Large 4D Gaussian Reconstruction Model📄
5SV4D: Dynamic 3D Content Generation with Multi-Frame and Multi-View Consistency📄

(⬆️ back to top)


🏗️ Naïve 3D Generation: LRM Based

Large Reconstruction Model (LRM) based feed-forward 3D generation.

#PaperLink
1LRM: Large Reconstruction Model for Single Image to 3D📄
2Instant3D: Fast Text-to-3D with Sparse-View Generation and Large Reconstruction Model📄
3DMV3D: Denoising Multi-View Diffusion Using 3D Large Reconstruction Model📄
4PF-LRM: Pose-Free Large Reconstruction Model for Joint Pose and Shape Prediction📄
5TripoSR: Fast 3D Object Reconstruction from a Single Image📄
6GRM: Large Gaussian Reconstruction Model for Efficient 3D Reconstruction and Generation📄
7InstantMesh: Efficient 3D Mesh Generation from a Single Image with Sparse-View Large Reconstruction Models📄
8M-LRM: Multi-View Large Reconstruction Model📄
9LRM-Zero: Training Large Reconstruction Models with Synthesized Data📄
10ControLRM: Fast and Controllable 3D Generation via Large Reconstruction Model📄
11Enhancing Single Image to 3D Generation Using Gaussian Splatting and Hybrid Diffusion Priors📄
123D-Adapter: Geometry-Consistent Multi-View Diffusion for High-Quality 3D Generation📄
13Fancy123: One Image to High-Quality 3D Mesh Generation via Plug-and-Play Deformation📄

(⬆️ back to top)


🎯 Naïve 3D Generation

Naïve 3D native generation methods (diffusion in 3D space, autoregressive, etc.).

#PaperLink
1DiffRF: Rendering-Guided 3D Radiance Field Diffusion📄
2Point-E: A System for Generating 3D Point Clouds from Complex Prompts📄
3Shap-E: Generating Conditional 3D Implicit Functions📄
4GSD: View-Guided Gaussian Splatting Diffusion for 3D Reconstruction📄
53DTopia-XL: Scaling High-Quality 3D Asset Generation via Primitive Diffusion📄
6SeMv-3D: Towards Semantic and Multi-View Consistency for General Text-to-3D Generation with Triplane Priors📄
7L3DG: Latent 3D Gaussian Diffusion📄
8LucidFusion: Generating 3D Gaussians with Arbitrary Unposed Images📄
9GaussianAnything: Interactive Point Cloud Latent Diffusion for 3D Generation📄
10SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-Scale 3D VQVAE📄
11A Lesson in Splats: Teacher-Guided Diffusion for 3D Gaussian Splats Generation with 2D Supervision📄
12TAR3D: Creating High-Quality 3D Assets via Next-Part Prediction📄
13RELATE3D: REfocusing Latent Adapter for Targeted Local Enhancement and Editing in 3D Generation📄
14AssetFormer: Modular 3D Assets Generation with Autoregressive Transformer📄
15VAR-3D: View-Aware Auto-Regressive Model for Text-to-3D Generation via a 3D Tokenizer📄
16Fuse3D: Generating 3D Assets Controlled by Multi-Image Fusion📄
17GaussianGPT: Towards Autoregressive 3D Gaussian Scene Generation📄
18Omni123: Exploring 3D Native Foundation Models with Limited 3D Data📄
19Hitem3D 2.0: Multi-View Guided Native 3D Texture Generation📄

(⬆️ back to top)


🌀 Implicit Latent Space

Implicit latent space representations for 3D shape generation.

#PaperLink
13DShape2VecSet: A 3D Shape Representation for Neural Fields and Generative Diffusion Models📄
2Michelangelo: Conditional 3D Shape Generation Based on Shape-Image-Text Aligned Latent Representation📄
3CraftsMan3D: High-Fidelity Mesh Generation with 3D Native Generation and Interactive Geometry Refiner📄
4CLAY: A Controllable Large-Scale Generative Model for Creating High-Quality 3D Assets📄
5Dora: Sampling and Benchmarking for 3D Shape Variational Auto-Encoders📄
6TripoSG: High-Fidelity 3D Shape Synthesis Using Large-Scale Rectified Flow Models📄
7Step1X-3D: Towards High-Fidelity and Controllable Generation of Textured 3D Assets📄
8Hunyuan3D 2.1: From Images to High-Fidelity 3D Assets with Production-Ready PBR Material📄
9Hunyuan3D 2.5: Towards High-Fidelity 3D Assets Generation with Ultimate Details📄
10Seed3D 1.0: From images to high-fidelity simulation-ready 3D assets📄
11LATTICE: Democratize high-fidelity 3D generation at scale📄
12UltraShape 1.0: High-Fidelity 3D Shape Generation via Scalable Geometric Refinement📄
13Seed3D 2.0: Advancing High-Fidelity Simulation-Ready 3D Content Generation📄
14Pose-Aware Diffusion for 3D Generation📄
15ROAR-3D: Routing arbitrary views for high-fidelity 3D generation📄

(⬆️ back to top)

🌀 4D (Implicit)

#PaperLink
1Sculpt4D: Generating 4D Shapes via Sparse-Attention Diffusion Transformers📄

(⬆️ back to top)


🔶 Explicit Latent Space

Explicit latent space representations (Gaussians, structured latents, etc.) for generation.

#PaperLink
1GaussianCube: A Structured and Explicit Radiance Representation for 3D Generative Modeling📄
2Structured 3D Latents for Scalable and Versatile 3D Generation📄
3SynCity: Training-Free Generation of 3D Worlds📄
4Hi3DGen: High-Fidelity 3D Geometry Generation from Images via Normal Bridging📄
5DSO: Aligning 3D Generators with Simulation Feedback for Physical Soundness📄
6Sparc3D: Sparse Representation and Construction for High-Resolution 3D Shapes Modeling📄
7Ultra3D: Efficient and High-Fidelity 3D Generation with Part Attention📄
8Few-Step Flow for 3D Generation via Marginal-Data Transport Distillation📄
9ReconViaGen: Towards Accurate Multi-View 3D Object Reconstruction via Generation📄
10SAM 3D: 3Dfy Anything in Images📄
11Wukong's 72 transformations: High-fidelity textured 3D morphing via flow models📄
12Native and Compact Structured Latents for 3D Generation📄
13MorphAny3D: Unleashing the power of structured latent in 3D morphing📄
14Muses: Designing, Composing, Generating Nonexistent Fantasy 3D Creatures without Training📄
15Interp3D: Correspondence-aware Interpolation for Generative Textured 3D Morphing📄
16RelaxFlow: Text-Driven Amodal 3D Generation📄
17MV-SAM3D: Adaptive multi-view fusion for layout-aware 3D generation📄
18Points-to-3D: Structure-Aware 3D Generation with Point Cloud Priors📄
19Pixal3D: Pixel-Aligned 3D Generation from Images📄

(⬆️ back to top)

🔶 4D (Explicit)

#PaperLink
1SS4D: Native 4D Generative Model via Structured Spacetime Latents📄

(⬆️ back to top)


🔺 Triplane

Triplane-based 3D representations for generation.

#PaperLink
1Efficient Geometry-Aware 3D Generative Adversarial Networks📄
2GET3D: A Generative Model of High Quality 3D Textured Shapes Learned from Images📄
3Rodin: A Generative Model for Sculpting 3D Digital Avatars Using Diffusion📄
43DGen: Triplane Latent Diffusion for Textured Mesh Generation📄
5Triplane Meets Gaussian Splatting: Fast and Generalizable Single-View 3D Reconstruction with Transformers📄
6AGG: Amortized Generative 3D Gaussians for Single Image to 3D📄
7LN3Diff: Scalable Latent Neural Fields Diffusion for Speedy 3D Generation📄
8Direct3D: Scalable Image-to-3D Generation via 3D Latent Diffusion Transformer📄
9DiffGS: Functional Gaussian Splatting Diffusion📄

(⬆️ back to top)


🕸️ Mesh Generation

Direct mesh generation via autoregressive or diffusion-based approaches.

#PaperLink
1MeshGPT: Generating Triangle Meshes with Decoder-Only Transformers📄
2MeshAnything: Artist-Created Mesh Generation with Autoregressive Transformers📄
3MeshAnything V2: Artist-Created Mesh Generation with Adjacent Mesh Tokenization📄
4EdgeRunner: Auto-Regressive Auto-Encoder for Artistic Mesh Generation📄
5Scaling Mesh Generation via Compressive Tokenization📄
6LLaMA-Mesh: Unifying 3D Mesh Generation with Language Models📄
7TreeMeshGPT: Artistic Mesh Generation with Autoregressive Tree Sequencing📄
8DeepMesh: Auto-Regressive Artist-Mesh Creation with Reinforcement Learning📄
9MeshCraft: Exploring Efficient and Controllable Mesh Generation with Flow-Based DiTs📄
10Mesh-RFT: Enhancing Mesh Generation via Fine-Grained Reinforcement Fine-Tuning📄
11Topology-Preserved Auto-Regressive Mesh Generation in the Manner of Weaving Silk📄
12XSpecMesh: Quality-Preserving Auto-Regressive Mesh Generation Acceleration📄
13VertexRegen: Mesh Generation with Continuous Level of Detail📄
14FastMesh: Efficient Artistic Mesh Generation via Component Decoupling📄
15MeshMosaic: Scaling Artist Mesh Generation via Local-to-Global Assembly📄
16ARMesh: Autoregressive Mesh Generation via Next-Level-of-Detail Prediction📄
17FlashMesh: Faster and Better Autoregressive Mesh Synthesis via Structured Speculation📄
18PartDiffuser: Part-Wise 3D Mesh Generation via Discrete Diffusion📄
19TopGen: Learning structural layouts and cross-fields for quadrilateral mesh generation📄
20FACE: A face-based autoregressive representation for high-fidelity and efficient mesh generation📄
21Strips as Tokens: Artist Mesh Generation with Native UV Segmentation📄

(⬆️ back to top)


🧩 Part Generation

Part-aware and compositional 3D generation.

#PaperLink
1Part123: Part-Aware 3D Reconstruction from a Single-View Image📄
2MIDI: Multi-Instance Diffusion for Single Image to 3D Scene Generation📄
3PartGen: Part-Level 3D Generation and Reconstruction with Multi-View Diffusion Models📄
4PartCrafter: Structured 3D Mesh Generation via Compositional Latent Diffusion Transformers📄
5Efficient Part-Level 3D Object Generation via Dual Volume Packing📄
6Assembler: Scalable 3D Part Assembly via Anchor Point Diffusion📄
7OmniPart: Part-Aware 3D Generation with Semantic Decoupling and Structural Cohesion📄
8From One to More: Contextual Part Latents for 3D Generation📄
9AutoPartGen: Autoregressive 3D Part Generation and Discovery📄
10X-Part: High Fidelity and Structure Coherent Shape Decomposition📄
11FullPart: Generating Each 3D Part at Full Resolution📄
12UniPart: Part-Level 3D Generation with Unified 3D Geom-Seg Latents📄
13EI-Part: Explode for Completion and Implode for Refinement📄
14DreamPartGen: Semantically Grounded Part-Level 3D Generation via Collaborative Latent Denoising📄

(⬆️ back to top)


🦾 Articulate

Articulated object generation, rigging, and animation.

#PaperLink
1URDFormer: A Pipeline for Constructing Articulated Simulation Environments from Real-World Images📄
2Puppet-Master: Scaling Interactive Video Generation as a Motion Prior for Part-Level Dynamics📄
3Articulate-Anything: Automatic Modeling of Articulated Objects via a Vision-Language Foundation Model📄
4SINGAPO: Single Image Controlled Generation of Articulated Parts in Objects📄
5ArtFormer: Controllable Generation of Diverse 3D Articulated Objects📄
6MeshArt: Generating Articulated Meshes with Structure-Guided Transformers📄
7RigAnything: Template-Free Autoregressive Rigging for Diverse 3D Assets📄
8MagicArticulate: Make Your 3D Models Articulation-Ready📄
9One Model to Rig Them All: Diverse Skeleton Rigging with UniRig📄
10Anymate: A Dataset and Baselines for Learning 3D Object Rigging📄
11DreamArt: Generating Interactable Articulated Objects from a Single Image📄
12Puppeteer: Rig and Animate Your 3D Models📄
13Stable Part Diffusion 4D: Multi-View RGB and Kinematic Parts Video Generation📄
14FreeArt3D: Training-Free Articulated Object Generation Using 3D Diffusion📄
15PhysX-Anything: Simulation-Ready Physical 3D Assets from Single Image📄
16Particulate: Feed-Forward 3D Object Articulation📄
17ART: Articulated Reconstruction Transformer📄
18Choreographing a World of Dynamic Objects📄
19RigMo: Unifying Rig and Motion Learning for Generative Animation📄
20PALUM: Part-Based Attention Learning for Unified Motion Retargeting📄
21Motion 3-to-4: 3D Motion Reconstruction for 4D Synthesis📄
22ActionMesh: Animated 3D Mesh Generation with Temporal 3D Diffusion📄
23Skin Tokens: A Learned Compact Representation for Unified Autoregressive Rigging📄
24PAct: Part-Decomposed Single-View Articulated Object Generation📄
25ArtLLM: Generating Articulated Assets via 3D LLM📄
26AniGen: Unified S3 Fields for Animatable 3D Asset Generation📄
27AnimateAnyMesh++: A Flexible 4D Foundation Model for High-Fidelity Text-Driven Mesh Animation📄
28PhysForge: Generating Physics-Grounded 3D Assets for Interactive Virtual World📄
29Rigel3D: Rig-aware Latents for Animation-Ready 3D Asset Generation📄
30R-DMesh: Video-Guided 3D Animation via Rectified Dynamic Mesh Flow📄
31Articraft: An Agentic System for Scalable Articulated 3D Asset Generation📄

(⬆️ back to top)


If you find this repository useful, please consider giving it a ⭐

Made with ❤️ for the 3D Vision Community

Contributors

2hiTee

15 commits

2hiTee/awesome-3D-Generation

This is a collective repository for all 3D and 4D Object Generation papers

20

15 commits

updated May 22, 2026

See the code

README

🚀 Awesome 3D/4D Generation 🚀

3D Generation

A curated collection of cutting-edge research papers on 3D and 4D Content Generation.

Awesome PRs Welcome If you find this repo useful, please consider giving it a ⭐!


🔗 Explore Our Other Curated Lists

ListDescription
✏️Awesome 3D/4D EditingPapers on 3D and 4D scene/object editing
🔄Awesome FeedForward 3D/4D ReconstructionPapers on feed-forward 3D/4D reconstruction

📋 Table of Contents


🔥 SDS Based

Score Distillation Sampling approaches for text/image-to-3D generation.

#PaperLink
1DreamFusion: Text-to-3D Using 2D Diffusion📄
2Latent-NeRF for Shape-Guided Generation of 3D Shapes and Textures📄
3Magic3D: High-Resolution Text-to-3D Content Creation📄
4NerfDiff: Single-Image View Synthesis with NeRF-Guided Distillation from 3D-Aware Diffusion📄
5DreamBooth3D: Subject-Driven Text-to-3D Generation📄
6Fantasia3D: Disentangling Geometry and Appearance for High-Quality Text-to-3D Content Creation📄
7Make-It-3D: High-Fidelity 3D Creation from a Single Image with Diffusion Prior📄
8TextMesh: Generation of Realistic 3D Meshes From Text Prompts📄
9IT3D: Improved Text-to-3D Generation with Explicit View Synthesis📄
10Text-to-3D Using Gaussian Splatting📄
11DreamGaussian: Generative Gaussian Splatting for Efficient 3D Content Creation📄
12GaussianDreamer: Fast Generation from Text to 3D Gaussians by Bridging 2D and 3D Diffusion Models📄
13DreamCraft3D: Hierarchical 3D Generation with Bootstrapped Diffusion Prior📄
14Learn to Optimize Denoising Scores for 3D Generation📄
15REPARO: Compositional 3D Assets Generation with Differentiable 3D Layout Alignment📄
16Recon3D: High Quality 3D Reconstruction from a Single Image Using Generated Back-View Explicit Priors📄
17COMOGen: A Controllable Text-to-3D Multi-Object Generation Framework📄
18Enhancing Single Image to 3D Generation Using Gaussian Splatting and Hybrid Diffusion Priors📄
19ModeDreamer: Mode Guiding Score Distillation for Text-to-3D Generation📄
20Enhanced 3D Generation by 2D Editing📄
21Diverse Score Distillation📄
22Chirpy3D: Continuous Part Latents for Creative 3D Bird Generation📄
23Probability-Flow Distillation: Exact Wasserstein Gradient Flow for High-Fidelity 3D Generation📄

(⬆️ back to top)

🔥 4D (SDS Based)

SDS-based dynamic 4D content generation.

#PaperLink
1Text-to-4D Dynamic Scene Generation📄
2Consistent4D: Consistent 360° Dynamic Object Generation from Monocular Video📄
3Animate124: Animating One Image to 4D Dynamic Scene📄
44D-fy: Text-to-4D Generation Using Hybrid Score Distillation Sampling📄
5DreamGaussian4D: Generative 4D Gaussian Splatting📄
6SC4D: Sparse-Controlled Video-to-4D Generation and Motion Transfer📄

(⬆️ back to top)


👁️ Multi-View Generation Based

Multi-view diffusion models for 3D-consistent image generation and reconstruction.

#PaperLink
1Zero-1-to-3: Zero-Shot One Image to 3D Object📄
2One-2-3-45: Any Single Image to 3D Mesh in 45 Seconds Without Per-Shape Optimization📄
3MVDiffusion: Enabling Holistic Multi-View Image Generation with Correspondence-Aware Diffusion📄
4MVDream: Multi-View Diffusion for 3D Generation📄
5SyncDreamer: Generating Multiview-Consistent Images from a Single-View Image📄
6Wonder3D: Single Image to 3D Using Cross-Domain Diffusion📄
7Zero123++: A Single Image to Consistent Multi-View Diffusion Base Model📄
8DMV3D: Denoising Multi-View Diffusion Using 3D Large Reconstruction Model📄
9ViVid-1-to-3: Novel View Synthesis with Video Diffusion Models📄
10ImageDream: Image-Prompt Multi-View Diffusion for 3D Generation📄
114DGen: Grounded 4D Content Generation with Spatial-Temporal Consistency📄
12EscherNet: A Generative Model for Scalable View Synthesis📄
13LGM: Large Multi-View Gaussian Model for High-Resolution 3D Content Creation📄
14IM-3D: Iterative Multiview Diffusion and Reconstruction for High-Quality 3D Generation📄
15MVDiffusion++: A Dense High-Resolution Multi-View Diffusion Model for 3D Object Reconstruction📄
16V3D: Video Diffusion Models Are Effective 3D Generators📄
17Make-Your-3D: Fast and Consistent Subject-Driven 3D Content Generation📄
18SV3D: Novel Multi-View Synthesis and 3D Generation from a Single Image Using Latent Video Diffusion📄
19VFusion3D: Learning Scalable 3D Generative Models from Video Diffusion Models📄
20CAT3D: Create Anything in 3D with Multi-View Diffusion Models📄
21Era3D: High-Resolution Multiview Diffusion Using Efficient Row-Wise Attention📄
22Hi3D: Pursuing High-Resolution Image-to-3D Generation with Video Diffusion Models📄
23Enhancing Single Image to 3D Generation Using Gaussian Splatting and Hybrid Diffusion Priors📄
24DreamCraft3D++: Efficient Hierarchical 3D Generation with Multi-Plane Reconstruction Model📄
25MVPaint: Synchronized Multi-View Diffusion for Painting Anything 3D📄
26Edify 3D: Scalable High-Quality 3D Asset Generation📄
27Direct and Explicit 3D Generation from a Single Image📄
28Fancy123: One Image to High-Quality 3D Mesh Generation via Plug-and-Play Deformation📄
29ModeDreamer: Mode Guiding Score Distillation for Text-to-3D Generation📄
30RIGI: Rectifying Image-to-3D Generation Inconsistency via Uncertainty-Aware Learning📄
31Turbo3D: Ultra-Fast Text-to-3D Generation📄
32Gen-3Diffusion: Realistic Image-to-3D Generation via 2D & 3D Diffusion Synergy📄
33PartGen: Part-Level 3D Generation and Reconstruction with Multi-View Diffusion Models📄

(⬆️ back to top)

👁️ 4D (Multi-View Based)

Multi-view generation approaches for dynamic 4D content.

#PaperLink
14DGen: Grounded 4D Content Generation with Spatial-Temporal Consistency📄
2STAG4D: Spatial-Temporal Anchored Generative 4D Gaussians📄
3Diffusion4D: Fast Spatial-Temporal Consistent 4D Generation via Video Diffusion Models📄
4L4GM: Large 4D Gaussian Reconstruction Model📄
5SV4D: Dynamic 3D Content Generation with Multi-Frame and Multi-View Consistency📄

(⬆️ back to top)


🏗️ Naïve 3D Generation: LRM Based

Large Reconstruction Model (LRM) based feed-forward 3D generation.

#PaperLink
1LRM: Large Reconstruction Model for Single Image to 3D📄
2Instant3D: Fast Text-to-3D with Sparse-View Generation and Large Reconstruction Model📄
3DMV3D: Denoising Multi-View Diffusion Using 3D Large Reconstruction Model📄
4PF-LRM: Pose-Free Large Reconstruction Model for Joint Pose and Shape Prediction📄
5TripoSR: Fast 3D Object Reconstruction from a Single Image📄
6GRM: Large Gaussian Reconstruction Model for Efficient 3D Reconstruction and Generation📄
7InstantMesh: Efficient 3D Mesh Generation from a Single Image with Sparse-View Large Reconstruction Models📄
8M-LRM: Multi-View Large Reconstruction Model📄
9LRM-Zero: Training Large Reconstruction Models with Synthesized Data📄
10ControLRM: Fast and Controllable 3D Generation via Large Reconstruction Model📄
11Enhancing Single Image to 3D Generation Using Gaussian Splatting and Hybrid Diffusion Priors📄
123D-Adapter: Geometry-Consistent Multi-View Diffusion for High-Quality 3D Generation📄
13Fancy123: One Image to High-Quality 3D Mesh Generation via Plug-and-Play Deformation📄

(⬆️ back to top)


🎯 Naïve 3D Generation

Naïve 3D native generation methods (diffusion in 3D space, autoregressive, etc.).

#PaperLink
1DiffRF: Rendering-Guided 3D Radiance Field Diffusion📄
2Point-E: A System for Generating 3D Point Clouds from Complex Prompts📄
3Shap-E: Generating Conditional 3D Implicit Functions📄
4GSD: View-Guided Gaussian Splatting Diffusion for 3D Reconstruction📄
53DTopia-XL: Scaling High-Quality 3D Asset Generation via Primitive Diffusion📄
6SeMv-3D: Towards Semantic and Multi-View Consistency for General Text-to-3D Generation with Triplane Priors📄
7L3DG: Latent 3D Gaussian Diffusion📄
8LucidFusion: Generating 3D Gaussians with Arbitrary Unposed Images📄
9GaussianAnything: Interactive Point Cloud Latent Diffusion for 3D Generation📄
10SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-Scale 3D VQVAE📄
11A Lesson in Splats: Teacher-Guided Diffusion for 3D Gaussian Splats Generation with 2D Supervision📄
12TAR3D: Creating High-Quality 3D Assets via Next-Part Prediction📄
13RELATE3D: REfocusing Latent Adapter for Targeted Local Enhancement and Editing in 3D Generation📄
14AssetFormer: Modular 3D Assets Generation with Autoregressive Transformer📄
15VAR-3D: View-Aware Auto-Regressive Model for Text-to-3D Generation via a 3D Tokenizer📄
16Fuse3D: Generating 3D Assets Controlled by Multi-Image Fusion📄
17GaussianGPT: Towards Autoregressive 3D Gaussian Scene Generation📄
18Omni123: Exploring 3D Native Foundation Models with Limited 3D Data📄
19Hitem3D 2.0: Multi-View Guided Native 3D Texture Generation📄

(⬆️ back to top)


🌀 Implicit Latent Space

Implicit latent space representations for 3D shape generation.

#PaperLink
13DShape2VecSet: A 3D Shape Representation for Neural Fields and Generative Diffusion Models📄
2Michelangelo: Conditional 3D Shape Generation Based on Shape-Image-Text Aligned Latent Representation📄
3CraftsMan3D: High-Fidelity Mesh Generation with 3D Native Generation and Interactive Geometry Refiner📄
4CLAY: A Controllable Large-Scale Generative Model for Creating High-Quality 3D Assets📄
5Dora: Sampling and Benchmarking for 3D Shape Variational Auto-Encoders📄
6TripoSG: High-Fidelity 3D Shape Synthesis Using Large-Scale Rectified Flow Models📄
7Step1X-3D: Towards High-Fidelity and Controllable Generation of Textured 3D Assets📄
8Hunyuan3D 2.1: From Images to High-Fidelity 3D Assets with Production-Ready PBR Material📄
9Hunyuan3D 2.5: Towards High-Fidelity 3D Assets Generation with Ultimate Details📄
10Seed3D 1.0: From images to high-fidelity simulation-ready 3D assets📄
11LATTICE: Democratize high-fidelity 3D generation at scale📄
12UltraShape 1.0: High-Fidelity 3D Shape Generation via Scalable Geometric Refinement📄
13Seed3D 2.0: Advancing High-Fidelity Simulation-Ready 3D Content Generation📄
14Pose-Aware Diffusion for 3D Generation📄
15ROAR-3D: Routing arbitrary views for high-fidelity 3D generation📄

(⬆️ back to top)

🌀 4D (Implicit)

#PaperLink
1Sculpt4D: Generating 4D Shapes via Sparse-Attention Diffusion Transformers📄

(⬆️ back to top)


🔶 Explicit Latent Space

Explicit latent space representations (Gaussians, structured latents, etc.) for generation.

#PaperLink
1GaussianCube: A Structured and Explicit Radiance Representation for 3D Generative Modeling📄
2Structured 3D Latents for Scalable and Versatile 3D Generation📄
3SynCity: Training-Free Generation of 3D Worlds📄
4Hi3DGen: High-Fidelity 3D Geometry Generation from Images via Normal Bridging📄
5DSO: Aligning 3D Generators with Simulation Feedback for Physical Soundness📄
6Sparc3D: Sparse Representation and Construction for High-Resolution 3D Shapes Modeling📄
7Ultra3D: Efficient and High-Fidelity 3D Generation with Part Attention📄
8Few-Step Flow for 3D Generation via Marginal-Data Transport Distillation📄
9ReconViaGen: Towards Accurate Multi-View 3D Object Reconstruction via Generation📄
10SAM 3D: 3Dfy Anything in Images📄
11Wukong's 72 transformations: High-fidelity textured 3D morphing via flow models📄
12Native and Compact Structured Latents for 3D Generation📄
13MorphAny3D: Unleashing the power of structured latent in 3D morphing📄
14Muses: Designing, Composing, Generating Nonexistent Fantasy 3D Creatures without Training📄
15Interp3D: Correspondence-aware Interpolation for Generative Textured 3D Morphing📄
16RelaxFlow: Text-Driven Amodal 3D Generation📄
17MV-SAM3D: Adaptive multi-view fusion for layout-aware 3D generation📄
18Points-to-3D: Structure-Aware 3D Generation with Point Cloud Priors📄
19Pixal3D: Pixel-Aligned 3D Generation from Images📄

(⬆️ back to top)

🔶 4D (Explicit)

#PaperLink
1SS4D: Native 4D Generative Model via Structured Spacetime Latents📄

(⬆️ back to top)


🔺 Triplane

Triplane-based 3D representations for generation.

#PaperLink
1Efficient Geometry-Aware 3D Generative Adversarial Networks📄
2GET3D: A Generative Model of High Quality 3D Textured Shapes Learned from Images📄
3Rodin: A Generative Model for Sculpting 3D Digital Avatars Using Diffusion📄
43DGen: Triplane Latent Diffusion for Textured Mesh Generation📄
5Triplane Meets Gaussian Splatting: Fast and Generalizable Single-View 3D Reconstruction with Transformers📄
6AGG: Amortized Generative 3D Gaussians for Single Image to 3D📄
7LN3Diff: Scalable Latent Neural Fields Diffusion for Speedy 3D Generation📄
8Direct3D: Scalable Image-to-3D Generation via 3D Latent Diffusion Transformer📄
9DiffGS: Functional Gaussian Splatting Diffusion📄

(⬆️ back to top)


🕸️ Mesh Generation

Direct mesh generation via autoregressive or diffusion-based approaches.

#PaperLink
1MeshGPT: Generating Triangle Meshes with Decoder-Only Transformers📄
2MeshAnything: Artist-Created Mesh Generation with Autoregressive Transformers📄
3MeshAnything V2: Artist-Created Mesh Generation with Adjacent Mesh Tokenization📄
4EdgeRunner: Auto-Regressive Auto-Encoder for Artistic Mesh Generation📄
5Scaling Mesh Generation via Compressive Tokenization📄
6LLaMA-Mesh: Unifying 3D Mesh Generation with Language Models📄
7TreeMeshGPT: Artistic Mesh Generation with Autoregressive Tree Sequencing📄
8DeepMesh: Auto-Regressive Artist-Mesh Creation with Reinforcement Learning📄
9MeshCraft: Exploring Efficient and Controllable Mesh Generation with Flow-Based DiTs📄
10Mesh-RFT: Enhancing Mesh Generation via Fine-Grained Reinforcement Fine-Tuning📄
11Topology-Preserved Auto-Regressive Mesh Generation in the Manner of Weaving Silk📄
12XSpecMesh: Quality-Preserving Auto-Regressive Mesh Generation Acceleration📄
13VertexRegen: Mesh Generation with Continuous Level of Detail📄
14FastMesh: Efficient Artistic Mesh Generation via Component Decoupling📄
15MeshMosaic: Scaling Artist Mesh Generation via Local-to-Global Assembly📄
16ARMesh: Autoregressive Mesh Generation via Next-Level-of-Detail Prediction📄
17FlashMesh: Faster and Better Autoregressive Mesh Synthesis via Structured Speculation📄
18PartDiffuser: Part-Wise 3D Mesh Generation via Discrete Diffusion📄
19TopGen: Learning structural layouts and cross-fields for quadrilateral mesh generation📄
20FACE: A face-based autoregressive representation for high-fidelity and efficient mesh generation📄
21Strips as Tokens: Artist Mesh Generation with Native UV Segmentation📄

(⬆️ back to top)


🧩 Part Generation

Part-aware and compositional 3D generation.

#PaperLink
1Part123: Part-Aware 3D Reconstruction from a Single-View Image📄
2MIDI: Multi-Instance Diffusion for Single Image to 3D Scene Generation📄
3PartGen: Part-Level 3D Generation and Reconstruction with Multi-View Diffusion Models📄
4PartCrafter: Structured 3D Mesh Generation via Compositional Latent Diffusion Transformers📄
5Efficient Part-Level 3D Object Generation via Dual Volume Packing📄
6Assembler: Scalable 3D Part Assembly via Anchor Point Diffusion📄
7OmniPart: Part-Aware 3D Generation with Semantic Decoupling and Structural Cohesion📄
8From One to More: Contextual Part Latents for 3D Generation📄
9AutoPartGen: Autoregressive 3D Part Generation and Discovery📄
10X-Part: High Fidelity and Structure Coherent Shape Decomposition📄
11FullPart: Generating Each 3D Part at Full Resolution📄
12UniPart: Part-Level 3D Generation with Unified 3D Geom-Seg Latents📄
13EI-Part: Explode for Completion and Implode for Refinement📄
14DreamPartGen: Semantically Grounded Part-Level 3D Generation via Collaborative Latent Denoising📄

(⬆️ back to top)


🦾 Articulate

Articulated object generation, rigging, and animation.

#PaperLink
1URDFormer: A Pipeline for Constructing Articulated Simulation Environments from Real-World Images📄
2Puppet-Master: Scaling Interactive Video Generation as a Motion Prior for Part-Level Dynamics📄
3Articulate-Anything: Automatic Modeling of Articulated Objects via a Vision-Language Foundation Model📄
4SINGAPO: Single Image Controlled Generation of Articulated Parts in Objects📄
5ArtFormer: Controllable Generation of Diverse 3D Articulated Objects📄
6MeshArt: Generating Articulated Meshes with Structure-Guided Transformers📄
7RigAnything: Template-Free Autoregressive Rigging for Diverse 3D Assets📄
8MagicArticulate: Make Your 3D Models Articulation-Ready📄
9One Model to Rig Them All: Diverse Skeleton Rigging with UniRig📄
10Anymate: A Dataset and Baselines for Learning 3D Object Rigging📄
11DreamArt: Generating Interactable Articulated Objects from a Single Image📄
12Puppeteer: Rig and Animate Your 3D Models📄
13Stable Part Diffusion 4D: Multi-View RGB and Kinematic Parts Video Generation📄
14FreeArt3D: Training-Free Articulated Object Generation Using 3D Diffusion📄
15PhysX-Anything: Simulation-Ready Physical 3D Assets from Single Image📄
16Particulate: Feed-Forward 3D Object Articulation📄
17ART: Articulated Reconstruction Transformer📄
18Choreographing a World of Dynamic Objects📄
19RigMo: Unifying Rig and Motion Learning for Generative Animation📄
20PALUM: Part-Based Attention Learning for Unified Motion Retargeting📄
21Motion 3-to-4: 3D Motion Reconstruction for 4D Synthesis📄
22ActionMesh: Animated 3D Mesh Generation with Temporal 3D Diffusion📄
23Skin Tokens: A Learned Compact Representation for Unified Autoregressive Rigging📄
24PAct: Part-Decomposed Single-View Articulated Object Generation📄
25ArtLLM: Generating Articulated Assets via 3D LLM📄
26AniGen: Unified S3 Fields for Animatable 3D Asset Generation📄
27AnimateAnyMesh++: A Flexible 4D Foundation Model for High-Fidelity Text-Driven Mesh Animation📄
28PhysForge: Generating Physics-Grounded 3D Assets for Interactive Virtual World📄
29Rigel3D: Rig-aware Latents for Animation-Ready 3D Asset Generation📄
30R-DMesh: Video-Guided 3D Animation via Rectified Dynamic Mesh Flow📄
31Articraft: An Agentic System for Scalable Articulated 3D Asset Generation📄

(⬆️ back to top)


If you find this repository useful, please consider giving it a ⭐

Made with ❤️ for the 3D Vision Community

Contributors

2hiTee

15 commits