zhangxulu1996/awesome-personalization

25

stars

20

commits

Python

primary language

Apr 10, 2025

updated

README

Awesome Personalization

example

This repository contains a collection of papers and resources on Personalized Content Synthesis (PCS) with Diffusion Model.

๐Ÿ”— Citation

If you find the information in our paper useful for your research, please consider citing it in your work. Thank you!

@misc{zhang2024survey,
      title={A Survey on Personalized Content Synthesis with Diffusion Models}, 
      author={Xulu Zhang and Xiao-Yong Wei and Wengyu Zhang and Jinlin Wu and Zhaoxiang Zhang and Zhen Lei and Qing Li},
      year={2024},
      eprint={2405.05538},
      archivePrefix={arXiv},
      primaryClass={cs.CV}
}

๐Ÿ“– Contents

๐Ÿ—‚๏ธ Unified Test Dataset

To uniformly evaluate Personalized Content Synthesis (PCS) tasks, we introduces a comprehensive evaluation dataset designed for the most common personalized generation tasks, object and face personalization.

Download Link: PCS-dataset

๐Ÿ”จ Evaluation

We evaluate existing representative PCS methods based our unified test dataset. The evaluation results and settings are shown in the below table. For more details, please see our survey paper.

TypeMethodsFrameworkBackboneCLIP-TCLIP-I
ObjectTextual InversionTTFSD 1.50.1990.749
DreamboothTTFSD 1.50.2860.772
P+TTFSD 1.40.2440.643
Custom DiffusionTTFSD 1.40.3070.722
NeTITTFSD 1.40.2830.801
SVDiffTTFSD 1.50.2820.776
PerfusionTTFSD 1.50.2730.691
ELITEPTASD 1.40.2920.765
BLIP-DiffusionPTASD 1.50.2920.772
IP-AdapterPTASD 1.50.2720.825
SSR EncoderPTASD 1.50.2880.792
MoMAPTASD 1.50.3220.748
Diptych PromptingPTAFLUX 1.0 dev0.3270.722
ฮป-eclipsePTAKandinsky 2.20.2720.824
MS-DiffusionPTASDXL0.2980.777
FaceCrossInitializationTTFSD 2.10.2610.469
Face2DiffusionPTASD 1.40.2650.588
SSR EncoderPTASD 1.50.2330.490
FastComposerPTASD 1.50.2300.516
IP-AdapterPTASD 1.50.2920.462
IP-AdapterPTASDXL0.2920.642
PhotoMakerPTASDXL0.3110.547
InstantIDPTASDXL0.2780.707

TTF: Test-time Fine-tuning, PTA: Pre-trained Adaptation

๐Ÿ“ƒ Paper List

Personalized Object Generation

TitleVenueDateLinks
An Image is Worth One Word: Personalizing Text-to-Image Generation using Textual InversionICLR 20232022
08-02
Code
Paper
DreamBooth: Fine-Tuning Text-to-Image Diffusion Models for Subject-Driven GenerationCVPR 20232022
08-25
Code
Paper
Re-Imagen: Retrieval-Augmented Text-to-Image GeneratorICLR 20232022
09-29
โ€”
Paper
Versatile Diffusion: Text, Images, and Variations All in One Diffusion ModelICCV 20232022
11-15
Code
Paper
DreamArtist: Towards Controllable One-Shot Text-to-Image Generation via Positive-Negative Prompt-TuningarXiv 20222022
11-21
Code
Paper
Is This Loss Informative? Faster Text-to-Image Customization by Tracking Objective DynamicsNeurIPS 20232023
02-09
Code
Paper
Encoder-Based Domain Tuning for Fast Personalization of Text-to-Image ModelsACM Trans on Graphics2023
02-23
Code
Paper
ELITE: Encoding Visual Concepts into Textual Embeddings for Customized Text-to-Image GenerationICCV 20232023
02-27
Code
Paper
Highly Personalized Text Embedding for Image Manipulation by Stable DiffusionarXiv2023
03-15
Code
Paper
Unified Multi-Modal Latent Diffusion for Joint Subject and Text Conditional Image GenerationarXiv2023
03-16
โ€”
Paper
P+: Extended Textual Conditioning in Text-to-Image GenerationarXiv2023
03-16
Code
Paper
A Closer Look at Parameter-Efficient Tuning in Diffusion ModelsarXiv2023
03-31
Code
Paper
Subject-Driven Text-to-Image Generation via Apprenticeship LearningNIPS 20232023
04-01
โ€”
Paper
Taming Encoder for Zero Fine-Tuning Image Customization with Text-to-Image Diffusion ModelsarXiv2023
04-05
โ€”
Paper
InstantBooth: Personalized Text-to-Image Generation Without Test-Time FinetuningarXiv2023
04-06
Code
Paper
Controllable Textual Inversion for Personalized Text-to-Image GenerationarXiv2023
04-11
Code
Paper
Gradient-Free Textual InversionACM MM 20232023
04-12
Code
Paper
Personalize Segment Anything Model with One ShotICLR 20242023
05-04
Code
Paper
DisenBooth: Identity-Preserving Disentangled Tuning for Subject-Driven Text-to-Image GenerationICLR 20242023
05-05
Code
Paper
BLIP-Diffusion: Pre-Trained Subject Representation for Controllable Text-to-Image Generation and EditingNIPS 20232023
05-24
Code
Paper
A Neural Space-Time Representation for Text-to-Image PersonalizationSIGGRAPH Asia 20232023
05-24
Code
Paper
Prospect: Prompt Spectrum for Attribute-Aware Personalization of Diffusion ModelsACM Trans on Graphics2023
05-25
Code
Paper
Break-a-Scene: Extracting Multiple Concepts from a Single ImageSIGGRAPH ASIA 20232023
05-25
Code
Paper
COMCAT: Towards Efficient Compression and Customization of Attention-Based Vision ModelsarXiv2023
05-26
Code
Paper
ViCo: Plug-and-Play Visual Condition for Personalized Text-to-Image GenerationarXiv2023
06-01
Code
Paper
Controlling Text-to-Image Diffusion by Orthogonal Fine-TuningarXiv2023
06-12
Code
Paper
Domain-Agnostic Tuning-Encoder for Fast Personalization of Text-to-Image ModelsSIGGRAPH 20232023
07-13
Code
Paper
IP-Adapter: Text Compatible Image Prompt Adapter for Text-to-Image Diffusion ModelsarXiv2023
08-13
Code
Paper
Navigating Text-to-Image Customization: From Lycoris Fine-Tuning to Model EvaluationICLR 20242023
09-26
Code
Paper
Kosmos-G: Generating Images in Context with Multimodal Large Language ModelsarXiv2023
10-04
Code
Paper
Personalized Text-to-Image Model Enhancement Strategies: SOD Preprocessing and CNN Local Feature IntegrationarXiv2023
10-26
โ€”
Paper
A Data Perspective on Enhanced Identity Preservation for Diffusion PersonalizationICLR 20242023
11-07
Code
Paper
DIFFNAT: Improving Diffusion Image Quality Using Natural Image StatisticsarXiv2023
11-16
โ€”
Paper
An Image is Worth Multiple Words: Multi-attribute Inversion for Constrained Text-to-Image SynthesisarXiv2023
11-20
โ€”
Paper
LEGO: Learning to Disentangle and Invert Concepts Beyond Object Appearance in Text-to-Image Diffusion ModelsarXiv2023
11-23
Code
Paper
Catversion: Concatenating Embeddings for Diffusion-Based Text-to-Image PersonalizationarXiv2023
11-24
Code
Paper
CLiC: Concept Learning in ContextCVPR 20242023
11-28
Code
Paper
HiFi Tuner: High-Fidelity Subject-Driven Fine-Tuning for Diffusion ModelsarXiv2023
11-30
โ€”
Paper
InstructBooth: Instruction-Following Personalized Text-to-Image GenerationarXiv2023
12-04
โ€”
Paper
Customization Assistant for Text-to-Image GenerationCVPR 20242023
12-05
โ€”
Paper
Decoupled Textual Embeddings for Customized Image GenerationAAAI 20242023
12-19
Code
Paper
Towards Accurate Guided Diffusion Sampling through Symplectic Adjoint MethodarXiv2023
12-19
Code
Paper
DreamDistribution: Prompt Distribution Learning for Text-to-Image Diffusion ModelsarXiv2023
12-21
Code
Paper
DreamTuner: Single Image is Enough for Subject-Driven GenerationarXiv2023
12-21
Code
Paper
BootPIG: Bootstrapping Zero-Shot Personalized Image Generation Capabilities in Pretrained Diffusion ModelsarXiv2024
01-25
โ€”
Paper
Object-Driven One-Shot Fine-Tuning of Text-to-Image Diffusion with Prototypical EmbeddingarXiv2024
01-28
โ€”
Paper
DisenDreamer: Subject-Driven Text-to-Image Generation with Sample-aware Disentangled TuningarXiv2024
02-26
โ€”
Paper
Infusion: Preventing Customized Text-to-Image Diffusion from OverfittingarXiv2024
04-22
โ€”
Paper
Customizing Text-to-Image Models with a Single Image PairarXiv2024
05-02
โ€”
Paper

Multi-concept Composition

TitleVenueDateLinks
Multi-Concept Customization of Text-to-Image DiffusionCVPR 20232022
12-08
Code
Paper
CONES: Concept Neurons in Diffusion Models for Customized GenerationICML 20232023
03-09
Code
Paper
SVDiff: Compact Parameter Space for Diffusion Fine-TuningICCV 20232023
03-20
Code
Paper
Key-Locked Rank One Editing for Text-to-Image PersonalizationSIGGRAPH 20232023
05-02
โ€”
Paper
Mix-of-Show: Decentralized Low-Rank Adaptation for Multi-Concept Customization of Diffusion ModelsNIPS 20232023
05-29
Code
Paper
CONES 2: Customizable Image Synthesis with Multiple SubjectsNIPS 20232023
05-30
Code
Paper
Generate Anything Anywhere in Any ScenearXiv2023
06-29
Code
Paper
AnyDoor: Zero-Shot Object-Level Image CustomizationarXiv2023
07-18
Code
Paper
Subject-Diffusion: Open Domain Personalized Text-to-Image Generation Without Test-Time Fine-TuningarXiv2023
07-21
Code
Paper
CustomNet: Zero-Shot Object Customization with Variable-Viewpoints in Text-to-Image Diffusion ModelsarXiv2023
10-30
Code
Paper
Compositional Inversion for Stable Diffusion ModelsAAAI 20242023
12-13
Code
Paper
Visual Concept-Driven Image Generation with Text-to-Image Diffusion ModelarXiv2024
02-18
โ€”
Paper
MIGC: Multi-Instance Generation Controller for Text-to-Image SynthesisarXiv2024
02-27
โ€”
Paper
Multi-Object Editing in Personalized Text-To-Image Diffusion Model Via Segmentation GuidancearXiv2024
03-18
โ€”
Paper
MC2: Multi-concept Guidance for Customized Multi-concept GenerationarXiv2024
04-12
โ€”
Paper
MultiBooth: Towards Generating All Your Concepts in an Image from TextarXiv2024
04-22
โ€”
Paper
MagicTailor: Component-Controllable Personalization in Text-to-Image Diffusion ModelsarXiv2024
10-06
Code
Paper

Personalized Style Generation

TitleVenueDateLinks
StyleDrop: Text-to-Image Synthesis of Any StyleNIPS 20232023
06-01
Code
Paper
StyleAdapter: A Single-Pass LoRA-Free Model for Stylized Image GenerationICLR 20242023
09-04
โ€”
Paper
StyleBoost: A Study of Personalizing Text-to-Image Generation in Any Style using DreamBoothICTC 20232023
10-13
โ€”
Paper
ArtAdapter: Text-to-Image Style Transfer Using Multi-Level Style Encoder and Explicit AdaptationarXiv2023
12-04
Code
Paper
Style Aligned Image Generation via Shared AttentionCVPR 20242023
12-04
Code
Paper
Generative Active Learning for Image Synthesis PersonalizationarXiv2024
03-22
Code
Paper
Text-to-Image Synthesis for Any Artistic Styles: Advancements in Personalized Artistic Image Generation via Subdivision and Dual BindingarXiv2024
04-08
โ€”
Paper

Personalized Face Generation

TitleVenueDateLinks
Identity Encoder for Personalized DiffusionCoRR 20232023
04-14
โ€”
Paper
FastComposer: Tuning-Free Multi-Subject Image Generation with Localized AttentionCoRR 20232023
05-21
Code
Paper
Enhancing Detail Preservation for Customized Text-to-Image Generation: A Regularization-Free ApproacharXiv2023
05-23
โ€”
Paper
Inserting Anybody in Diffusion Models via Celeb BasisNIPS 20232023
06-01
Code
Paper
Face0: Instantaneously Conditioning a Text-to-Image Model on a FaceSIGGRAPH 20232023
06-11
โ€”
Paper
DreamIdentity: Improved Editability for Efficient Face-Identity Preserved Image GenerationarXiv2023
07-01
Code
Paper
HyperDreamBooth: Hypernetworks for Fast Personalization of Text-to-Image ModelsarXiv2023
07-13
Code
Paper
Identity-Preserving Aging of Face Images via Latent Diffusion ModelsIJCB 20232023
07-17
Code
Paper
Magicapture: High-Resolution Multi-Concept Portrait CustomizationarXiv2023
09-13
Code
Paper
High-Fidelity Person-Centric Subject-to-Image SynthesisCVPR 20242023
11-17
โ€”
Paper
When StyleGAN Meets Stable Diffusion: A W+ Adapter for Personalized Image GenerationarXiv2023
11-29
Code
Paper
Portrait Diffusion: Training-Free Face Stylization with Chain-of-PaintingarXiv2023
12-03
Code
Paper
Retrieving Conditions from Reference Images for Diffusion ModelsarXiv2023
12-05
โ€”
Paper
FaceStudio: Put Your Face Everywhere in SecondsarXiv2023
12-05
Code
Paper
Personalized Face Inpainting with Diffusion Models by Parallel Visual AttentionWACV 20242023
12-06
Code
Paper
PhotoMaker: Customizing Realistic Human Photos via Stacked ID EmbeddingarXiv2023
12-07
Code
Paper
DemoCaricature: Democratising Caricature Generation with a Rough SketchCVPR 20242023
12-07
Code
Paper
Stellar: Systematic Evaluation of Human-Centric Personalized Text-to-Image MethodsCoRR 20232023
12-11
Code
Paper
PortraitBooth: A Versatile Portrait Model for Fast Identity-Preserved PersonalizationCVPR 20242023
12-11
Code
Paper
Concept-Centric Personalization with Large-Scale Diffusion PriorsarXiv2023
12-13
Code
Paper
Cross Initialization for Personalized Text-to-Image GenerationarXiv2023
12-26
Code
Paper
InstantID: Zero-Shot Identity-Preserving Generation in SecondsarXiv2024
01-15
Code
Paper
Face2Diffusion for Fast and Editable Face PersonalizationarXiv2024
03-08
โ€”
Paper
OMG: Occlusion-friendly Personalized Multi-concept Generation in Diffusion ModelsarXiv2024
03-16
Code
Paper
Infinite-ID: Identity-preserved Personalization via ID-semantics Decoupling ParadigmarXiv2024
03-18
โ€”
Paper
IDAdapter: Learning Mixed Features for Tuning-Free Personalization of Text-to-Image ModelsarXiv2024
03-21
โ€”
Paper
MoA: Mixture-of-Attention for Subject-Context Disentanglement in Personalized Image GenerationarXiv2024
04-17
โ€”
Paper
ID-Aligner: Enhancing Identity-Preserving Text-to-Image Generation with Reward Feedback LearningarXiv2024
04-23
โ€”
Paper
InstantFamily: Masked Attention for Zero-shot Multi-ID Image GenerationarXiv2024
04-30
โ€”
Paper

Personalization with Extra Condition

TitleVenueDateLinks
Training-free layout control with cross-attention guidanceWACV 20242023
04-06
Code
Paper
Prompt-Free Diffusion: Taking "Text" Out of Text-to-Image Diffusion ModelsCVPR 20242023
05-25
Code
Paper
Uni-ControlNet: All-in-One Control to Text-to-Image Diffusion ModelsNeurIPS 20232023
05-25
Code
Paper
PhotoSwap: Personalized Subject Swapping in ImagesNIPS 20232023
05-29
Code
Paper
TryonDiffusion: A Tale of Two UNetsCVPR 20232023
06-14
Code
Paper
ViscoNet: Bridging and Harmonizing Visual and Textual Conditioning for ControlNetarXiv2023
12-05
Code
Paper
Context Diffusion: In-Context Aware Image GenerationarXiv2023
12-06
Code
Paper
FreeControl: Training-Free Spatial Control of Any Text-to-Image Diffusion Model with Any ConditionCVPR 20242023
12-12
Code
Paper
A Two-Stage Personalized Virtual Try-On Framework with Shape Control and Texture GuidanceCoRR 20232023
12-24
โ€”
Paper
Tuning-Free Image Customization with Image and Text GuidancearXiv2024
03-19
โ€”
Paper
SWAPANYTHING: Enabling Arbitrary Object Swapping in Personalized Visual EditingarXiv2024
04-08
โ€”
Paper
Customizing Text-to-Image Diffusion with Camera Viewpoint ControlarXiv2024
04-18
โ€”
Paper

Personalized Video Generation

TitleVenueDateLinks
Tune-A-Video: One-Shot Tuning of Image Diffusion Models for Text-to-Video GenerationICCV 20232022
12-22
Code
Paper
Structure and Content-Guided Video Synthesis with Diffusion ModelsICCV 20232023
02-06
โ€”
Paper
Make-A-Protagonist: Generic Video Editing with Visual and Textual CluesarXiv2023
05-15
Code
Paper
Animate-A-Story: Storytelling with Retrieval-Augmented Video GenerationarXiv2023
07-13
Code
Paper
MotionDirector: Motion Customization of Text-to-Video Diffusion ModelsarXiv2023
10-12
Code
Paper
LAMP: Learn a Motion Pattern for Few-Shot-Based Video GenerationarXiv2023
10-16
Code
Paper
VideoDreamer: Customized Multi-Subject Text-to-Video Generation with Disen-Mix FinetuningarXiv2023
11-02
Code
Paper
VideoAssembler: Identity-Consistent Video Generation with Reference Entities Using Diffusion ModelarXiv2023
11-29
Code
Paper
VMC: Video Motion Customization using Temporal Attention Adaption for Text-to-Video Diffusion ModelsarXiv2023
12-01
Code
Paper
VideoBooth: Diffusion-Based Video Generation with Image PromptsCVPR 20242023
12-01
Code
Paper
StyleCrafter: Enhancing Stylized Text-to-Video Generation with Style AdapterarXiv2023
12-01
Code
Paper
SAVE: Protagonist Diversification with Structure Agnostic Video EditingarXiv2023
12-05
Code
Paper
Customizing Motion in Text-to-Video Diffusion ModelsarXiv2023
12-07
Code
Paper
DreamVideo: Composing Your Dream Videos with Customized Subject and MotionarXiv2023
12-07
Code
Paper
MotionCrafter: One-Shot Motion Customization of Diffusion ModelsarXiv2023
12-08
Code
Paper
DreaMoving: A Human Video Generation Framework Based on Diffusion ModelsarXiv2023
12-08
Code
Paper
CustomVideo: Customizing Text-to-Video Generation with Multiple SubjectsarXiv2024
01-18
Code
Paper
Magic-Me: Identity-Specific Video Customized DiffusionarXiv2024
02-14
โ€”
Paper
Customize-A-Video: One-Shot Motion Customization of Text-to-Video Diffusion ModelsarXiv2024
02-22
โ€”
Paper
ID-Animator: Zero-Shot Identity-Preserving Human Video GenerationarXiv2024
04-23
โ€”
Paper

Personalized 3D Generation

TitleVenueDateLinks
Magic3D: High-Resolution Text-to-3D Content CreationCVPR 20232022
11-18
Code
Paper
DreamBooth3D: Subject-Driven Text-to-3D GenerationICCV 20232023
03-23
Code
Paper
Text-Conditional Contextualized Avatars For Zero-Shot PersonalizationarXiv2023
04-14
โ€”
Paper
StyleAvatar3D: Leveraging Image-Text Diffusion Models for High-Fidelity 3D Avatar GenerationarXiv2023
05-30
โ€”
Paper
AvatarBooth: High-Quality and Customizable 3D Human Avatar GenerationarXiv2023
06-16
Code
Paper
MVDREAM: MULTI-VIEW DIFFUSION FOR 3D GENERATIONICLR 20242023
08-31
Code
Paper
Chasing Consistency in Text-to-3D Generation from a Single ImagearXiv2023
09-07
โ€”
Paper
Animate124: Animating One Image to 4D Dynamic ScenearXiv2023
11-24
Code
Paper
A Unified Approach for Text- and Image-guided 4D Scene GenerationarXiv2023
11-28
Code
Paper
TextureDreamer: Image-guided Texture Synthesis through Geometry-aware DiffusionarXiv2024
01-17
โ€”
Paper
TIP-Editor: An Accurate 3D Editor Following Both Text-Prompts And Image-PromptsarXiv2024
01-26
Code
Paper

Others

TitleVenueDateLinks
Anti-DreamBooth: Protecting Users from Personalized Text-to-Image SynthesisarXiv2023
03-27
Code
Paper
Backdooring Textual Inversion for Concept CensorshiparXiv2023
08-21
Code
Paper
Personalization as a Shortcut for Few-Shot Backdoor Attack against Text-to-Image Diffusion ModelsAAAI2024
03-24
โ€”
Paper
ReVersion: Diffusion-Based Relation Inversion from ImagesarXiv2023
03-23
Code
Paper
Learning Disentangled Identifiers for Action-Customized Text-to-Image GenerationarXiv2023
11-30
Code
Paper
Inv-ReVersion: Enhanced Relation Inversion Based on Text-to-Image Diffusion ModelsMDPI2024
04-15
โ€”
Paper
Continual Diffusion: Continual Customization of Text-to-Image Diffusion with C-LoRAarXiv2023
04-12
Code
Paper
Text-Guided Vector Graphics CustomizationSIGGRAPH 20232023
09-21
Code
Paper
Customizing 360-Degree Panoramas Through Text-to-Image Diffusion ModelsWACV 20242023
10-28
Code
Paper

๐Ÿ“ฎ Contact Us

If you find any missing work, please report it by creating an Issue in the repository to contribute the community together.

Contributors

zhangxulu1996

10 commits

Huenao

7 commits

sunmayuan

3 commits

zhangxulu1996/awesome-personalization

25

stars

20

commits

Python

primary language

Apr 10, 2025

updated

README

Awesome Personalization

example

This repository contains a collection of papers and resources on Personalized Content Synthesis (PCS) with Diffusion Model.

๐Ÿ”— Citation

If you find the information in our paper useful for your research, please consider citing it in your work. Thank you!

@misc{zhang2024survey,
      title={A Survey on Personalized Content Synthesis with Diffusion Models}, 
      author={Xulu Zhang and Xiao-Yong Wei and Wengyu Zhang and Jinlin Wu and Zhaoxiang Zhang and Zhen Lei and Qing Li},
      year={2024},
      eprint={2405.05538},
      archivePrefix={arXiv},
      primaryClass={cs.CV}
}

๐Ÿ“– Contents

๐Ÿ—‚๏ธ Unified Test Dataset

To uniformly evaluate Personalized Content Synthesis (PCS) tasks, we introduces a comprehensive evaluation dataset designed for the most common personalized generation tasks, object and face personalization.

Download Link: PCS-dataset

๐Ÿ”จ Evaluation

We evaluate existing representative PCS methods based our unified test dataset. The evaluation results and settings are shown in the below table. For more details, please see our survey paper.

TypeMethodsFrameworkBackboneCLIP-TCLIP-I
ObjectTextual InversionTTFSD 1.50.1990.749
DreamboothTTFSD 1.50.2860.772
P+TTFSD 1.40.2440.643
Custom DiffusionTTFSD 1.40.3070.722
NeTITTFSD 1.40.2830.801
SVDiffTTFSD 1.50.2820.776
PerfusionTTFSD 1.50.2730.691
ELITEPTASD 1.40.2920.765
BLIP-DiffusionPTASD 1.50.2920.772
IP-AdapterPTASD 1.50.2720.825
SSR EncoderPTASD 1.50.2880.792
MoMAPTASD 1.50.3220.748
Diptych PromptingPTAFLUX 1.0 dev0.3270.722
ฮป-eclipsePTAKandinsky 2.20.2720.824
MS-DiffusionPTASDXL0.2980.777
FaceCrossInitializationTTFSD 2.10.2610.469
Face2DiffusionPTASD 1.40.2650.588
SSR EncoderPTASD 1.50.2330.490
FastComposerPTASD 1.50.2300.516
IP-AdapterPTASD 1.50.2920.462
IP-AdapterPTASDXL0.2920.642
PhotoMakerPTASDXL0.3110.547
InstantIDPTASDXL0.2780.707

TTF: Test-time Fine-tuning, PTA: Pre-trained Adaptation

๐Ÿ“ƒ Paper List

Personalized Object Generation

TitleVenueDateLinks
An Image is Worth One Word: Personalizing Text-to-Image Generation using Textual InversionICLR 20232022
08-02
Code
Paper
DreamBooth: Fine-Tuning Text-to-Image Diffusion Models for Subject-Driven GenerationCVPR 20232022
08-25
Code
Paper
Re-Imagen: Retrieval-Augmented Text-to-Image GeneratorICLR 20232022
09-29
โ€”
Paper
Versatile Diffusion: Text, Images, and Variations All in One Diffusion ModelICCV 20232022
11-15
Code
Paper
DreamArtist: Towards Controllable One-Shot Text-to-Image Generation via Positive-Negative Prompt-TuningarXiv 20222022
11-21
Code
Paper
Is This Loss Informative? Faster Text-to-Image Customization by Tracking Objective DynamicsNeurIPS 20232023
02-09
Code
Paper
Encoder-Based Domain Tuning for Fast Personalization of Text-to-Image ModelsACM Trans on Graphics2023
02-23
Code
Paper
ELITE: Encoding Visual Concepts into Textual Embeddings for Customized Text-to-Image GenerationICCV 20232023
02-27
Code
Paper
Highly Personalized Text Embedding for Image Manipulation by Stable DiffusionarXiv2023
03-15
Code
Paper
Unified Multi-Modal Latent Diffusion for Joint Subject and Text Conditional Image GenerationarXiv2023
03-16
โ€”
Paper
P+: Extended Textual Conditioning in Text-to-Image GenerationarXiv2023
03-16
Code
Paper
A Closer Look at Parameter-Efficient Tuning in Diffusion ModelsarXiv2023
03-31
Code
Paper
Subject-Driven Text-to-Image Generation via Apprenticeship LearningNIPS 20232023
04-01
โ€”
Paper
Taming Encoder for Zero Fine-Tuning Image Customization with Text-to-Image Diffusion ModelsarXiv2023
04-05
โ€”
Paper
InstantBooth: Personalized Text-to-Image Generation Without Test-Time FinetuningarXiv2023
04-06
Code
Paper
Controllable Textual Inversion for Personalized Text-to-Image GenerationarXiv2023
04-11
Code
Paper
Gradient-Free Textual InversionACM MM 20232023
04-12
Code
Paper
Personalize Segment Anything Model with One ShotICLR 20242023
05-04
Code
Paper
DisenBooth: Identity-Preserving Disentangled Tuning for Subject-Driven Text-to-Image GenerationICLR 20242023
05-05
Code
Paper
BLIP-Diffusion: Pre-Trained Subject Representation for Controllable Text-to-Image Generation and EditingNIPS 20232023
05-24
Code
Paper
A Neural Space-Time Representation for Text-to-Image PersonalizationSIGGRAPH Asia 20232023
05-24
Code
Paper
Prospect: Prompt Spectrum for Attribute-Aware Personalization of Diffusion ModelsACM Trans on Graphics2023
05-25
Code
Paper
Break-a-Scene: Extracting Multiple Concepts from a Single ImageSIGGRAPH ASIA 20232023
05-25
Code
Paper
COMCAT: Towards Efficient Compression and Customization of Attention-Based Vision ModelsarXiv2023
05-26
Code
Paper
ViCo: Plug-and-Play Visual Condition for Personalized Text-to-Image GenerationarXiv2023
06-01
Code
Paper
Controlling Text-to-Image Diffusion by Orthogonal Fine-TuningarXiv2023
06-12
Code
Paper
Domain-Agnostic Tuning-Encoder for Fast Personalization of Text-to-Image ModelsSIGGRAPH 20232023
07-13
Code
Paper
IP-Adapter: Text Compatible Image Prompt Adapter for Text-to-Image Diffusion ModelsarXiv2023
08-13
Code
Paper
Navigating Text-to-Image Customization: From Lycoris Fine-Tuning to Model EvaluationICLR 20242023
09-26
Code
Paper
Kosmos-G: Generating Images in Context with Multimodal Large Language ModelsarXiv2023
10-04
Code
Paper
Personalized Text-to-Image Model Enhancement Strategies: SOD Preprocessing and CNN Local Feature IntegrationarXiv2023
10-26
โ€”
Paper
A Data Perspective on Enhanced Identity Preservation for Diffusion PersonalizationICLR 20242023
11-07
Code
Paper
DIFFNAT: Improving Diffusion Image Quality Using Natural Image StatisticsarXiv2023
11-16
โ€”
Paper
An Image is Worth Multiple Words: Multi-attribute Inversion for Constrained Text-to-Image SynthesisarXiv2023
11-20
โ€”
Paper
LEGO: Learning to Disentangle and Invert Concepts Beyond Object Appearance in Text-to-Image Diffusion ModelsarXiv2023
11-23
Code
Paper
Catversion: Concatenating Embeddings for Diffusion-Based Text-to-Image PersonalizationarXiv2023
11-24
Code
Paper
CLiC: Concept Learning in ContextCVPR 20242023
11-28
Code
Paper
HiFi Tuner: High-Fidelity Subject-Driven Fine-Tuning for Diffusion ModelsarXiv2023
11-30
โ€”
Paper
InstructBooth: Instruction-Following Personalized Text-to-Image GenerationarXiv2023
12-04
โ€”
Paper
Customization Assistant for Text-to-Image GenerationCVPR 20242023
12-05
โ€”
Paper
Decoupled Textual Embeddings for Customized Image GenerationAAAI 20242023
12-19
Code
Paper
Towards Accurate Guided Diffusion Sampling through Symplectic Adjoint MethodarXiv2023
12-19
Code
Paper
DreamDistribution: Prompt Distribution Learning for Text-to-Image Diffusion ModelsarXiv2023
12-21
Code
Paper
DreamTuner: Single Image is Enough for Subject-Driven GenerationarXiv2023
12-21
Code
Paper
BootPIG: Bootstrapping Zero-Shot Personalized Image Generation Capabilities in Pretrained Diffusion ModelsarXiv2024
01-25
โ€”
Paper
Object-Driven One-Shot Fine-Tuning of Text-to-Image Diffusion with Prototypical EmbeddingarXiv2024
01-28
โ€”
Paper
DisenDreamer: Subject-Driven Text-to-Image Generation with Sample-aware Disentangled TuningarXiv2024
02-26
โ€”
Paper
Infusion: Preventing Customized Text-to-Image Diffusion from OverfittingarXiv2024
04-22
โ€”
Paper
Customizing Text-to-Image Models with a Single Image PairarXiv2024
05-02
โ€”
Paper

Multi-concept Composition

TitleVenueDateLinks
Multi-Concept Customization of Text-to-Image DiffusionCVPR 20232022
12-08
Code
Paper
CONES: Concept Neurons in Diffusion Models for Customized GenerationICML 20232023
03-09
Code
Paper
SVDiff: Compact Parameter Space for Diffusion Fine-TuningICCV 20232023
03-20
Code
Paper
Key-Locked Rank One Editing for Text-to-Image PersonalizationSIGGRAPH 20232023
05-02
โ€”
Paper
Mix-of-Show: Decentralized Low-Rank Adaptation for Multi-Concept Customization of Diffusion ModelsNIPS 20232023
05-29
Code
Paper
CONES 2: Customizable Image Synthesis with Multiple SubjectsNIPS 20232023
05-30
Code
Paper
Generate Anything Anywhere in Any ScenearXiv2023
06-29
Code
Paper
AnyDoor: Zero-Shot Object-Level Image CustomizationarXiv2023
07-18
Code
Paper
Subject-Diffusion: Open Domain Personalized Text-to-Image Generation Without Test-Time Fine-TuningarXiv2023
07-21
Code
Paper
CustomNet: Zero-Shot Object Customization with Variable-Viewpoints in Text-to-Image Diffusion ModelsarXiv2023
10-30
Code
Paper
Compositional Inversion for Stable Diffusion ModelsAAAI 20242023
12-13
Code
Paper
Visual Concept-Driven Image Generation with Text-to-Image Diffusion ModelarXiv2024
02-18
โ€”
Paper
MIGC: Multi-Instance Generation Controller for Text-to-Image SynthesisarXiv2024
02-27
โ€”
Paper
Multi-Object Editing in Personalized Text-To-Image Diffusion Model Via Segmentation GuidancearXiv2024
03-18
โ€”
Paper
MC2: Multi-concept Guidance for Customized Multi-concept GenerationarXiv2024
04-12
โ€”
Paper
MultiBooth: Towards Generating All Your Concepts in an Image from TextarXiv2024
04-22
โ€”
Paper
MagicTailor: Component-Controllable Personalization in Text-to-Image Diffusion ModelsarXiv2024
10-06
Code
Paper

Personalized Style Generation

TitleVenueDateLinks
StyleDrop: Text-to-Image Synthesis of Any StyleNIPS 20232023
06-01
Code
Paper
StyleAdapter: A Single-Pass LoRA-Free Model for Stylized Image GenerationICLR 20242023
09-04
โ€”
Paper
StyleBoost: A Study of Personalizing Text-to-Image Generation in Any Style using DreamBoothICTC 20232023
10-13
โ€”
Paper
ArtAdapter: Text-to-Image Style Transfer Using Multi-Level Style Encoder and Explicit AdaptationarXiv2023
12-04
Code
Paper
Style Aligned Image Generation via Shared AttentionCVPR 20242023
12-04
Code
Paper
Generative Active Learning for Image Synthesis PersonalizationarXiv2024
03-22
Code
Paper
Text-to-Image Synthesis for Any Artistic Styles: Advancements in Personalized Artistic Image Generation via Subdivision and Dual BindingarXiv2024
04-08
โ€”
Paper

Personalized Face Generation

TitleVenueDateLinks
Identity Encoder for Personalized DiffusionCoRR 20232023
04-14
โ€”
Paper
FastComposer: Tuning-Free Multi-Subject Image Generation with Localized AttentionCoRR 20232023
05-21
Code
Paper
Enhancing Detail Preservation for Customized Text-to-Image Generation: A Regularization-Free ApproacharXiv2023
05-23
โ€”
Paper
Inserting Anybody in Diffusion Models via Celeb BasisNIPS 20232023
06-01
Code
Paper
Face0: Instantaneously Conditioning a Text-to-Image Model on a FaceSIGGRAPH 20232023
06-11
โ€”
Paper
DreamIdentity: Improved Editability for Efficient Face-Identity Preserved Image GenerationarXiv2023
07-01
Code
Paper
HyperDreamBooth: Hypernetworks for Fast Personalization of Text-to-Image ModelsarXiv2023
07-13
Code
Paper
Identity-Preserving Aging of Face Images via Latent Diffusion ModelsIJCB 20232023
07-17
Code
Paper
Magicapture: High-Resolution Multi-Concept Portrait CustomizationarXiv2023
09-13
Code
Paper
High-Fidelity Person-Centric Subject-to-Image SynthesisCVPR 20242023
11-17
โ€”
Paper
When StyleGAN Meets Stable Diffusion: A W+ Adapter for Personalized Image GenerationarXiv2023
11-29
Code
Paper
Portrait Diffusion: Training-Free Face Stylization with Chain-of-PaintingarXiv2023
12-03
Code
Paper
Retrieving Conditions from Reference Images for Diffusion ModelsarXiv2023
12-05
โ€”
Paper
FaceStudio: Put Your Face Everywhere in SecondsarXiv2023
12-05
Code
Paper
Personalized Face Inpainting with Diffusion Models by Parallel Visual AttentionWACV 20242023
12-06
Code
Paper
PhotoMaker: Customizing Realistic Human Photos via Stacked ID EmbeddingarXiv2023
12-07
Code
Paper
DemoCaricature: Democratising Caricature Generation with a Rough SketchCVPR 20242023
12-07
Code
Paper
Stellar: Systematic Evaluation of Human-Centric Personalized Text-to-Image MethodsCoRR 20232023
12-11
Code
Paper
PortraitBooth: A Versatile Portrait Model for Fast Identity-Preserved PersonalizationCVPR 20242023
12-11
Code
Paper
Concept-Centric Personalization with Large-Scale Diffusion PriorsarXiv2023
12-13
Code
Paper
Cross Initialization for Personalized Text-to-Image GenerationarXiv2023
12-26
Code
Paper
InstantID: Zero-Shot Identity-Preserving Generation in SecondsarXiv2024
01-15
Code
Paper
Face2Diffusion for Fast and Editable Face PersonalizationarXiv2024
03-08
โ€”
Paper
OMG: Occlusion-friendly Personalized Multi-concept Generation in Diffusion ModelsarXiv2024
03-16
Code
Paper
Infinite-ID: Identity-preserved Personalization via ID-semantics Decoupling ParadigmarXiv2024
03-18
โ€”
Paper
IDAdapter: Learning Mixed Features for Tuning-Free Personalization of Text-to-Image ModelsarXiv2024
03-21
โ€”
Paper
MoA: Mixture-of-Attention for Subject-Context Disentanglement in Personalized Image GenerationarXiv2024
04-17
โ€”
Paper
ID-Aligner: Enhancing Identity-Preserving Text-to-Image Generation with Reward Feedback LearningarXiv2024
04-23
โ€”
Paper
InstantFamily: Masked Attention for Zero-shot Multi-ID Image GenerationarXiv2024
04-30
โ€”
Paper

Personalization with Extra Condition

TitleVenueDateLinks
Training-free layout control with cross-attention guidanceWACV 20242023
04-06
Code
Paper
Prompt-Free Diffusion: Taking "Text" Out of Text-to-Image Diffusion ModelsCVPR 20242023
05-25
Code
Paper
Uni-ControlNet: All-in-One Control to Text-to-Image Diffusion ModelsNeurIPS 20232023
05-25
Code
Paper
PhotoSwap: Personalized Subject Swapping in ImagesNIPS 20232023
05-29
Code
Paper
TryonDiffusion: A Tale of Two UNetsCVPR 20232023
06-14
Code
Paper
ViscoNet: Bridging and Harmonizing Visual and Textual Conditioning for ControlNetarXiv2023
12-05
Code
Paper
Context Diffusion: In-Context Aware Image GenerationarXiv2023
12-06
Code
Paper
FreeControl: Training-Free Spatial Control of Any Text-to-Image Diffusion Model with Any ConditionCVPR 20242023
12-12
Code
Paper
A Two-Stage Personalized Virtual Try-On Framework with Shape Control and Texture GuidanceCoRR 20232023
12-24
โ€”
Paper
Tuning-Free Image Customization with Image and Text GuidancearXiv2024
03-19
โ€”
Paper
SWAPANYTHING: Enabling Arbitrary Object Swapping in Personalized Visual EditingarXiv2024
04-08
โ€”
Paper
Customizing Text-to-Image Diffusion with Camera Viewpoint ControlarXiv2024
04-18
โ€”
Paper

Personalized Video Generation

TitleVenueDateLinks
Tune-A-Video: One-Shot Tuning of Image Diffusion Models for Text-to-Video GenerationICCV 20232022
12-22
Code
Paper
Structure and Content-Guided Video Synthesis with Diffusion ModelsICCV 20232023
02-06
โ€”
Paper
Make-A-Protagonist: Generic Video Editing with Visual and Textual CluesarXiv2023
05-15
Code
Paper
Animate-A-Story: Storytelling with Retrieval-Augmented Video GenerationarXiv2023
07-13
Code
Paper
MotionDirector: Motion Customization of Text-to-Video Diffusion ModelsarXiv2023
10-12
Code
Paper
LAMP: Learn a Motion Pattern for Few-Shot-Based Video GenerationarXiv2023
10-16
Code
Paper
VideoDreamer: Customized Multi-Subject Text-to-Video Generation with Disen-Mix FinetuningarXiv2023
11-02
Code
Paper
VideoAssembler: Identity-Consistent Video Generation with Reference Entities Using Diffusion ModelarXiv2023
11-29
Code
Paper
VMC: Video Motion Customization using Temporal Attention Adaption for Text-to-Video Diffusion ModelsarXiv2023
12-01
Code
Paper
VideoBooth: Diffusion-Based Video Generation with Image PromptsCVPR 20242023
12-01
Code
Paper
StyleCrafter: Enhancing Stylized Text-to-Video Generation with Style AdapterarXiv2023
12-01
Code
Paper
SAVE: Protagonist Diversification with Structure Agnostic Video EditingarXiv2023
12-05
Code
Paper
Customizing Motion in Text-to-Video Diffusion ModelsarXiv2023
12-07
Code
Paper
DreamVideo: Composing Your Dream Videos with Customized Subject and MotionarXiv2023
12-07
Code
Paper
MotionCrafter: One-Shot Motion Customization of Diffusion ModelsarXiv2023
12-08
Code
Paper
DreaMoving: A Human Video Generation Framework Based on Diffusion ModelsarXiv2023
12-08
Code
Paper
CustomVideo: Customizing Text-to-Video Generation with Multiple SubjectsarXiv2024
01-18
Code
Paper
Magic-Me: Identity-Specific Video Customized DiffusionarXiv2024
02-14
โ€”
Paper
Customize-A-Video: One-Shot Motion Customization of Text-to-Video Diffusion ModelsarXiv2024
02-22
โ€”
Paper
ID-Animator: Zero-Shot Identity-Preserving Human Video GenerationarXiv2024
04-23
โ€”
Paper

Personalized 3D Generation

TitleVenueDateLinks
Magic3D: High-Resolution Text-to-3D Content CreationCVPR 20232022
11-18
Code
Paper
DreamBooth3D: Subject-Driven Text-to-3D GenerationICCV 20232023
03-23
Code
Paper
Text-Conditional Contextualized Avatars For Zero-Shot PersonalizationarXiv2023
04-14
โ€”
Paper
StyleAvatar3D: Leveraging Image-Text Diffusion Models for High-Fidelity 3D Avatar GenerationarXiv2023
05-30
โ€”
Paper
AvatarBooth: High-Quality and Customizable 3D Human Avatar GenerationarXiv2023
06-16
Code
Paper
MVDREAM: MULTI-VIEW DIFFUSION FOR 3D GENERATIONICLR 20242023
08-31
Code
Paper
Chasing Consistency in Text-to-3D Generation from a Single ImagearXiv2023
09-07
โ€”
Paper
Animate124: Animating One Image to 4D Dynamic ScenearXiv2023
11-24
Code
Paper
A Unified Approach for Text- and Image-guided 4D Scene GenerationarXiv2023
11-28
Code
Paper
TextureDreamer: Image-guided Texture Synthesis through Geometry-aware DiffusionarXiv2024
01-17
โ€”
Paper
TIP-Editor: An Accurate 3D Editor Following Both Text-Prompts And Image-PromptsarXiv2024
01-26
Code
Paper

Others

TitleVenueDateLinks
Anti-DreamBooth: Protecting Users from Personalized Text-to-Image SynthesisarXiv2023
03-27
Code
Paper
Backdooring Textual Inversion for Concept CensorshiparXiv2023
08-21
Code
Paper
Personalization as a Shortcut for Few-Shot Backdoor Attack against Text-to-Image Diffusion ModelsAAAI2024
03-24
โ€”
Paper
ReVersion: Diffusion-Based Relation Inversion from ImagesarXiv2023
03-23
Code
Paper
Learning Disentangled Identifiers for Action-Customized Text-to-Image GenerationarXiv2023
11-30
Code
Paper
Inv-ReVersion: Enhanced Relation Inversion Based on Text-to-Image Diffusion ModelsMDPI2024
04-15
โ€”
Paper
Continual Diffusion: Continual Customization of Text-to-Image Diffusion with C-LoRAarXiv2023
04-12
Code
Paper
Text-Guided Vector Graphics CustomizationSIGGRAPH 20232023
09-21
Code
Paper
Customizing 360-Degree Panoramas Through Text-to-Image Diffusion ModelsWACV 20242023
10-28
Code
Paper

๐Ÿ“ฎ Contact Us

If you find any missing work, please report it by creating an Issue in the repository to contribute the community together.

Contributors

zhangxulu1996

10 commits

Huenao

7 commits

sunmayuan

3 commits

Languages

Python

75.1%

Jupyter Notebook

23.6%

Shell

1.3%