harlanhong/awesome-talking-head-generation

1,933

123 commits

updated Apr 27, 2026

See the code

README

Awesome Talking Head Generation Awesome Last Updated

A curated list of papers and resources for Talking Head Generation, including face animation, audio-driven synthesis, portrait animation, and related tasks.

Contributions welcome! Open an issue, pull request, or contact fatinghong@gmail.com. Discord: Fa-Ting Hong#6563

🔍 I'm actively looking for remote full-time positions and full-time postdoc opportunities in talking head generation, generative models, or related areas. If you're interested, feel free to reach out at fatinghong@gmail.com.

:fire: New: ACTalker — portrait video generation driven by audio and expression simultaneously. ICCV 2025

Table of Contents

Datasets

  1. VoxCeleb1 [Download link].
  2. VoxCeleb2 [Download link].
  3. Faceforensics++ [Download link].
  4. CelebV [Download link].
  5. TalkingHead-1KH [Download link].
  6. LRW (Lip Reading in the Wild) [Download link].
  7. MEAD [Download link].
  8. CelebV-HQ [Download link].
  9. CHDTF [Download link].
  10. MultiTalk [Download link].
  11. VFHQ [Download link].
  12. Hallo3 [Download link].
  13. AVSpeech [Download link].

Survey


Image-driven

YearPaperVenueLinks
2026PersonaLive! Expressive Portrait Image Animation for Live StreamingCVPR 2026Code
2026PortraitDirector: A Hierarchical Disentanglement Framework for Controllable and Real-time Facial ReenactmentarXiv 2026
2025HunyuanPortrait: Implicit Condition Control for Enhanced Portrait AnimationCVPR 2025Code · Project
2025Robust Deepfake Detection for Electronic Know Your Customer Systems Using Registered ImagesarXiv 2025
2025Towards Interactive Intelligence for Digital HumansarXiv 2025
2025FlashPortrait: 6x Faster Infinite Portrait Animation with Adaptive Latent PredictionarXiv 2025
2024X-Portrait: Expressive Portrait Animation with Hierarchical Motion AttentionSIGGRAPH 2024Code
2024Follow-Your-Emoji: Fine-Controllable and Expressive Freestyle Portrait AnimationSIGGRAPH Asia 2024Code
2024LivePortrait: Efficient Portrait Animation with Stitching and Retargeting ControlarXiv 2024Code · Project
2024EMOPortraits: Emotion-enhanced Multimodal One-shot Head AvatarsCVPR 2024Code · Project
2024Synergizing Motion and Appearance: Multi-Scale Compensatory Codebooks for Talking Head Video GenerationCVPR 2024Project
2023Audio-Visual Face ReenactmentWACV 2023Code · Project
2023Cross-identity Video Motion Retargeting with Joint Transformation and SynthesisWACV 2023Code
2023Implicit Identity Representation Conditioned Memory Compensation Network for Talking Head Video GenerationICCV 2023Project · Code
2023StyleLipSync: Style-based Personalized Lip-sync Video GenerationICCV 2023Code
2022Depth-Aware Generative Adversarial Network for Talking Head Video GenerationCVPR 2022Code · Project
2022Thin-Plate Spline Motion Model for Image AnimationCVPR 2022Code
2022StyleHEAT: One-Shot High-Resolution Editable Talking Face Generation via Pretrained StyleGANECCV 2022Code · Project
2022MegaPortraits: One-shot Megapixel Neural Head AvatarsACM MM 2022Project
2022Structure-Aware Motion Transfer with Deformable Anchor ModelCVPR 2022Code
2022StyleMask: Disentangling the Style Space of StyleGAN2 for Neural Face ReenactmentFG, 2023Code
2022Controllable Radiance Fields for Dynamic Face SynthesisArxiv 2022
2022Animatable 3D-Aware Face Image Generation for Video AvatarsNeurIPS 2022Project
2022Implicit Warping for Animation with Image SetsNeurIPS 2022Project
2022HifiHead: One-Shot High Fidelity Neural Head Synthesis with 3D ControlIJCAI 2022
2022Face Animation with Multiple Source ImagesArxiv 2022
2022MetaPortrait: Identity-Preserving Talking Head Generation with Fast Personalized AdaptationArxiv 2022
2022Compressing Video Calls using Synthetic Talking HeadsBMVC 2022Project
2022Finding Directions in GAN’s Latent Space for Neural Face ReenactmentBMVC 2022Project · Code
2022Latent Image Animator: Learning to Animate Images via Latent Space NavigationICLR 2022Project · Code
2021One-Shot Free-View Neural Talking-Head Synthesis for Video ConferencingCVPR 2021 OralProject
2021Sparse to Dense Motion Transfer for Face Image AnimationICCV 2021
2021SAFA: Structure Aware Face Animation3DV 2021Code
2021Self-appearance-aided Differential Evolution for Motion TransferarXiv 2021
2021PIRenderer: Controllable Portrait Image Generation via Semantic Neural RenderingICCV 2021Code
2021FACEGAN: Facial Attribute Controllable rEenactment GANWACV 2021
2021F3A-GAN: Facial Flow for Face Animation With Generative Adversarial NetworksIEEE TIP 2021
2021FACIAL: Synthesizing Dynamic Talking Face with Implicit Attribute LearningICCV 2021
2021 Motion Representations for Articulated AnimationCVPR 2021Code
2021HeadGAN: One-shot Neural Head Synthesis and EditingICCV 2021Project
2020Mesh Guided One-shot Face Reenactment Using Graph Convolutional NetworksACM Multimedia 2020Code
2020MarioNETte: Few-shot Face Reenactment Preserving Identity of Unseen TargetsAAAI 2020Project
2020Learning Identity-Invariant Motion Representations for Cross-ID Face ReenactmentCVPR 2020
2019First order motion model for image animationNeurIPS 2019Code
2019Few-Shot Adversarial Learning of Realistic Neural Talking Head ModelsICCV 2019Code
2019Animating Arbitrary Objects via Deep Motion TransferCVPR 2019 OralCode · Project
2019Few-shot Video-to-Video SynthesisNeurIPS 2019Code · Project
2018ReenactGAN: Learning to Reenact Faces via Boundary TransferECCV 2018Code
2018X2Face: A network for controlling face generation by using images, audio, and pose codesECCV 2018Code · Project
2016Face2Face: Real-time face capture and reenactment of RGB videosCVPR 2016

Audio-driven

YearPaperVenueLinks
2026FunCineForge: A Unified Dataset Toolkit and Model for Zero-Shot Movie Dubbing in Diverse Cinematic ScenesarXiv 2026
2026TurboTalk: Progressive Distillation for One-Step Audio-Driven Talking Avatar GenerationarXiv 2026
2026SEDTalker: Emotion-Aware 3D Facial Animation Using Frame-Level Speech Emotion DiarizationarXiv 2026Code
2026AUHead: Realistic Emotional Talking Head Generation via Action Units ControlICLR 2026
2026UniTalking: A Unified Audio-Video Framework for Talking Portrait GenerationCVPR 2026
2026DiFlowDubber: Discrete Flow Matching for Automated Video Dubbing via Cross-Modal Alignment and SynchronizationCVPR 2026
2026ActAvatar: Temporally-Aware Precise Action Control for Talking AvatarsCVPR 2026
2026Cross-Modal Emotion Transfer for Emotion Editing in Talking Face VideoCVPR 2026
2026MMFace-DiT: A Dual-Stream Diffusion Transformer for High-Fidelity Multimodal Face GenerationCVPR 2026
2025OmniHuman-1: Rethinking the Scaling-Up of One-Stage Conditioned Human Animation ModelsarXiv 2025Project
2025Audio-visual Controlled Video Diffusion with Masked Selective State Spaces Modelling for Natural Talking Head GenerationICCV 2025Project
2025OmniAvatar: Efficient Audio-Driven Avatar Video Generation with Adaptive Body AnimationarXiv 2025Code · Project
2025Teller: Real-Time Streaming Audio-Driven Portrait Animation with Autoregressive Motion GenerationCVPR 2025
2025EmotiveTalk: Expressive Talking Head Generation through Audio Information Decoupling and Emotional Video DiffusionCVPR 2025Project
2025INFP: Audio-Driven Interactive Head Generation in Dyadic ConversationsCVPR 2025Project
2025Hallo2: Long-Duration and High-Resolution Audio-Driven Portrait Image AnimationICLR 2025Code
2025Loopy: Taming Audio-Driven Portrait Avatar with Long-Term Motion DependencyICLR 2025Project
2025DAWN: Dynamic Frame Avatar with Non-autoregressive Diffusion Framework for Talking Head Video GenerationICLR 2025Project · Code
2025AnyTalk: Multi-modal Driven Multi-domain Talking Head GenerationAAAI 2025
2025Occlusion-Insensitive Talking Head Video Generation via Facelet CompensationAAAI 2025
2025FixTalk: Taming Identity Leakage for High-Quality Talking Head GenerationICCV 2025
2025FLOAT: Generative Motion Latent Flow Matching for Audio-driven Talking PortraitICCV 2025Project
2025MoEE: Mixture of Emotion Experts for Audio-Driven Portrait AnimationCVPR 2025
2025Live Avatar: Streaming Real-time Audio-Driven Avatar Generation with Infinite LengtharXiv 2025Code
2025DreamTalk: When Expressive Talking Head Generation Meets Diffusion Probabilistic ModelsarXiv 2025
2025GAIA: Zero-shot Talking Avatar GenerationarXiv 2025
2024Real3D-Portrait: One-shot Realistic 3D Talking Portrait SynthesisICLR 2024Project · Code
2024Emote Portrait Alive - Generating Expressive Portrait Videos with Audio2Video Diffusion Model under Weak ConditionsarXiv 2024Project · Code
2024Style2Talker: High-Resolution Talking Head Generation with Emotion Style and Art StyleAAAI 2024
2024Say Anything with Any StyleAAAI 2024
2024[MuseTalk] Real-Time High Quality Lip Synchorization with Latent Space Inpainting, [Code].Code
2024VASA-1: Lifelike Audio-Driven Talking Faces Generated in Real TimeNeurIPS 2024Project
2024THQA: A Perceptual Quality Assessment Database for Talking HeadsarXiv 2024Code
2024Talk3D: High-Fidelity Talking Portrait Synthesis via Personalized 3D Generative PriorarXiv 2024Code · Project
2024EDTalk: Efficient Disentanglement for Emotional Talking Head SynthesisarXiv 2024Code · Project
2024AniPortrait: Audio-Driven Synthesis of Photorealistic Portrait AnimationsarXiv 2024Code
2024FlowVQTalker: High-Quality Emotional Talking Face Generation through Normalizing Flow and QuantizationarXiv 2024
2024FaceChain-ImagineID: Freely Crafting High-Fidelity Diverse Talking Faces from Disentangled AudioarXiv 2024Code
2024Hallo: Hierarchical Audio-Driven Visual Synthesis for Portrait Image AnimationarXiv 2024Code
2024EchoMimic: Lifelike Audio-Driven Portrait Animations through Editable Landmark ConditionsarXiv 2024Code · Project
2024RealTalk: Real-time and Realistic Audio-driven Face Generation with 3D Facial Prior-guided Identity Alignment NetworkarXiv 2024
2024Emotional Conversation: Empowering Talking Faces with Cohesive Expression, Gaze and Pose GenerationarXiv 2024
2024Make Your Actor Talk: Generalizable and High-Fidelity Lip Sync with Motion and Appearance DisentanglementarXiv 2024
2024FD2Talk: Towards Generalized Talking Head Generation with Facial Decoupled Diffusion ModelarXiv 2024
2024ReSyncer: Rewiring Style-based Generator for Unified Audio-Visually Synced Facial PerformerarXiv 2024
2024Style-Preserving Lip Sync via Audio-Aware Style ReferencearXiv 2024
2024EchoMimicV2: Towards Striking, Simplified, and Semi-Body Human AnimationarXiv 2024Code · Project
2024Latent Diffusion Transformer for Talking Video SynthesisarXiv 2024Code · Project
2024IF-MDM: Implicit Face Motion Diffusion Model for High-Fidelity Realtime Talking Head GenerationarXiv 2024Project
2024Memory-Guided Diffusion for Expressive Talking Video GenerationarXiv 2024Project · Code
2024Highly Dynamic and Realistic Portrait Image Animation with Diffusion Transformer NetworksarXiv 2024
2024VQTalker: Towards Multilingual Talking Avatars through Facial Motion TokenizationarXiv 2024
2024Towards Customizable One-Shot Audio-to-Talking Face GenerationarXiv 2024
2024LatentSync: Audio Conditioned Latent Diffusion Models for Lip SyncarXiv 2024Code
2024Media2Face: Co-speech Facial Animation Generation with Multi-Modality GuidanceSIGGRAPH 2024
2024PersonaTalk: Bring Attention to Your Persona in Visual DubbingSIGGRAPH Asia 2024
2024StyleTalk++: A Unified Framework for Controlling the Speaking Styles of Talking HeadsTPAMI 2024
2024JEAN: Joint Expression and Audio-guided NeRF-based Talking Face GenerationBMVC 2024
2024Ditto: Motion-Space Diffusion for Controllable Realtime Talking Head SynthesisarXiv 2024
2024JoyVASA: Portrait and Animal Image Animation with Diffusion-Based Audio-Driven Facial DynamicsarXiv 2024Code
2024HelloMeme: Integrating Spatial Knitting Attentions to Embed High-Level Conditions in Diffusion ModelsarXiv 2024Code
2024LaDTalk: Latent Denoising for Synthesizing Talking Head Videos with High Frequency DetailsarXiv 2024
2023Diffused Heads: Diffusion Models Beat GANs on Talking-Face GenerationArxiv 2023Project
2023DiffTalk: Crafting Diffusion Models for Generalized Talking Head SynthesisArxiv 2023Project · Code
2023[READ Avatars: Realistic Emotion-controllable Audio Driven Avatars](READ Avatars: Realistic Emotion-controllable Audio Driven Avatars)Arxiv 2023
2023DAE-Talker: High Fidelity Speech-Driven Talking Face Generation with Diffusion AutoencoderArxiv 2023
2023Emotionally Enhanced Talking Face GenerationArxiv 2023Code
2023Seeing What You Said: Talking Face Generation Guided by a Lip Reading ExpertCVPR 2023Code
2023StyleSync: High-Fidelity Generalized and Personalized Lip Sync in Style-based GeneratorCVPR 2023Project · Code
2023GeneFace++: Generalized and Stable Real-Time Audio-Driven 3D Talking Face GenerationarXiv 2023Project · Code
2023MODA: Mapping-Once Audio-driven Portrait Animation with Dual AttentionsICCV 2023
2023VividTalk: One-Shot Audio-Driven Talking Head Generation Based on 3D Hybrid PriorArxiv 2023Project · Code
2023IP_LAP: Identity-Preserving Talking Face Generation with Landmark and Appearance PriorsCVPR 2023Code
2023HyperLips: Hyper Control Lips with High Resolution Decoder for Talking Face GenerationCVPR 2023Code
2023Efficient Emotional Adaptation for Audio-Driven Talking-Head GenerationICCV 2023Project · Code
2023SadTalker: Learning Realistic 3D Motion Coefficients for Stylized Audio-Driven Talking Head AnimationCVPR 2023Project · Code
2023DINet: Deformation Inpainting Network for Realistic Face Visually Dubbing on High Resolution VideoAAAI 2023Code
2023EMMN: Emotional Motion Memory Network for Audio-driven Emotional Talking Face GenerationICCV 2023
2023ToonTalker: Cross-Domain Face ReenactmentICCV 2023
2023High-fidelity Generalized Emotional Talking Face Generation with Multi-modal Emotion Space LearningCVPR 2023
2023DisCoHead: Audio-and-Video-Driven Talking Head Generation by Disentangled Control of Head Pose and Facial ExpressionsICASSP 2023Code
2022Expressive Talking Head Generation with Granular Audio-Visual Control CVPR 2022
2022Talking Face Generation with Multilingual TTSCVPR 2022Demo
2022EAMM: One-Shot Emotional Talking Face via Audio-Based Emotion-Aware Motion ModelSIGGRAPH 2022
2022SPACEx 🚀: Speech-driven Portrait Animation with Controllable ExpressionarXiv 2022Project
2022Masked Lip-Sync Prediction by Audio-Visual Contextual Exploitation in TransformersSIGGRAPH Asia 2022
2022Memories are One-to-Many Mapping Alleviators in Talking Face GenerationarXiv 2022
2021Pose-Controllable Talking Face Generation by Implicitly Modularized Audio-Visual RepresentationCVPR 2021Code · Project
2021Imitating Arbitrary Talking Style for Realistic Audio-Driven Talking Face SynthesisACM Multimedia 2021
2021Audio-Driven Emotional Video PortraitsCVPR 2021Code
2021Talking Head Generation with Audio and Speech Related Facial Action Unitsarxiv 2021
2021Speech2Talking-Face: Inferring and Driving a Face with Synchronized Audio-Visual RepresentationIJCAI 2021
2021Imitating Arbitrary Talking Style for Realistic Audio-Driven Talking Face SynthesisACM MM 2021
2021Live Speech Portraits: Real-Time Photorealistic Talking-Head AnimationACM TOG 2021Code
2021Audio2head: Audio-driven one-shot talking-head generation with natural head motionArXiv 2021
2020A Lip Sync Expert Is All You Need for Speech to Lip Generation In The WildACM Multimedia 2020Code · Project
2020Talking-head Generation with Rhythmic Head MotionECCV 2020Code
2020MakeItTalk: Speaker-Aware Talking-Head AnimationSIGGRAPH Asia 2020Code · Project
2020Neural Voice Puppetry: Audio-driven Facial ReenactmentECCV 2020Code · Project
2020MEAD: A Large-scale Audio-visual Dataset for Emotional Talking-face GenerationECCV 2020Code · Project
2020Realistic Speech-Driven Facial Animation with GANsIJCV 2020
2019Talking Face Generation by Adversarially Disentangled Audio-Visual RepresentationAAAI 2019Code
2019Hierarchical Cross-modal Talking Face Generation with Dynamic Pixel-wise LossCVPR 2019Code
2018Lip Movements Generation at a GlanceECCV 2018Code
2018VisemeNet: Audio-Driven Animator-Centric Speech AnimationSIGGRAPH 2018
2017Synthesizing Obama: Learning Lip Sync From AudioSIGGRAPH 2017Project
2017You Said That?: Synthesising Talking Faces From AudioIJCV 2019Code
2017Audio-Driven Facial Animation by Joint End-to-End Learning of Pose and EmotionSIGGRAPH 2017
2017A Deep Learning Approach for Generalized Speech AnimationSIGGRAPH 2017
2016Lip Reading in the WildACCV 2016

Nerf & 3D

YearPaperVenueLinks
2026MMTalker: Multiresolution 3D Talking Head Synthesis with Multimodal Feature FusionarXiv 2026
2026EmoTaG: Emotion-Aware Talking Head Synthesis on Gaussian Splatting with Few-Shot PersonalizationCVPR 2026
2026GenFaceTalk: Generalizable One-Shot Talking-Head Generation for Diverse StylesICLR 2026
2026FG-Portrait: 3D Flow Guided Editable Portrait AnimationCVPR 2026
2026Giving Faces Their Feelings Back: Explicit Emotion Control for Feedforward Single-Image 3D Head AvatarsarXiv 2026
20263DRealHead: Few-Shot Detailed Head AvatararXiv 2026
2026FHAvatar: Fast and High-Fidelity Reconstruction of Face-and-Hair Composable 3D Head Avatar from Few Casual CapturesarXiv 2026
2026NBAvatar: Neural Billboards Avatars with Realistic Hand-Face InteractionarXiv 2026
2026OMG-Avatar: One-shot Multi-LOD Gaussian Head AvatararXiv 2026
2026OFERA: Blendshape-driven 3D Gaussian Control for Occluded Facial Expression to Realistic Avatars in VRarXiv 2026
2026CAG-Avatar: Cross-Attention Guided Gaussian Avatars for High-Fidelity Head ReconstructionarXiv 2026
2026Uncertainty-Aware 3D Emotional Talking Face Synthesis with Emotion Prior DistillationarXiv 2026
20263DXTalker: Unifying Identity, Lip Sync, Emotion, and Spatial Dynamics in Expressive 3D Talking AvatarsarXiv 2026
2025IM-Portrait: Learning 3D-aware Video Diffusion for Photorealistic Talking Heads from Monocular VideosCVPR 2025
2025Perceptually Accurate 3D Talking Head Generation: New Definitions, Speech-Mesh Representation, and Evaluation MetricsCVPR 2025
2025GaussianSpeech: Audio-Driven Personalized 3D Gaussian AvatarsICCV 2025Code
2025MemoryTalker: Personalized Speech-Driven 3D Facial Animation via Audio-Guided StylizationICCV 2025Code
2025CAFE-TALK: Generating 3D Talking FaceICLR 2025
2025InsTaG: Learning Personalized 3D Talking Head from Few-Second VideoCVPR 2025Code
2025Monocular and Generalizable Gaussian Talking Head AnimationCVPR 2025Project
2025DualTalk: Dual-Speaker Interaction for 3D Talking Head ConversationsCVPR 2025Code
2025VASA-3D: Lifelike Audio-Driven Gaussian Head Avatars from a Single ImageNeurIPS 2025
2025PointTalk: Audio-Driven Dynamic Lip Point Cloud for 3D Gaussian-based Talking Head SynthesisAAAI 2025
2025PTalker: Personalized Speech-Driven 3D Talking Head Animation via Style DisentanglementarXiv 2025
2025SynergyWarpNet: Attention-Guided Cooperative Warping for Neural Portrait AnimationarXiv 2025
2025From Autoencoders to CycleGAN: Robust Unpaired Face Manipulation via Adversarial LearningarXiv 2025
2024CVTHead: One-shot Controllable Head Avatar with Vertex-feature TransformerWACV 2024Code
20243D-Aware Talking-Head Video Motion TransferWACV 2024
2024TalkingGaussian: Structure-Persistent 3D Talking Head Synthesis via Gaussian SplattingECCV 2024Code
2024GaussianTalker: Real-Time High-Fidelity Talking Head Synthesis with Audio-Driven 3D Gaussian SplattingACM MM 2024Code
2024Generalizable and Animatable Gaussian Head AvatarNeurIPS 2024Code
2024MimicTalk: Mimicking a Personalized and Expressive 3D Talking Face in MinutesNeurIPS 2024
2024EmoTalk3D: High-Fidelity Free-View Synthesis of Emotional 3D Talking HeadECCV 2024
2024Learning Dynamic Tetrahedra for High-Quality Talking Head SynthesisCVPR 2024Code
2024GPAvatar: Generalizable and Precise Head Avatar from Image(s)ICLR 2024Code
2024UniTalker: Scaling up Audio-Driven 3D Facial Animation through A Unified ModelarXiv 2024Code
2023Efficient Region-Aware Neural Radiance Fields for High-Fidelity Talking Portrait SynthesisICCV 2023Code
2023EmoTalk: Speech-Driven Emotional Disentanglement for 3D Face AnimationICCV 2023Code
2023CodeTalker: Speech-Driven 3D Facial Animation with Discrete Motion PriorCVPR 2023Code
2023GANHead: Towards Generative Animatable Neural Head AvatarsCVPR 2023Code
2023OTAvatar: One-shot Talking Face Avatar with Controllable Tri-plane RenderingCVPR 2023Code
2023GeneFace: Generalized and High-Fidelity Audio-Driven 3D Talking Face SynthesisICLR 2023Code
2023SyncTalk: The Devil is in the Synchronization for Talking Head SynthesisCVPR 2024Code
2022Semantic-Aware Implicit Neural Audio-Driven Video Portrait Generationarxiv, 2022
2022HeadNeRF: A Real-time NeRF-based Parametric Head ModelCVPR 2022Code · Project
2022I M Avatar: Implicit Morphable Head Avatars from VideosCVPR 2022Code
2022Realistic One-shot Mesh-based Head AvatarsECCV 2022
2022FNeVR: Neural Volume Rendering for Face AnimationArxiv 2022Code
20223DFaceShop: Explicitly Controllable 3D-Aware Portrait GenerationArxiv 2022Code · Project
2022Generative Neural Texture Rasterization for 3D-Aware Head AvatarsArxiv 2022Project
2022NeRFInvertor: High Fidelity NeRF-GAN Inversion for Single-shot Real Image AnimationArxiv 2022
2022Learning Dynamic Facial Radiance Fields for Few-Shot Talking Head SynthesisECCV 2022Code
2021DFA-NeRF: Personalized Talking Head Generation via Disentangled Face Attributes Neural Renderingarxiv, 2021
2021NerFACE: Dynamic Neural Radiance Fields for Monocular 4D Facial Avatar ReconstructionCVPR 2021 OralCode · Project
2021AD-NeRF: Audio Driven Neural Radiance Fields for Talking Head SynthesisICCV 2021Code · Code
2020Disentangled and Controllable Face Image Generation via 3D Imitative-Contrastive LearningCVPR 2020 OralCode

Star History

Star History Chart

face-reenactment
image-animation
motion-transfer
talking-head

Contributors

harlanhong

87 commits

ykk648

22 commits

kkakkkka

3 commits

yuangan

2 commits

harlanhong/awesome-talking-head-generation

1,933

123 commits

updated Apr 27, 2026

See the code

README

Awesome Talking Head Generation Awesome Last Updated

A curated list of papers and resources for Talking Head Generation, including face animation, audio-driven synthesis, portrait animation, and related tasks.

Contributions welcome! Open an issue, pull request, or contact fatinghong@gmail.com. Discord: Fa-Ting Hong#6563

🔍 I'm actively looking for remote full-time positions and full-time postdoc opportunities in talking head generation, generative models, or related areas. If you're interested, feel free to reach out at fatinghong@gmail.com.

:fire: New: ACTalker — portrait video generation driven by audio and expression simultaneously. ICCV 2025

Table of Contents

Datasets

  1. VoxCeleb1 [Download link].
  2. VoxCeleb2 [Download link].
  3. Faceforensics++ [Download link].
  4. CelebV [Download link].
  5. TalkingHead-1KH [Download link].
  6. LRW (Lip Reading in the Wild) [Download link].
  7. MEAD [Download link].
  8. CelebV-HQ [Download link].
  9. CHDTF [Download link].
  10. MultiTalk [Download link].
  11. VFHQ [Download link].
  12. Hallo3 [Download link].
  13. AVSpeech [Download link].

Survey


Image-driven

YearPaperVenueLinks
2026PersonaLive! Expressive Portrait Image Animation for Live StreamingCVPR 2026Code
2026PortraitDirector: A Hierarchical Disentanglement Framework for Controllable and Real-time Facial ReenactmentarXiv 2026
2025HunyuanPortrait: Implicit Condition Control for Enhanced Portrait AnimationCVPR 2025Code · Project
2025Robust Deepfake Detection for Electronic Know Your Customer Systems Using Registered ImagesarXiv 2025
2025Towards Interactive Intelligence for Digital HumansarXiv 2025
2025FlashPortrait: 6x Faster Infinite Portrait Animation with Adaptive Latent PredictionarXiv 2025
2024X-Portrait: Expressive Portrait Animation with Hierarchical Motion AttentionSIGGRAPH 2024Code
2024Follow-Your-Emoji: Fine-Controllable and Expressive Freestyle Portrait AnimationSIGGRAPH Asia 2024Code
2024LivePortrait: Efficient Portrait Animation with Stitching and Retargeting ControlarXiv 2024Code · Project
2024EMOPortraits: Emotion-enhanced Multimodal One-shot Head AvatarsCVPR 2024Code · Project
2024Synergizing Motion and Appearance: Multi-Scale Compensatory Codebooks for Talking Head Video GenerationCVPR 2024Project
2023Audio-Visual Face ReenactmentWACV 2023Code · Project
2023Cross-identity Video Motion Retargeting with Joint Transformation and SynthesisWACV 2023Code
2023Implicit Identity Representation Conditioned Memory Compensation Network for Talking Head Video GenerationICCV 2023Project · Code
2023StyleLipSync: Style-based Personalized Lip-sync Video GenerationICCV 2023Code
2022Depth-Aware Generative Adversarial Network for Talking Head Video GenerationCVPR 2022Code · Project
2022Thin-Plate Spline Motion Model for Image AnimationCVPR 2022Code
2022StyleHEAT: One-Shot High-Resolution Editable Talking Face Generation via Pretrained StyleGANECCV 2022Code · Project
2022MegaPortraits: One-shot Megapixel Neural Head AvatarsACM MM 2022Project
2022Structure-Aware Motion Transfer with Deformable Anchor ModelCVPR 2022Code
2022StyleMask: Disentangling the Style Space of StyleGAN2 for Neural Face ReenactmentFG, 2023Code
2022Controllable Radiance Fields for Dynamic Face SynthesisArxiv 2022
2022Animatable 3D-Aware Face Image Generation for Video AvatarsNeurIPS 2022Project
2022Implicit Warping for Animation with Image SetsNeurIPS 2022Project
2022HifiHead: One-Shot High Fidelity Neural Head Synthesis with 3D ControlIJCAI 2022
2022Face Animation with Multiple Source ImagesArxiv 2022
2022MetaPortrait: Identity-Preserving Talking Head Generation with Fast Personalized AdaptationArxiv 2022
2022Compressing Video Calls using Synthetic Talking HeadsBMVC 2022Project
2022Finding Directions in GAN’s Latent Space for Neural Face ReenactmentBMVC 2022Project · Code
2022Latent Image Animator: Learning to Animate Images via Latent Space NavigationICLR 2022Project · Code
2021One-Shot Free-View Neural Talking-Head Synthesis for Video ConferencingCVPR 2021 OralProject
2021Sparse to Dense Motion Transfer for Face Image AnimationICCV 2021
2021SAFA: Structure Aware Face Animation3DV 2021Code
2021Self-appearance-aided Differential Evolution for Motion TransferarXiv 2021
2021PIRenderer: Controllable Portrait Image Generation via Semantic Neural RenderingICCV 2021Code
2021FACEGAN: Facial Attribute Controllable rEenactment GANWACV 2021
2021F3A-GAN: Facial Flow for Face Animation With Generative Adversarial NetworksIEEE TIP 2021
2021FACIAL: Synthesizing Dynamic Talking Face with Implicit Attribute LearningICCV 2021
2021 Motion Representations for Articulated AnimationCVPR 2021Code
2021HeadGAN: One-shot Neural Head Synthesis and EditingICCV 2021Project
2020Mesh Guided One-shot Face Reenactment Using Graph Convolutional NetworksACM Multimedia 2020Code
2020MarioNETte: Few-shot Face Reenactment Preserving Identity of Unseen TargetsAAAI 2020Project
2020Learning Identity-Invariant Motion Representations for Cross-ID Face ReenactmentCVPR 2020
2019First order motion model for image animationNeurIPS 2019Code
2019Few-Shot Adversarial Learning of Realistic Neural Talking Head ModelsICCV 2019Code
2019Animating Arbitrary Objects via Deep Motion TransferCVPR 2019 OralCode · Project
2019Few-shot Video-to-Video SynthesisNeurIPS 2019Code · Project
2018ReenactGAN: Learning to Reenact Faces via Boundary TransferECCV 2018Code
2018X2Face: A network for controlling face generation by using images, audio, and pose codesECCV 2018Code · Project
2016Face2Face: Real-time face capture and reenactment of RGB videosCVPR 2016

Audio-driven

YearPaperVenueLinks
2026FunCineForge: A Unified Dataset Toolkit and Model for Zero-Shot Movie Dubbing in Diverse Cinematic ScenesarXiv 2026
2026TurboTalk: Progressive Distillation for One-Step Audio-Driven Talking Avatar GenerationarXiv 2026
2026SEDTalker: Emotion-Aware 3D Facial Animation Using Frame-Level Speech Emotion DiarizationarXiv 2026Code
2026AUHead: Realistic Emotional Talking Head Generation via Action Units ControlICLR 2026
2026UniTalking: A Unified Audio-Video Framework for Talking Portrait GenerationCVPR 2026
2026DiFlowDubber: Discrete Flow Matching for Automated Video Dubbing via Cross-Modal Alignment and SynchronizationCVPR 2026
2026ActAvatar: Temporally-Aware Precise Action Control for Talking AvatarsCVPR 2026
2026Cross-Modal Emotion Transfer for Emotion Editing in Talking Face VideoCVPR 2026
2026MMFace-DiT: A Dual-Stream Diffusion Transformer for High-Fidelity Multimodal Face GenerationCVPR 2026
2025OmniHuman-1: Rethinking the Scaling-Up of One-Stage Conditioned Human Animation ModelsarXiv 2025Project
2025Audio-visual Controlled Video Diffusion with Masked Selective State Spaces Modelling for Natural Talking Head GenerationICCV 2025Project
2025OmniAvatar: Efficient Audio-Driven Avatar Video Generation with Adaptive Body AnimationarXiv 2025Code · Project
2025Teller: Real-Time Streaming Audio-Driven Portrait Animation with Autoregressive Motion GenerationCVPR 2025
2025EmotiveTalk: Expressive Talking Head Generation through Audio Information Decoupling and Emotional Video DiffusionCVPR 2025Project
2025INFP: Audio-Driven Interactive Head Generation in Dyadic ConversationsCVPR 2025Project
2025Hallo2: Long-Duration and High-Resolution Audio-Driven Portrait Image AnimationICLR 2025Code
2025Loopy: Taming Audio-Driven Portrait Avatar with Long-Term Motion DependencyICLR 2025Project
2025DAWN: Dynamic Frame Avatar with Non-autoregressive Diffusion Framework for Talking Head Video GenerationICLR 2025Project · Code
2025AnyTalk: Multi-modal Driven Multi-domain Talking Head GenerationAAAI 2025
2025Occlusion-Insensitive Talking Head Video Generation via Facelet CompensationAAAI 2025
2025FixTalk: Taming Identity Leakage for High-Quality Talking Head GenerationICCV 2025
2025FLOAT: Generative Motion Latent Flow Matching for Audio-driven Talking PortraitICCV 2025Project
2025MoEE: Mixture of Emotion Experts for Audio-Driven Portrait AnimationCVPR 2025
2025Live Avatar: Streaming Real-time Audio-Driven Avatar Generation with Infinite LengtharXiv 2025Code
2025DreamTalk: When Expressive Talking Head Generation Meets Diffusion Probabilistic ModelsarXiv 2025
2025GAIA: Zero-shot Talking Avatar GenerationarXiv 2025
2024Real3D-Portrait: One-shot Realistic 3D Talking Portrait SynthesisICLR 2024Project · Code
2024Emote Portrait Alive - Generating Expressive Portrait Videos with Audio2Video Diffusion Model under Weak ConditionsarXiv 2024Project · Code
2024Style2Talker: High-Resolution Talking Head Generation with Emotion Style and Art StyleAAAI 2024
2024Say Anything with Any StyleAAAI 2024
2024[MuseTalk] Real-Time High Quality Lip Synchorization with Latent Space Inpainting, [Code].Code
2024VASA-1: Lifelike Audio-Driven Talking Faces Generated in Real TimeNeurIPS 2024Project
2024THQA: A Perceptual Quality Assessment Database for Talking HeadsarXiv 2024Code
2024Talk3D: High-Fidelity Talking Portrait Synthesis via Personalized 3D Generative PriorarXiv 2024Code · Project
2024EDTalk: Efficient Disentanglement for Emotional Talking Head SynthesisarXiv 2024Code · Project
2024AniPortrait: Audio-Driven Synthesis of Photorealistic Portrait AnimationsarXiv 2024Code
2024FlowVQTalker: High-Quality Emotional Talking Face Generation through Normalizing Flow and QuantizationarXiv 2024
2024FaceChain-ImagineID: Freely Crafting High-Fidelity Diverse Talking Faces from Disentangled AudioarXiv 2024Code
2024Hallo: Hierarchical Audio-Driven Visual Synthesis for Portrait Image AnimationarXiv 2024Code
2024EchoMimic: Lifelike Audio-Driven Portrait Animations through Editable Landmark ConditionsarXiv 2024Code · Project
2024RealTalk: Real-time and Realistic Audio-driven Face Generation with 3D Facial Prior-guided Identity Alignment NetworkarXiv 2024
2024Emotional Conversation: Empowering Talking Faces with Cohesive Expression, Gaze and Pose GenerationarXiv 2024
2024Make Your Actor Talk: Generalizable and High-Fidelity Lip Sync with Motion and Appearance DisentanglementarXiv 2024
2024FD2Talk: Towards Generalized Talking Head Generation with Facial Decoupled Diffusion ModelarXiv 2024
2024ReSyncer: Rewiring Style-based Generator for Unified Audio-Visually Synced Facial PerformerarXiv 2024
2024Style-Preserving Lip Sync via Audio-Aware Style ReferencearXiv 2024
2024EchoMimicV2: Towards Striking, Simplified, and Semi-Body Human AnimationarXiv 2024Code · Project
2024Latent Diffusion Transformer for Talking Video SynthesisarXiv 2024Code · Project
2024IF-MDM: Implicit Face Motion Diffusion Model for High-Fidelity Realtime Talking Head GenerationarXiv 2024Project
2024Memory-Guided Diffusion for Expressive Talking Video GenerationarXiv 2024Project · Code
2024Highly Dynamic and Realistic Portrait Image Animation with Diffusion Transformer NetworksarXiv 2024
2024VQTalker: Towards Multilingual Talking Avatars through Facial Motion TokenizationarXiv 2024
2024Towards Customizable One-Shot Audio-to-Talking Face GenerationarXiv 2024
2024LatentSync: Audio Conditioned Latent Diffusion Models for Lip SyncarXiv 2024Code
2024Media2Face: Co-speech Facial Animation Generation with Multi-Modality GuidanceSIGGRAPH 2024
2024PersonaTalk: Bring Attention to Your Persona in Visual DubbingSIGGRAPH Asia 2024
2024StyleTalk++: A Unified Framework for Controlling the Speaking Styles of Talking HeadsTPAMI 2024
2024JEAN: Joint Expression and Audio-guided NeRF-based Talking Face GenerationBMVC 2024
2024Ditto: Motion-Space Diffusion for Controllable Realtime Talking Head SynthesisarXiv 2024
2024JoyVASA: Portrait and Animal Image Animation with Diffusion-Based Audio-Driven Facial DynamicsarXiv 2024Code
2024HelloMeme: Integrating Spatial Knitting Attentions to Embed High-Level Conditions in Diffusion ModelsarXiv 2024Code
2024LaDTalk: Latent Denoising for Synthesizing Talking Head Videos with High Frequency DetailsarXiv 2024
2023Diffused Heads: Diffusion Models Beat GANs on Talking-Face GenerationArxiv 2023Project
2023DiffTalk: Crafting Diffusion Models for Generalized Talking Head SynthesisArxiv 2023Project · Code
2023[READ Avatars: Realistic Emotion-controllable Audio Driven Avatars](READ Avatars: Realistic Emotion-controllable Audio Driven Avatars)Arxiv 2023
2023DAE-Talker: High Fidelity Speech-Driven Talking Face Generation with Diffusion AutoencoderArxiv 2023
2023Emotionally Enhanced Talking Face GenerationArxiv 2023Code
2023Seeing What You Said: Talking Face Generation Guided by a Lip Reading ExpertCVPR 2023Code
2023StyleSync: High-Fidelity Generalized and Personalized Lip Sync in Style-based GeneratorCVPR 2023Project · Code
2023GeneFace++: Generalized and Stable Real-Time Audio-Driven 3D Talking Face GenerationarXiv 2023Project · Code
2023MODA: Mapping-Once Audio-driven Portrait Animation with Dual AttentionsICCV 2023
2023VividTalk: One-Shot Audio-Driven Talking Head Generation Based on 3D Hybrid PriorArxiv 2023Project · Code
2023IP_LAP: Identity-Preserving Talking Face Generation with Landmark and Appearance PriorsCVPR 2023Code
2023HyperLips: Hyper Control Lips with High Resolution Decoder for Talking Face GenerationCVPR 2023Code
2023Efficient Emotional Adaptation for Audio-Driven Talking-Head GenerationICCV 2023Project · Code
2023SadTalker: Learning Realistic 3D Motion Coefficients for Stylized Audio-Driven Talking Head AnimationCVPR 2023Project · Code
2023DINet: Deformation Inpainting Network for Realistic Face Visually Dubbing on High Resolution VideoAAAI 2023Code
2023EMMN: Emotional Motion Memory Network for Audio-driven Emotional Talking Face GenerationICCV 2023
2023ToonTalker: Cross-Domain Face ReenactmentICCV 2023
2023High-fidelity Generalized Emotional Talking Face Generation with Multi-modal Emotion Space LearningCVPR 2023
2023DisCoHead: Audio-and-Video-Driven Talking Head Generation by Disentangled Control of Head Pose and Facial ExpressionsICASSP 2023Code
2022Expressive Talking Head Generation with Granular Audio-Visual Control CVPR 2022
2022Talking Face Generation with Multilingual TTSCVPR 2022Demo
2022EAMM: One-Shot Emotional Talking Face via Audio-Based Emotion-Aware Motion ModelSIGGRAPH 2022
2022SPACEx 🚀: Speech-driven Portrait Animation with Controllable ExpressionarXiv 2022Project
2022Masked Lip-Sync Prediction by Audio-Visual Contextual Exploitation in TransformersSIGGRAPH Asia 2022
2022Memories are One-to-Many Mapping Alleviators in Talking Face GenerationarXiv 2022
2021Pose-Controllable Talking Face Generation by Implicitly Modularized Audio-Visual RepresentationCVPR 2021Code · Project
2021Imitating Arbitrary Talking Style for Realistic Audio-Driven Talking Face SynthesisACM Multimedia 2021
2021Audio-Driven Emotional Video PortraitsCVPR 2021Code
2021Talking Head Generation with Audio and Speech Related Facial Action Unitsarxiv 2021
2021Speech2Talking-Face: Inferring and Driving a Face with Synchronized Audio-Visual RepresentationIJCAI 2021
2021Imitating Arbitrary Talking Style for Realistic Audio-Driven Talking Face SynthesisACM MM 2021
2021Live Speech Portraits: Real-Time Photorealistic Talking-Head AnimationACM TOG 2021Code
2021Audio2head: Audio-driven one-shot talking-head generation with natural head motionArXiv 2021
2020A Lip Sync Expert Is All You Need for Speech to Lip Generation In The WildACM Multimedia 2020Code · Project
2020Talking-head Generation with Rhythmic Head MotionECCV 2020Code
2020MakeItTalk: Speaker-Aware Talking-Head AnimationSIGGRAPH Asia 2020Code · Project
2020Neural Voice Puppetry: Audio-driven Facial ReenactmentECCV 2020Code · Project
2020MEAD: A Large-scale Audio-visual Dataset for Emotional Talking-face GenerationECCV 2020Code · Project
2020Realistic Speech-Driven Facial Animation with GANsIJCV 2020
2019Talking Face Generation by Adversarially Disentangled Audio-Visual RepresentationAAAI 2019Code
2019Hierarchical Cross-modal Talking Face Generation with Dynamic Pixel-wise LossCVPR 2019Code
2018Lip Movements Generation at a GlanceECCV 2018Code
2018VisemeNet: Audio-Driven Animator-Centric Speech AnimationSIGGRAPH 2018
2017Synthesizing Obama: Learning Lip Sync From AudioSIGGRAPH 2017Project
2017You Said That?: Synthesising Talking Faces From AudioIJCV 2019Code
2017Audio-Driven Facial Animation by Joint End-to-End Learning of Pose and EmotionSIGGRAPH 2017
2017A Deep Learning Approach for Generalized Speech AnimationSIGGRAPH 2017
2016Lip Reading in the WildACCV 2016

Nerf & 3D

YearPaperVenueLinks
2026MMTalker: Multiresolution 3D Talking Head Synthesis with Multimodal Feature FusionarXiv 2026
2026EmoTaG: Emotion-Aware Talking Head Synthesis on Gaussian Splatting with Few-Shot PersonalizationCVPR 2026
2026GenFaceTalk: Generalizable One-Shot Talking-Head Generation for Diverse StylesICLR 2026
2026FG-Portrait: 3D Flow Guided Editable Portrait AnimationCVPR 2026
2026Giving Faces Their Feelings Back: Explicit Emotion Control for Feedforward Single-Image 3D Head AvatarsarXiv 2026
20263DRealHead: Few-Shot Detailed Head AvatararXiv 2026
2026FHAvatar: Fast and High-Fidelity Reconstruction of Face-and-Hair Composable 3D Head Avatar from Few Casual CapturesarXiv 2026
2026NBAvatar: Neural Billboards Avatars with Realistic Hand-Face InteractionarXiv 2026
2026OMG-Avatar: One-shot Multi-LOD Gaussian Head AvatararXiv 2026
2026OFERA: Blendshape-driven 3D Gaussian Control for Occluded Facial Expression to Realistic Avatars in VRarXiv 2026
2026CAG-Avatar: Cross-Attention Guided Gaussian Avatars for High-Fidelity Head ReconstructionarXiv 2026
2026Uncertainty-Aware 3D Emotional Talking Face Synthesis with Emotion Prior DistillationarXiv 2026
20263DXTalker: Unifying Identity, Lip Sync, Emotion, and Spatial Dynamics in Expressive 3D Talking AvatarsarXiv 2026
2025IM-Portrait: Learning 3D-aware Video Diffusion for Photorealistic Talking Heads from Monocular VideosCVPR 2025
2025Perceptually Accurate 3D Talking Head Generation: New Definitions, Speech-Mesh Representation, and Evaluation MetricsCVPR 2025
2025GaussianSpeech: Audio-Driven Personalized 3D Gaussian AvatarsICCV 2025Code
2025MemoryTalker: Personalized Speech-Driven 3D Facial Animation via Audio-Guided StylizationICCV 2025Code
2025CAFE-TALK: Generating 3D Talking FaceICLR 2025
2025InsTaG: Learning Personalized 3D Talking Head from Few-Second VideoCVPR 2025Code
2025Monocular and Generalizable Gaussian Talking Head AnimationCVPR 2025Project
2025DualTalk: Dual-Speaker Interaction for 3D Talking Head ConversationsCVPR 2025Code
2025VASA-3D: Lifelike Audio-Driven Gaussian Head Avatars from a Single ImageNeurIPS 2025
2025PointTalk: Audio-Driven Dynamic Lip Point Cloud for 3D Gaussian-based Talking Head SynthesisAAAI 2025
2025PTalker: Personalized Speech-Driven 3D Talking Head Animation via Style DisentanglementarXiv 2025
2025SynergyWarpNet: Attention-Guided Cooperative Warping for Neural Portrait AnimationarXiv 2025
2025From Autoencoders to CycleGAN: Robust Unpaired Face Manipulation via Adversarial LearningarXiv 2025
2024CVTHead: One-shot Controllable Head Avatar with Vertex-feature TransformerWACV 2024Code
20243D-Aware Talking-Head Video Motion TransferWACV 2024
2024TalkingGaussian: Structure-Persistent 3D Talking Head Synthesis via Gaussian SplattingECCV 2024Code
2024GaussianTalker: Real-Time High-Fidelity Talking Head Synthesis with Audio-Driven 3D Gaussian SplattingACM MM 2024Code
2024Generalizable and Animatable Gaussian Head AvatarNeurIPS 2024Code
2024MimicTalk: Mimicking a Personalized and Expressive 3D Talking Face in MinutesNeurIPS 2024
2024EmoTalk3D: High-Fidelity Free-View Synthesis of Emotional 3D Talking HeadECCV 2024
2024Learning Dynamic Tetrahedra for High-Quality Talking Head SynthesisCVPR 2024Code
2024GPAvatar: Generalizable and Precise Head Avatar from Image(s)ICLR 2024Code
2024UniTalker: Scaling up Audio-Driven 3D Facial Animation through A Unified ModelarXiv 2024Code
2023Efficient Region-Aware Neural Radiance Fields for High-Fidelity Talking Portrait SynthesisICCV 2023Code
2023EmoTalk: Speech-Driven Emotional Disentanglement for 3D Face AnimationICCV 2023Code
2023CodeTalker: Speech-Driven 3D Facial Animation with Discrete Motion PriorCVPR 2023Code
2023GANHead: Towards Generative Animatable Neural Head AvatarsCVPR 2023Code
2023OTAvatar: One-shot Talking Face Avatar with Controllable Tri-plane RenderingCVPR 2023Code
2023GeneFace: Generalized and High-Fidelity Audio-Driven 3D Talking Face SynthesisICLR 2023Code
2023SyncTalk: The Devil is in the Synchronization for Talking Head SynthesisCVPR 2024Code
2022Semantic-Aware Implicit Neural Audio-Driven Video Portrait Generationarxiv, 2022
2022HeadNeRF: A Real-time NeRF-based Parametric Head ModelCVPR 2022Code · Project
2022I M Avatar: Implicit Morphable Head Avatars from VideosCVPR 2022Code
2022Realistic One-shot Mesh-based Head AvatarsECCV 2022
2022FNeVR: Neural Volume Rendering for Face AnimationArxiv 2022Code
20223DFaceShop: Explicitly Controllable 3D-Aware Portrait GenerationArxiv 2022Code · Project
2022Generative Neural Texture Rasterization for 3D-Aware Head AvatarsArxiv 2022Project
2022NeRFInvertor: High Fidelity NeRF-GAN Inversion for Single-shot Real Image AnimationArxiv 2022
2022Learning Dynamic Facial Radiance Fields for Few-Shot Talking Head SynthesisECCV 2022Code
2021DFA-NeRF: Personalized Talking Head Generation via Disentangled Face Attributes Neural Renderingarxiv, 2021
2021NerFACE: Dynamic Neural Radiance Fields for Monocular 4D Facial Avatar ReconstructionCVPR 2021 OralCode · Project
2021AD-NeRF: Audio Driven Neural Radiance Fields for Talking Head SynthesisICCV 2021Code · Code
2020Disentangled and Controllable Face Image Generation via 3D Imitative-Contrastive LearningCVPR 2020 OralCode

Star History

Star History Chart

face-reenactment
image-animation
motion-transfer
talking-head

Contributors

harlanhong

87 commits

ykk648

22 commits

kkakkkka

3 commits

yuangan

2 commits