Purdue-M2/Detect-LAIM-generated-Multimedia-Survey

This repository contains a collection of resources and papers on Detecting Multimedia Generated by Large AI Models

119

158 commits

updated Jul 24, 2025

See the code

README

Detect-LAIM-generated-Multimedia-Survey

This repository contains a collection of resources and papers on Detecting Multimedia Generated by Large AI Models: A Survey

timeline

The references of those works are displayed in Generation Works and Detection Works.

Please let us know if you find a mistake, or if we have missed your wonderful work by e-mail: lin1785@purdue.edu, hu968@purdue.edu, gupt1031@purdue.edu

If you find our survey useful for your research, please cite the following Paper

@article{lin2024detecting,
  title={Detecting Multimedia Generated by Large AI Models: A Survey},
  author={Lin, Li and Gupta, Neeraj and Zhang, Yue and Ren, Hainan and Liu, Chun-Hao and Ding, Feng and Wang, Xin and Li, Xin and Verdoliva, Luisa and Hu, Shu},
  journal={arXiv preprint arXiv:2402.00045},
  year={2024}
}

💻 Contents

         - A Survey on Detection of LLMs-Generated Content Paper GitHub

         - A Survey on LLM-generated Text Detection: Necessity, Methods, and Future Directions Paper GitHub

         - Towards possibilities & impossibilities of ai-generated text detection: A survey Paper

         - Machine-generated text: A comprehensive survey of threat models and detection methods Paper

         - The Age of Synthetic Realities: Challenges and Opportunities Paper

         - GenAI against humanity: Nefarious applications of generative artificial intelligence and large language models Paper

         - A Comprehensive Survey of Fake Text Detection on Misinformation and LM-Generated Texts Paper

         - Recent Advances on Generalizable Diffusion-generated Image Detection Paper

         - Survey on AI-Generated Media Detection: From Non-MLLM to MLLM Paper

         - Passive Deepfake Detection Across Multi-modalities: A Comprehensive Survey Paper

Generation

Generation Processes Illustrations of different types of multimedia generation process based on LAIMs.

Public Datasets for Detection

Please read the column I20(Input-to-Output) with these abbreviations:

  • T2T: Text-to-Text
  • V2T: Video-to-Text
  • T2I: Text-to-Image
  • I2I: Image-to-Image
  • T2A: Text-to-Audio
  • I.A2V: (Image conditioned with Audio)-to-Video
ModalityDatasetYearBMContentLinkI2O#Real#GeneratedSource of Real MediaGenerative Method
TextTuringBench2021NewsLinkT2T8,854159,758News MediaGPT-1&2&3, CTRL, GROVER
Paraphrase2022EssaysLinkT2T98,280163,710Arxiv, Wikipedia, ThesesGPT-3, T5
SynSCiPass2022PassagesLinkT2T99,98999,989Scientific papersGPT-2, BLOOM
MAGE2023GeneralLinkT2T154,078294,381Reddit, EL15, Yelp, XSum27 LLMs
Stu.Essays2023EssaysLinkT2T1,0006,000Ivy PandaChatGPT
Writing2023StoriesLinkT2T1,0006,000Reddit WritingPromptsChatGPT
News2023NewsLinkT2T1,0006,000Reuters 50-50ChatGPT
OUTFOX2023NewsLinkT2T15,40015,400Feedback PrizeChatGPT, GPT-3.5, T5
MULTITuDE2023EssaysLinkT2T7,99216,005MassiveSummGPT-3&4, ChatGPT
MGTDetect-CoCo2023NewsLinkT2T10,48610,484News OutletsGPT-3.5
HPPT2023AbstractsLinkT2T1,0001,000ACL AnthologyChatGPT
HC-Var2023GeneralLinkT2T90,09690,096XSum , IMDb, Yelp, FiQAChatGPT
HC32023GeneralLinkT2T26,90358,546FiQA , EL15 , MediaDialogChatGPT
M42023GeneralLinkT2T32,79958,803WikiHow , Arxiv, RedditChatGPT, GPT-3.5, LLaMA, T5, BLOOM
F32023Social MediaLinkT2T12,72327,667Politifact , SnopesGPT-3.5
MixSet2024GeneralLinkT2T3,6003,600Email , BBC News, ArXivGPT-4, LLaMA2
GPABench2024WritingLinkT2T150,000450,000ArxivGPT-3.5
M4GT-Bench2024GeneralLinkT2T119,771119,388Wikipedia, WikiHow, Reddit, ArXiv, News10 LLMs
RAID2024GeneralLinkT2T14,9176,287,820Public datasets from 8 domains11 LLMs
DetectRL2024GeneralLinkT2T100,800134,400Writing Prompts , YelpGPT-3.5, PaLM2, Claude, LLaMA2
MultiSocial2024Social Media-T2T58,000414,000Gab, Discord, WhatsApp7 LLMs
SM-D2024Social Media-T2T--Medium, Quora, RedditSourced from social media
ImageDFF2023FaceLinkT/I2I30,00090,000IMDB-WIKISDMs, InsightFace
RealFaces2023FaceLinkT2I258258PromptsSDMs
OHImg2023OverheadLinkT/I2I6,4756,675MapBox , Google MapsGLIDE, DDPM
Western Blot2022BiologyLinkT/I2I~14,000-Western BlotDDPM, Pix2pix, CycleGAN
Synthbuster2023GeneralLinkT2I-9,000Raise-1KDALL-E 2&3, Midjourney, SDMs, GLIDE
GenImage2023GeneralLinkT/I2I1,331,1671,350,000ImageNetSDMs, Midjourney, BigGAN
CIFAKE2023GeneralLinkT/I2I60,00060,000CIFAR-10SD-V1.4
AutoSplice2023GeneralLinkT2I2,2753,621Visual NewsDALL-E 2
DiffusionDB2023GeneralLinkT/I2I3,300,00016,000,000DiscordChatExporterSD
Artifact2023GeneralLinkT/I2I1,749960,894COCO, FFHQ , COCO, LSUNSDMs, DDPM, LDM, CIPS
HiFi-FIDL2023GeneralLinkT/I2I~60,0001,300,000FFHQ , COCO, LSUNDDPM, GLIDE, LDM, GANs
DiffForensics2023GeneralLinkT/I2I232,000232,000LSUN, ImageNetLDM, DDPM, VQDM, ADM
CocoGlide2023GeneralLinkT/I2I512512COCOGLIDE
LSUNDB2023GeneralLinkT/I2I250,000250,000LSUNDDPM, LDM, StyleGAN
UniFake2023GeneralLinkT/I2I8,0008,000LAION-400MLDM, GLIDE
REGM2023GeneralLinkT/I2I116,000116,000CelebA , LSUN116 publicly available GMs
DMImage2023GeneralLinkT/I2I200,000200,000COCO, LSUNLDM
AIGCD2023GeneralLinkT/I2I360,000580,000LSUN, COCO, FFHQSDMs, GANs, ADM, DALL-E 2, GLIDE
DIF2023GeneralLinkT/I2I34,80054,500LAION-585SDMs, DALL-E 2, GLIDE, GANS
Fake2M2024GeneralLinkT/I2I2,300,000-CC3MSD-V1.5, IOI, IF , StyleGAN3
SID-Set2024Social MediaLinkT/I2I100,000100,000COCO, Flickr30K, MagicBrushFLUX
Chameleon2024Social MediaLinkT/I2I14,86311,170UnsplashGANs, SDMs, DALL-E 2, GLIDE
DF402024FaceLinkT/I2I~1,100~1,000,000FF++, CDF, FFHQ, CelebASDMs, GANs, Midjourney, DDPM
FakeBench2024GeneralLinkT/I2I3,0003,00010 Public Datasets10 Generative Models
AI-Face2024FaceLinkT/I2I400,8851,245,6606 Public datasetsSDMs, GANs, Midjourney, IF
VideoWildDeepfake2021FaceLinkI.A2V3,8053,509Social MediaSocial Media
DiffHead2023FaceLinkI.A2V820-CREMADiffused Heads: build on DDPM
DVF2024GeneralLinkI/T2V2,7503,938Intervid , Youtube-8M8 Diffusion Models
GenVideo2024GeneralLinkI/T2V1,223,5111,078,838Kinetics-400 , Youku-mPLUG, MSR-VTT20 Generative Models
GenVidBench2025GeneralLinkI/T2V33,931110,400Vript , HDL-VG-130M8 Generative Models
PDID2024Social MediaLink---Social MediaSocial Media
AudioIn-the-Wild2022SpeechLinkT2A20.7 hours17.2 hoursSocial Media, Video Streaming PlatformsSocial Media, Video Streaming Platforms
LibriSeVoc2023SpeechLinkT2A13,20179,206LibriTTSDiffWave, WaveNet
SONAR2024SpeechLinkT2A-2,274LibriTTSOpenAI, Seed-TTS, AudioGen
ASVspoof 20242024SpeechLinkT/A2A~289,527~1,211,186MLS-English32 Manipulation Methods
Multi-modalDGM^42023NewsLinkT/I2T77,426152,574Visual NewsB-GST, StyleCLIP, HFGI
COCOFake2023GeneralLinkT/I2T113,287566,435COCOSDMs
AV-Deepfake1M2023FaceLinkT2A286,721860,039Voxceleb2VITS, YourTTS, TalkLip
2024GeneralLinkT2I~2,300,000~9,200,000LAION-400MSDMs, IF
M³A2024News-T2T/I/T/V/V/A/T
T/I/V/V/A2T
708,4256,566,38660 News OutletsLLaMA2, GPT-4, GLIDE, SD, Tango
LOKI2024GeneralLinkT2T/I/T/V/V/T
T/I/V/V/A2T
~9,000~9,00021 Public Datasets43 Generative Models
MMFakeBench2024Social MediaLinkT2T/I-~11,000MS-COCO, VisualNews, Reddit, FEVERGPT-3.5, SD-XL, DALL-E 3, Midjourney
Deepfake-Eval2024Social MediaLinkT2T/A/V3,3902,441Social MediaSocial Media
ILLUSION2025GeneralLinkT2A/I
I2I
139,7401,232,246CelebV-Text [158], COCO, MusicCaps , Social Media28 Generative Methods

:mag_right: Detection :fire:

:page_facing_up: Text


Pure Detection

text_pure Illustrations of pure detection methodologies for LAIM-generated text.

  ♣️ Easy Explainable Methods

        ▶️ Watermarking

         - Distillation-Resistant Watermarking for Model Protection in NLP Paper

         - Three bricks to consolidate watermarks for large language models Paper GitHub

         - Robust multi-bit natural language watermarking through invariant features Paper

         - Undetectable Watermarks for Language Models Paper

         - Robust distortion-free watermarks for language models Paper

         - Provable robust watermarking for ai-generated text Paper GitHub

         - A Private Watermark for Large Language Models Paper

        ▶️ Non-watermarking

         - Unraveling the mystery of artifacts in machine generated text Paper

         - Stylometric detection of ai-generated text in twitter timelines Paper

         - CoCo: Coherence-Enhanced Machine-Generated Text Detection Under Data Limitation With Contrastive Learning Paper

         - Beat LLMs at Their Own Game: Zero-Shot LLM-Generated Text Detection via Querying ChatGPT Paper GitHub

         - Who Wrote This? The Key to Zero-Shot LLM-Generated Text Detection Is GECScore Paper GitHub

  ♣️ Hard Explainable Methods

         - HowkGPT: Investigating the Detection of ChatGPT-generated University Student Homework through Context-Aware Perplexity Analysis Paper

         - GPTZero Tool

         - Detectgpt: Zero-shot machine-generated text detection using probability curvature Paper GitHub

         - Spotting LLMs With Binoculars: Zero-Shot Detection of Machine-Generated Text Paper GitHub

         - Multiscale Positive-Unlabeled Detection of AI-Generated Texts Paper GitHub

Beyond Detection

text_beyond Illustrations of beyond detection methodologies for LAIM-generated text.

  ♣️ Efficiency

         - Efficient Detection of LLM-generated Texts with a Bayesian Surrogate Model Paper

         - Fast-DetectGPT: Efficient Zero-Shot Detection of Machine-Generated Text via Conditional Probability Curvature Paper GitHub

         - DetectLLM: Leveraging Log Rank Information for Zero-Shot Detection of Machine-Generated Text Paper GitHub

         - SeqXGPT: Sentence-Level AI-Generated Text Detection Paper GitHub

         - Glimpse: Enabling White-Box Methods to Use Proprietary Models for Zero-Shot LLM-Generated Text Detection Paper GitHub

  ♣️ Attribution

         - TURINGBENCH: A Benchmark Environment for Turing Test in the Age of Neural Text Generation Paper Turingbench

         - Whodunit? Learning to Contrast for Authorship Attribution Paper

         - Through the looking glass: Learning to attribute synthetic text generated by language models Paper

         - TopRoBERTa: Topology-Aware Authorship Attribution of Deepfake Texts Paper

         - Authorship attribution for neural text generation Paper GitHub

         - Gpt-who: An information density-based machine-generated text detector Paper

         - LLMDet: A Third Party Large Language Models Generated Text Detection Tool Paper GitHub

         - Few-Shot Detection of Machine-Generated Text using Style Representations Paper

         - Origin Tracing and Detecting of LLMs Paper

  ♣️ Generalization

         - Ghostbuster: Detecting Text Ghostwritten by Large Language Models Paper

         - Conda: Contrastive domain adaptation for ai-generated text detection Paper GitHub

         - Text Fluoroscopy: Detecting LLM-Generated Text through Intrinsic Features Paper GitHub

         - DeTeCtive: Detecting AI-generated Text via Multi-Level Contrastive Learning Paper GitHub

         - Intrinsic Dimension Estimation for Robust Detection of AI-Generated Texts Paper GitHub

  ♣️ Interpretability

         - DNA-GPT: Divergent N-Gram Analysis for Training-Free Detection of GPT-Generated Text Paper GitHub

         - A Watermark for Large Language Models Paper GitHub

         - Chatgpt or human? detect and explain. explaining decisions of machine learning model for detecting short chatgpt-generated text Paper

         - Check Me If You Can: Detecting ChatGPT-Generated Academic Writing using CheckGPT Paper

         - Is chatgpt involved in texts? measure the polish ratio to detect chatgpt-generated text Paper

  ♣️ Robustness

        ▶️ Adversarial Attack Robustness

         - Red Teaming Language Model Detectors with Language Models Paper

         - Radar: Robust ai-text detection via adversarial learning Paper Project Page

         - J-guard: Journalism guided adversarially robust detection of ai-generated news Paper

         - Outfox: Llm-generated essay detection through in-context learning with adversarially generated examples Paper

        ▶️ LAIM-Polished Robustness

         - Is chatgpt involved in texts? measure the polish ratio to detect chatgpt-generated text Paper

  ♣️ Empirical Study

         - ChatLog: Recording and Analyzing ChatGPT Across Time Paper GitHub

         - On the Zero-Shot Generalization of Machine-Generated Text Detectors Paper

         - On the Generalization of Training-based ChatGPT Detection Methods Paper

         - Supervised Machine-Generated Text Detectors: Family and Scale Matters Paper GitHub

         - Deepfake Text Detection in the Wild Paper GitHub

         - How large language models are transforming machine-paraphrased plagiarism Paper

         - Paraphrase Detection: Human vs. Machine Content Paper

         - MGTBench: Benchmarking Machine-Generated Text Detection Paper GitHub

         - How close is chatgpt to human experts? comparison corpus, evaluation, and detection Paper GitHub

         - Can LLM-Generated Misinformation Be Detected? Paper GitHub

         - From Text to Source: Results in Detecting Large Language Model-Generated Content Paper

📸 Image


Pure Detection

image_pure Illustrations of pure detection methodologies for LAIM-generated image.

  ♣️ Physical/Physiological based Methods

         - Qualitative Failures of Image Generation Models and Their Application in Detecting Deepfakes Paper

         - Perspective (in) consistency of paint by text Paper

         - Lighting (in) consistency of paint by text Paper

  ♣️ Diffuser Fingerprints based Methods

         - Deep Image Fingerprint: Accurate And Low Budget Synthetic Image Detector Paper

         - DIRE for Diffusion-Generated Image Detection Paper GitHub

         - Exposing the Fake: Effective Diffusion-Generated Images Detection Paper

         - LaRE^2: Latent Reconstruction Error Based Method for Diffusion-Generated Image Detection Paper GitHub

         - Aligned Datasets Improve Detection of Latent Diffusion-Generated Images Paper GitHub

         - Manifold Induced Biases for Zero-shot and Few-shot Detection of Generated Images Paper GitHub

  ♣️ Spatial-based Methods

         - Rich and Poor Texture Contrast: A Simple yet Effective Approach for AI-generated Image Detection Paper Project Page

         - Unmasking The Artist: Discriminating Human-Drawn And AI-Generated Human Face Art Through Facial Feature Analysis Paper

         - Detecting images generated by deep diffusion models using their local intrinsic dimensionality Paper

  ♣️ Frequency-based Methods

         - Wavelet-packets for deepfake image analysis and detection Paper GitHub

         - AUSOME: authenticating social media images using frequency analysis Paper

         - AI-Generated Image Detection using a Cross-Attention Enhanced Dual-Stream Network Paper

         - Synthbuster: Towards Detection of Diffusion Model Generated Images Paper

         - Faster Than Lies: Real-time Deepfake Detection using Binary Neural Networks Paper GitHub

  ♣️ Distribution-based Methods

         - Zero-Shot Detection of AI-Generated Images Paper GitHub

Beyond Detection

image_beyond Illustrations of beyond detection methodologies for LAIM-generated image.

  ♣️ Attribution and Model Parsing

        ▶️ Attribution and Model Parsing

         - Level up the deepfake detection: a method to effectively discriminate images generated by gan architectures and diffusion models Paper

         - Reverse engineering of generative models: Inferring model hyperparameters from generated images Paper

  ♣️ Generalization

         - Online Detection of AI-Generated Images Paper

         - Towards universal fake image detectors that generalize across generative models Paper GitHub

         - Raising the Bar of AI-generated Image Detection with CLIP Paper

         - Transcending Forgery Specificity with Latent Space Augmentation for Generalizable Deepfake Detection Paper

         - Fingerprintnet: Synthesized fingerprints for generated image detection Paper

         - Detecting Deepfakes Without Seeing Any Paper GitHub

         - Improving Synthetically Generated Image Detection in Cross-Concept Settings Paper

         - Diffusion Noise Feature: Accurate and Fast Generated Image Detection Paper

         - Contrasting Deepfakes Diffusion via Contrastive Learning and Global-Local Similarities Paper GitHub

         - HRR: Hierarchical Retrospection Refinement for Generated Image Detection Paper

         - A Sanity Check for AI-generated Image Detection Paper GitHub

         - Stacking Brick by Brick: Aligned Feature Isolation for Incremental Face Forgery Detection Paper GitHub

         - A Bias-Free Training Paradigm for More General AI-generated Image Detection Paper GitHub

         - Breaking Semantic Artifacts for Generalized AI-generated Image Detection Paper GitHub

         - Dual Data Alignment Makes AI‑Generated Image Detector Easier Generalizable Paper

  ♣️ Interpretability

         - Interpretable-through-prototypes deepfake detection for diffusion models Paper GitHub

         - Did You Note My Palette? Unveiling Synthetic Images Through Color Statistics Paper

  ♣️ Localization

        ▶️ Fully-supervised

         - Hierarchical fine-grained image forgery detection and localization Paper GitHub

         - Perceptual Artifacts Localization for Image Synthesis Tasks Paper GitHub

         - TruFor: Leveraging all-round clues for trustworthy image forgery detection and localization Paper GitHub

         - UnionFormer: Unified-Learning Transformer with Multi-View Representation for Image Manipulation Detection and Localization Paper

        ▶️ Weakly-supervised

         - Weakly-supervised deepfake localization in diffusion-generated images Paper

  ♣️ Robustness

        ▶️ Adversarial Attack Robustness

         - D4: Detection of Adversarial Diffusion Deepfakes Using Disjoint Ensembles Paper

         - Exploring the Adversarial Robustness of CLIP for AI-generated Image Detection Paper

         - All Patches Matter, More Patches Better: Enhance AI‑Generated Image Detection via Panoptic Patch Learning Paper

        ▶️ Post-Processing Robustness

         - GLFF: Global and Local Feature Fusion for AI-synthesized Image Detection Paper

         - Exposing fake images generated by text-to-image diffusion models Paper

         - Local Statistics for Generative Image Detection Paper

  ♣️ Empirical Study

         - On the detection of synthetic images generated by diffusion models Paper GitHub

         - Intriguing properties of synthetic images: from generative adversarial networks to diffusion models Paper

         - Towards the detection of diffusion model deepfakes Paper

         - Unveiling the Impact of Image Transformations on Deepfake Detection: An Experimental Analysis Paper

         - On the use of Stable Diffusion for creating realistic faces: from generation to detection Paper

         - Finding AI-Generated Faces in the Wild Paper

         - Forensic analysis of synthetically generated western blot images Paper

         - Beyond Human Forgeries: An Investigation into Detecting Diffusion-Generated Handwriting Paper

🎞️ Video


Video Detection

Illustration of detection methodology in generalization task for LAIM-generated video.

Pure Detection

  ♣️ Spatial & Temporal based Methods

         - Distinguish Any Fake Videos: Unleashing the Power of Large-scale Data and Motion Features Paper

         - Exposing AI-generated Videos: A Benchmark Dataset and a Local-and-Global Temporal Defect Based Detection Method Paper

Beyond Detection

  ♣️ Generalization

         - Revisiting Generalizability in Deepfake Detection: Improving Metrics and Stabilizing Transfer Paper

  ♣️ Empirical Study

         - Beyond Deepfake Images: Detecting AI-Generated Videos Paper

🎵 Audio


Pure Detection

Audio Detection

The artifacts introduced by DM-based neural vocoders (WaveGrad and DiffWave) to a voice signal. The differences in mel-spectrograms between real and generated ones are illustrated in the third and fifth columns.

  ♣️ Vocoder-based

         - AI-Synthesized Voice Detection Using Neural Vocoder Artifacts Paper GitHub

Beyond Detection

  ♣️ Generalization

         - Improving Generalization for AI-Synthesized Voice Detection Paper GitHub

🍯 Multimodal


Pure Detection

Multimodal Detection

Illustrations of pure detection methodologies for LAIM-generated multimodal media.

  ♣️ Prompt-guided

         - Parents and Children: Distinguishing Multimodal DeepFakes from Natural Images Paper

         - On Learning Multi-Modal Forgery Representation for Diffusion Generated Video Detection Paper GitHub

         - Human Action CLIPS: Detecting AI-generated Human Motion Paper

  ♣️ Text-image Inconsistency

         - Detecting Cross-Modal Inconsistency to Defend Against Neural Fake News Paper GitHub

         - Exposing Text-Image Inconsistency Using Diffusion Models Paper

Beyond Detection

Multimodal Detection

Illustrations of beyond detection methodologies for LAIM-generated multimodal media.

  ♣️ Attribution

         - De-fake: Detection and attribution of fake images generated by text-to-image generation models Paper

         - FIDAVL: Fake Image Detection and Attribution using Vision-Language Model Paper GitHub

  ♣️ Generalization

        ▶️ Prompt Tuning

         - AntifakePrompt: Prompt-Tuned Vision-Language Models are Fake Image Detectors Paper GitHub

        ▶️ Contrastive Learning

         - Generalizable Synthetic Image Detection via Language-guided Contrastive Learning Paper GitHub

  ♣️ Interpretability

         - Combating Misinformation in the Era of Generative AI Models Paper

         - FFAA: Multimodal Large Language Model based Explainable Open-World Face Forgery Analysis Assistant Paper GitHub

         - X^2-DFD: A framework for eXplainable and eXtendable deepfake detection Paper

  ♣️ Localization

        ▶️ Spatial-based

         - Detecting and grounding multi-modal media manipulation Paper

         - Exploiting Modality-Specific Features For Multi-Modal Manipulation Detection And Grounding Paper

        ▶️ Frequency-based

         - Unified Frequency-Assisted Transformer Framework for Detecting and Grounding Multi-Modal Manipulation Paper

        ▶️ MLLM-based

         - FakeShield: Explainable Image Forgery Detection and Localization via Multi-modal Large Language Models Paper GitHub

         - ForgeryGPT: Multimodal Large Language Model For Explainable Image Forgery Detection and Localization Paper

  ♣️ Empirical Study

         - Detecting Images Generated by Diffusers Paper GitHub

         - CLIPping the Deception: Adapting Vision-Language Models for Universal Deepfake Detection Paper

         - VERITE: a Robust benchmark for multimodal misinformation detection accounting for unimodal bias Paper GitHub

         - Can ChatGPT Detect DeepFakes? A Study of Using Multimodal Large Language Models for Media Forensics Paper

Detection Tools

ModalityToolCompanyLinkTypeOpen SourceCost
TextAI Content DetectorCopyleaksLinkWebapp & APILimited free usage
AI Content Detector, ChatGPT detectorZeroGPTLinkWebapp & APIFree usage
AI DetectorGPTZeroLinkMulti-platformLimited free usage
AI Content DetectorWinston AILinkWebapp & APILimited free usage
AI Content DetectorCrossplagLinkWebappLimited free usage
Giant Language model Test RoomGLTRLinkWebappFree usage
The AI DetectorBrandwellLinkWebappFree usage
AI CheckerOriginality aiLinkWebapp & APILimited free usage
Advanced AI Detector and HumanizerUndetectable aiLinkWebapp & APILimited free usage
AI Content DetectorWriterLinkWebapp & APILimited free usage
AI Content DetectorConchLinkWebappLimited free usage
Illuminarty TextIlluminartyLinkWebapp & APILimited free usage
AI-Generated Text DetectorIs it AILinkWebapp & APILimited free usage
ImageLiveness Detection, Facial RecognitionIncodeLinkMulti-platformPaid
AI or Not imageAI or NotLinkWebapp & APILimited free usage
AI-Generated Image DetectorIs it AILinkWebapp & APILimited free usage
Illuminarty ImageIlluminartyLinkWebapp & APILimited free usage
AI Image DetectorUndetectable aiLinkWebapp & APILimited free usage
SynthIDGoogleLinkWebappFree usage
The AI image detectorWinstonLinkWebapp & APILimited free usage
Advanced AI Image DetectorBrandwellLinkWebappLimited free usage
VideoDeepware ScannerDeepwareLinkWebapp & APIFree usage
Attestiv Deepfake Video DetectionAttestivLinkWebapp & APILimited free usage
AudioPulse InspectPindropLinkMulti-platformPaid
AI Voice DetectorAI Voice DetectorLinkWebapp & APILimited free usage
AI Speech ClassifierElevenLabsLinkWebapp & APILimited free usage
AI or Not audioAI or NotLinkMulti-platformLimited free usage
Multi-modalVideo, Image, and Audio DetectorDeep MediaLinkMulti-platformLimited free usage
Deepfake DetectionSensity AILinkMulti-platformPaid
Hive AI’s Deepfake Detection APIHive AILinkAPILimited free usage
Resemble DetectResemble AILinkWebapp & APILimited free usage
DuckDuckGoose AI (Phocus)DuckDuckGoose AILinkWebappPaid
SentinelSentinelLinkWebappPaid
Deepfake DetectorDeepfake DetectorLinkMulti-platformFree usage
DeepFake-o-meterU of BuffaloLinkWebappFree usage
BioIDBioIDLinkWebapp & APILimited free usage
Get Real ProtectGet RealLinkMulti-platformPaid
Reality DefenderReality DefenderLinkMulti-platformPaid

Generation Works

WorksTimeModalityLinks
T5Q4 2019TextLink
GPT-3Q2 2020TextLink
Wave-Grad2Q1 2021AudioLink
PanGuQ2 2021TextLink
LDMsQ4 2021ImageLink
GLIDEQ4 2021ImageLink
ImagenQ2 2022ImageLink
PaLMQ2 2022TextLink
OPTQ2 2022TextLink
Make-A-VideoQ3 2022VideoLink
GLMQ3 2022TextLink
HuggingGPTQ3 2022MultimodalLink
WhisperQ3 2022AudioLink
ChatGPTQ4 2022TextLink
DALL-E 2Q4 2022ImageLink
SDQ4 2022ImageLink
mT0Q4 2022TextLink
BLOOMQ4 2022TextLink
Make-An-AudioQ1 2023AudioLink
GPT-4Q1 2023MultimodalLink
BardQ1 2023TextLink
LLaMAQ1 2023TextLink
GEN-1Q1 2023VideoLink
ImageRewardQ2 2023ImageLink
PaLM2Q2 2023TextLink
CodeGen2Q2 2023TextLink
IFQ2 2023ImageLink
VideoGenQ3 2023VideoLink
DALL-E 3Q3 2023ImageLink
LLaMA 2Q3 2023TextLink
GeminiQ4 2023TextLink
Emu EditQ4 2023ImageLink
Emu VideoQ4 2023VideoLink
TitanQ4 2023ImageLink
Stable VideoQ4 2023VideoLink
MidjourneyV6Q4 2023ImageLink
Imagen 2Q4 2023ImageLink
Claude 3.5Q1 2024MultimodalLink
aMUSEdQ1 2024ImageLink
Synthesia 2Q2 2024VideoLink
MultiBoothQ2 2024ImageLink
GPT-4oQ2 2024MultimodalLink
LLaMA 3Q2 2024MultimodalLink
GLM-4Q2 2024MultimodalLink
CustomCrafterQ3 2024VideoLink
MegaFusionQ3 2024ImageLink
Qwen 2Q3 2024MultimodalLink
Tri-ErgonQ4 2024AudioLink
Veo 2Q4 2024VideoLink
SoraQ4 2024VideoLink
AudioXQ1 2025AudioLink
Grok 3Q1 2025MultimodalLink
DeepSeek-V3Q1 2025MultimodalLink
Gemini 2.5 ProQ1 2025MultimodalLink
LLaMA 4Q2 2025MultimodalLink
Qwen 3Q2 2025MultimodalLink

Detection Works

WorksTimeModalityLinks
LinguisticQ4 2020TextLink
XLNet-FTQ1 2021TextLink
Turing-BenchQ4 2021TextLink
UnravelingQ2 2022TextLink
WaveletQ3 2022ImageLink
WhodunitQ3 2022TextLink
De-FakeQ4 2022MultimodalLink
TowardsQ4 2022ImageLink
TruForQ4 2022ImageLink
DIREQ1 2023ImageLink
GPTZeroQ1 2023TextLink
DetectGPTQ1 2023TextLink
HAMMERQ2 2023MultimodalLink
DetectVocoderQ2 2023AudioLink
SeDIDQ3 2023ImageLink
RADARQ3 2023TextLink
OUTFOXQ3 2023ImageLink
SeqXGPTQ4 2023TextLink
RevisitVideoQ4 2023VideoLink
RaisingQ4 2023ImageLink
BinocularsQ1 2024TextLink
AI FaceQ2 2024ImageLink
DuB3DQ2 2024VideoLink
GECScoreQ2 2024TextLink
FFAAQ3 2024MultimodalLink
BreakingQ4 2024ImageLink
B-FreeQ4 2024ImageLink
ForgeryGPTQ4 2024MultimodalLink
FakeShieldQ4 2024MultimodalLink
GenVidBenchQ1 2025VideoLink

Contributors

Purdue-M2

80 commits

ll-cqu

56 commits

gneeraj97

18 commits

discovershu

4 commits

Purdue-M2/Detect-LAIM-generated-Multimedia-Survey

This repository contains a collection of resources and papers on Detecting Multimedia Generated by Large AI Models

119

158 commits

updated Jul 24, 2025

See the code

README

Detect-LAIM-generated-Multimedia-Survey

This repository contains a collection of resources and papers on Detecting Multimedia Generated by Large AI Models: A Survey

timeline

The references of those works are displayed in Generation Works and Detection Works.

Please let us know if you find a mistake, or if we have missed your wonderful work by e-mail: lin1785@purdue.edu, hu968@purdue.edu, gupt1031@purdue.edu

If you find our survey useful for your research, please cite the following Paper

@article{lin2024detecting,
  title={Detecting Multimedia Generated by Large AI Models: A Survey},
  author={Lin, Li and Gupta, Neeraj and Zhang, Yue and Ren, Hainan and Liu, Chun-Hao and Ding, Feng and Wang, Xin and Li, Xin and Verdoliva, Luisa and Hu, Shu},
  journal={arXiv preprint arXiv:2402.00045},
  year={2024}
}

💻 Contents

         - A Survey on Detection of LLMs-Generated Content Paper GitHub

         - A Survey on LLM-generated Text Detection: Necessity, Methods, and Future Directions Paper GitHub

         - Towards possibilities & impossibilities of ai-generated text detection: A survey Paper

         - Machine-generated text: A comprehensive survey of threat models and detection methods Paper

         - The Age of Synthetic Realities: Challenges and Opportunities Paper

         - GenAI against humanity: Nefarious applications of generative artificial intelligence and large language models Paper

         - A Comprehensive Survey of Fake Text Detection on Misinformation and LM-Generated Texts Paper

         - Recent Advances on Generalizable Diffusion-generated Image Detection Paper

         - Survey on AI-Generated Media Detection: From Non-MLLM to MLLM Paper

         - Passive Deepfake Detection Across Multi-modalities: A Comprehensive Survey Paper

Generation

Generation Processes Illustrations of different types of multimedia generation process based on LAIMs.

Public Datasets for Detection

Please read the column I20(Input-to-Output) with these abbreviations:

  • T2T: Text-to-Text
  • V2T: Video-to-Text
  • T2I: Text-to-Image
  • I2I: Image-to-Image
  • T2A: Text-to-Audio
  • I.A2V: (Image conditioned with Audio)-to-Video
ModalityDatasetYearBMContentLinkI2O#Real#GeneratedSource of Real MediaGenerative Method
TextTuringBench2021NewsLinkT2T8,854159,758News MediaGPT-1&2&3, CTRL, GROVER
Paraphrase2022EssaysLinkT2T98,280163,710Arxiv, Wikipedia, ThesesGPT-3, T5
SynSCiPass2022PassagesLinkT2T99,98999,989Scientific papersGPT-2, BLOOM
MAGE2023GeneralLinkT2T154,078294,381Reddit, EL15, Yelp, XSum27 LLMs
Stu.Essays2023EssaysLinkT2T1,0006,000Ivy PandaChatGPT
Writing2023StoriesLinkT2T1,0006,000Reddit WritingPromptsChatGPT
News2023NewsLinkT2T1,0006,000Reuters 50-50ChatGPT
OUTFOX2023NewsLinkT2T15,40015,400Feedback PrizeChatGPT, GPT-3.5, T5
MULTITuDE2023EssaysLinkT2T7,99216,005MassiveSummGPT-3&4, ChatGPT
MGTDetect-CoCo2023NewsLinkT2T10,48610,484News OutletsGPT-3.5
HPPT2023AbstractsLinkT2T1,0001,000ACL AnthologyChatGPT
HC-Var2023GeneralLinkT2T90,09690,096XSum , IMDb, Yelp, FiQAChatGPT
HC32023GeneralLinkT2T26,90358,546FiQA , EL15 , MediaDialogChatGPT
M42023GeneralLinkT2T32,79958,803WikiHow , Arxiv, RedditChatGPT, GPT-3.5, LLaMA, T5, BLOOM
F32023Social MediaLinkT2T12,72327,667Politifact , SnopesGPT-3.5
MixSet2024GeneralLinkT2T3,6003,600Email , BBC News, ArXivGPT-4, LLaMA2
GPABench2024WritingLinkT2T150,000450,000ArxivGPT-3.5
M4GT-Bench2024GeneralLinkT2T119,771119,388Wikipedia, WikiHow, Reddit, ArXiv, News10 LLMs
RAID2024GeneralLinkT2T14,9176,287,820Public datasets from 8 domains11 LLMs
DetectRL2024GeneralLinkT2T100,800134,400Writing Prompts , YelpGPT-3.5, PaLM2, Claude, LLaMA2
MultiSocial2024Social Media-T2T58,000414,000Gab, Discord, WhatsApp7 LLMs
SM-D2024Social Media-T2T--Medium, Quora, RedditSourced from social media
ImageDFF2023FaceLinkT/I2I30,00090,000IMDB-WIKISDMs, InsightFace
RealFaces2023FaceLinkT2I258258PromptsSDMs
OHImg2023OverheadLinkT/I2I6,4756,675MapBox , Google MapsGLIDE, DDPM
Western Blot2022BiologyLinkT/I2I~14,000-Western BlotDDPM, Pix2pix, CycleGAN
Synthbuster2023GeneralLinkT2I-9,000Raise-1KDALL-E 2&3, Midjourney, SDMs, GLIDE
GenImage2023GeneralLinkT/I2I1,331,1671,350,000ImageNetSDMs, Midjourney, BigGAN
CIFAKE2023GeneralLinkT/I2I60,00060,000CIFAR-10SD-V1.4
AutoSplice2023GeneralLinkT2I2,2753,621Visual NewsDALL-E 2
DiffusionDB2023GeneralLinkT/I2I3,300,00016,000,000DiscordChatExporterSD
Artifact2023GeneralLinkT/I2I1,749960,894COCO, FFHQ , COCO, LSUNSDMs, DDPM, LDM, CIPS
HiFi-FIDL2023GeneralLinkT/I2I~60,0001,300,000FFHQ , COCO, LSUNDDPM, GLIDE, LDM, GANs
DiffForensics2023GeneralLinkT/I2I232,000232,000LSUN, ImageNetLDM, DDPM, VQDM, ADM
CocoGlide2023GeneralLinkT/I2I512512COCOGLIDE
LSUNDB2023GeneralLinkT/I2I250,000250,000LSUNDDPM, LDM, StyleGAN
UniFake2023GeneralLinkT/I2I8,0008,000LAION-400MLDM, GLIDE
REGM2023GeneralLinkT/I2I116,000116,000CelebA , LSUN116 publicly available GMs
DMImage2023GeneralLinkT/I2I200,000200,000COCO, LSUNLDM
AIGCD2023GeneralLinkT/I2I360,000580,000LSUN, COCO, FFHQSDMs, GANs, ADM, DALL-E 2, GLIDE
DIF2023GeneralLinkT/I2I34,80054,500LAION-585SDMs, DALL-E 2, GLIDE, GANS
Fake2M2024GeneralLinkT/I2I2,300,000-CC3MSD-V1.5, IOI, IF , StyleGAN3
SID-Set2024Social MediaLinkT/I2I100,000100,000COCO, Flickr30K, MagicBrushFLUX
Chameleon2024Social MediaLinkT/I2I14,86311,170UnsplashGANs, SDMs, DALL-E 2, GLIDE
DF402024FaceLinkT/I2I~1,100~1,000,000FF++, CDF, FFHQ, CelebASDMs, GANs, Midjourney, DDPM
FakeBench2024GeneralLinkT/I2I3,0003,00010 Public Datasets10 Generative Models
AI-Face2024FaceLinkT/I2I400,8851,245,6606 Public datasetsSDMs, GANs, Midjourney, IF
VideoWildDeepfake2021FaceLinkI.A2V3,8053,509Social MediaSocial Media
DiffHead2023FaceLinkI.A2V820-CREMADiffused Heads: build on DDPM
DVF2024GeneralLinkI/T2V2,7503,938Intervid , Youtube-8M8 Diffusion Models
GenVideo2024GeneralLinkI/T2V1,223,5111,078,838Kinetics-400 , Youku-mPLUG, MSR-VTT20 Generative Models
GenVidBench2025GeneralLinkI/T2V33,931110,400Vript , HDL-VG-130M8 Generative Models
PDID2024Social MediaLink---Social MediaSocial Media
AudioIn-the-Wild2022SpeechLinkT2A20.7 hours17.2 hoursSocial Media, Video Streaming PlatformsSocial Media, Video Streaming Platforms
LibriSeVoc2023SpeechLinkT2A13,20179,206LibriTTSDiffWave, WaveNet
SONAR2024SpeechLinkT2A-2,274LibriTTSOpenAI, Seed-TTS, AudioGen
ASVspoof 20242024SpeechLinkT/A2A~289,527~1,211,186MLS-English32 Manipulation Methods
Multi-modalDGM^42023NewsLinkT/I2T77,426152,574Visual NewsB-GST, StyleCLIP, HFGI
COCOFake2023GeneralLinkT/I2T113,287566,435COCOSDMs
AV-Deepfake1M2023FaceLinkT2A286,721860,039Voxceleb2VITS, YourTTS, TalkLip
2024GeneralLinkT2I~2,300,000~9,200,000LAION-400MSDMs, IF
M³A2024News-T2T/I/T/V/V/A/T
T/I/V/V/A2T
708,4256,566,38660 News OutletsLLaMA2, GPT-4, GLIDE, SD, Tango
LOKI2024GeneralLinkT2T/I/T/V/V/T
T/I/V/V/A2T
~9,000~9,00021 Public Datasets43 Generative Models
MMFakeBench2024Social MediaLinkT2T/I-~11,000MS-COCO, VisualNews, Reddit, FEVERGPT-3.5, SD-XL, DALL-E 3, Midjourney
Deepfake-Eval2024Social MediaLinkT2T/A/V3,3902,441Social MediaSocial Media
ILLUSION2025GeneralLinkT2A/I
I2I
139,7401,232,246CelebV-Text [158], COCO, MusicCaps , Social Media28 Generative Methods

:mag_right: Detection :fire:

:page_facing_up: Text


Pure Detection

text_pure Illustrations of pure detection methodologies for LAIM-generated text.

  ♣️ Easy Explainable Methods

        ▶️ Watermarking

         - Distillation-Resistant Watermarking for Model Protection in NLP Paper

         - Three bricks to consolidate watermarks for large language models Paper GitHub

         - Robust multi-bit natural language watermarking through invariant features Paper

         - Undetectable Watermarks for Language Models Paper

         - Robust distortion-free watermarks for language models Paper

         - Provable robust watermarking for ai-generated text Paper GitHub

         - A Private Watermark for Large Language Models Paper

        ▶️ Non-watermarking

         - Unraveling the mystery of artifacts in machine generated text Paper

         - Stylometric detection of ai-generated text in twitter timelines Paper

         - CoCo: Coherence-Enhanced Machine-Generated Text Detection Under Data Limitation With Contrastive Learning Paper

         - Beat LLMs at Their Own Game: Zero-Shot LLM-Generated Text Detection via Querying ChatGPT Paper GitHub

         - Who Wrote This? The Key to Zero-Shot LLM-Generated Text Detection Is GECScore Paper GitHub

  ♣️ Hard Explainable Methods

         - HowkGPT: Investigating the Detection of ChatGPT-generated University Student Homework through Context-Aware Perplexity Analysis Paper

         - GPTZero Tool

         - Detectgpt: Zero-shot machine-generated text detection using probability curvature Paper GitHub

         - Spotting LLMs With Binoculars: Zero-Shot Detection of Machine-Generated Text Paper GitHub

         - Multiscale Positive-Unlabeled Detection of AI-Generated Texts Paper GitHub

Beyond Detection

text_beyond Illustrations of beyond detection methodologies for LAIM-generated text.

  ♣️ Efficiency

         - Efficient Detection of LLM-generated Texts with a Bayesian Surrogate Model Paper

         - Fast-DetectGPT: Efficient Zero-Shot Detection of Machine-Generated Text via Conditional Probability Curvature Paper GitHub

         - DetectLLM: Leveraging Log Rank Information for Zero-Shot Detection of Machine-Generated Text Paper GitHub

         - SeqXGPT: Sentence-Level AI-Generated Text Detection Paper GitHub

         - Glimpse: Enabling White-Box Methods to Use Proprietary Models for Zero-Shot LLM-Generated Text Detection Paper GitHub

  ♣️ Attribution

         - TURINGBENCH: A Benchmark Environment for Turing Test in the Age of Neural Text Generation Paper Turingbench

         - Whodunit? Learning to Contrast for Authorship Attribution Paper

         - Through the looking glass: Learning to attribute synthetic text generated by language models Paper

         - TopRoBERTa: Topology-Aware Authorship Attribution of Deepfake Texts Paper

         - Authorship attribution for neural text generation Paper GitHub

         - Gpt-who: An information density-based machine-generated text detector Paper

         - LLMDet: A Third Party Large Language Models Generated Text Detection Tool Paper GitHub

         - Few-Shot Detection of Machine-Generated Text using Style Representations Paper

         - Origin Tracing and Detecting of LLMs Paper

  ♣️ Generalization

         - Ghostbuster: Detecting Text Ghostwritten by Large Language Models Paper

         - Conda: Contrastive domain adaptation for ai-generated text detection Paper GitHub

         - Text Fluoroscopy: Detecting LLM-Generated Text through Intrinsic Features Paper GitHub

         - DeTeCtive: Detecting AI-generated Text via Multi-Level Contrastive Learning Paper GitHub

         - Intrinsic Dimension Estimation for Robust Detection of AI-Generated Texts Paper GitHub

  ♣️ Interpretability

         - DNA-GPT: Divergent N-Gram Analysis for Training-Free Detection of GPT-Generated Text Paper GitHub

         - A Watermark for Large Language Models Paper GitHub

         - Chatgpt or human? detect and explain. explaining decisions of machine learning model for detecting short chatgpt-generated text Paper

         - Check Me If You Can: Detecting ChatGPT-Generated Academic Writing using CheckGPT Paper

         - Is chatgpt involved in texts? measure the polish ratio to detect chatgpt-generated text Paper

  ♣️ Robustness

        ▶️ Adversarial Attack Robustness

         - Red Teaming Language Model Detectors with Language Models Paper

         - Radar: Robust ai-text detection via adversarial learning Paper Project Page

         - J-guard: Journalism guided adversarially robust detection of ai-generated news Paper

         - Outfox: Llm-generated essay detection through in-context learning with adversarially generated examples Paper

        ▶️ LAIM-Polished Robustness

         - Is chatgpt involved in texts? measure the polish ratio to detect chatgpt-generated text Paper

  ♣️ Empirical Study

         - ChatLog: Recording and Analyzing ChatGPT Across Time Paper GitHub

         - On the Zero-Shot Generalization of Machine-Generated Text Detectors Paper

         - On the Generalization of Training-based ChatGPT Detection Methods Paper

         - Supervised Machine-Generated Text Detectors: Family and Scale Matters Paper GitHub

         - Deepfake Text Detection in the Wild Paper GitHub

         - How large language models are transforming machine-paraphrased plagiarism Paper

         - Paraphrase Detection: Human vs. Machine Content Paper

         - MGTBench: Benchmarking Machine-Generated Text Detection Paper GitHub

         - How close is chatgpt to human experts? comparison corpus, evaluation, and detection Paper GitHub

         - Can LLM-Generated Misinformation Be Detected? Paper GitHub

         - From Text to Source: Results in Detecting Large Language Model-Generated Content Paper

📸 Image


Pure Detection

image_pure Illustrations of pure detection methodologies for LAIM-generated image.

  ♣️ Physical/Physiological based Methods

         - Qualitative Failures of Image Generation Models and Their Application in Detecting Deepfakes Paper

         - Perspective (in) consistency of paint by text Paper

         - Lighting (in) consistency of paint by text Paper

  ♣️ Diffuser Fingerprints based Methods

         - Deep Image Fingerprint: Accurate And Low Budget Synthetic Image Detector Paper

         - DIRE for Diffusion-Generated Image Detection Paper GitHub

         - Exposing the Fake: Effective Diffusion-Generated Images Detection Paper

         - LaRE^2: Latent Reconstruction Error Based Method for Diffusion-Generated Image Detection Paper GitHub

         - Aligned Datasets Improve Detection of Latent Diffusion-Generated Images Paper GitHub

         - Manifold Induced Biases for Zero-shot and Few-shot Detection of Generated Images Paper GitHub

  ♣️ Spatial-based Methods

         - Rich and Poor Texture Contrast: A Simple yet Effective Approach for AI-generated Image Detection Paper Project Page

         - Unmasking The Artist: Discriminating Human-Drawn And AI-Generated Human Face Art Through Facial Feature Analysis Paper

         - Detecting images generated by deep diffusion models using their local intrinsic dimensionality Paper

  ♣️ Frequency-based Methods

         - Wavelet-packets for deepfake image analysis and detection Paper GitHub

         - AUSOME: authenticating social media images using frequency analysis Paper

         - AI-Generated Image Detection using a Cross-Attention Enhanced Dual-Stream Network Paper

         - Synthbuster: Towards Detection of Diffusion Model Generated Images Paper

         - Faster Than Lies: Real-time Deepfake Detection using Binary Neural Networks Paper GitHub

  ♣️ Distribution-based Methods

         - Zero-Shot Detection of AI-Generated Images Paper GitHub

Beyond Detection

image_beyond Illustrations of beyond detection methodologies for LAIM-generated image.

  ♣️ Attribution and Model Parsing

        ▶️ Attribution and Model Parsing

         - Level up the deepfake detection: a method to effectively discriminate images generated by gan architectures and diffusion models Paper

         - Reverse engineering of generative models: Inferring model hyperparameters from generated images Paper

  ♣️ Generalization

         - Online Detection of AI-Generated Images Paper

         - Towards universal fake image detectors that generalize across generative models Paper GitHub

         - Raising the Bar of AI-generated Image Detection with CLIP Paper

         - Transcending Forgery Specificity with Latent Space Augmentation for Generalizable Deepfake Detection Paper

         - Fingerprintnet: Synthesized fingerprints for generated image detection Paper

         - Detecting Deepfakes Without Seeing Any Paper GitHub

         - Improving Synthetically Generated Image Detection in Cross-Concept Settings Paper

         - Diffusion Noise Feature: Accurate and Fast Generated Image Detection Paper

         - Contrasting Deepfakes Diffusion via Contrastive Learning and Global-Local Similarities Paper GitHub

         - HRR: Hierarchical Retrospection Refinement for Generated Image Detection Paper

         - A Sanity Check for AI-generated Image Detection Paper GitHub

         - Stacking Brick by Brick: Aligned Feature Isolation for Incremental Face Forgery Detection Paper GitHub

         - A Bias-Free Training Paradigm for More General AI-generated Image Detection Paper GitHub

         - Breaking Semantic Artifacts for Generalized AI-generated Image Detection Paper GitHub

         - Dual Data Alignment Makes AI‑Generated Image Detector Easier Generalizable Paper

  ♣️ Interpretability

         - Interpretable-through-prototypes deepfake detection for diffusion models Paper GitHub

         - Did You Note My Palette? Unveiling Synthetic Images Through Color Statistics Paper

  ♣️ Localization

        ▶️ Fully-supervised

         - Hierarchical fine-grained image forgery detection and localization Paper GitHub

         - Perceptual Artifacts Localization for Image Synthesis Tasks Paper GitHub

         - TruFor: Leveraging all-round clues for trustworthy image forgery detection and localization Paper GitHub

         - UnionFormer: Unified-Learning Transformer with Multi-View Representation for Image Manipulation Detection and Localization Paper

        ▶️ Weakly-supervised

         - Weakly-supervised deepfake localization in diffusion-generated images Paper

  ♣️ Robustness

        ▶️ Adversarial Attack Robustness

         - D4: Detection of Adversarial Diffusion Deepfakes Using Disjoint Ensembles Paper

         - Exploring the Adversarial Robustness of CLIP for AI-generated Image Detection Paper

         - All Patches Matter, More Patches Better: Enhance AI‑Generated Image Detection via Panoptic Patch Learning Paper

        ▶️ Post-Processing Robustness

         - GLFF: Global and Local Feature Fusion for AI-synthesized Image Detection Paper

         - Exposing fake images generated by text-to-image diffusion models Paper

         - Local Statistics for Generative Image Detection Paper

  ♣️ Empirical Study

         - On the detection of synthetic images generated by diffusion models Paper GitHub

         - Intriguing properties of synthetic images: from generative adversarial networks to diffusion models Paper

         - Towards the detection of diffusion model deepfakes Paper

         - Unveiling the Impact of Image Transformations on Deepfake Detection: An Experimental Analysis Paper

         - On the use of Stable Diffusion for creating realistic faces: from generation to detection Paper

         - Finding AI-Generated Faces in the Wild Paper

         - Forensic analysis of synthetically generated western blot images Paper

         - Beyond Human Forgeries: An Investigation into Detecting Diffusion-Generated Handwriting Paper

🎞️ Video


Video Detection

Illustration of detection methodology in generalization task for LAIM-generated video.

Pure Detection

  ♣️ Spatial & Temporal based Methods

         - Distinguish Any Fake Videos: Unleashing the Power of Large-scale Data and Motion Features Paper

         - Exposing AI-generated Videos: A Benchmark Dataset and a Local-and-Global Temporal Defect Based Detection Method Paper

Beyond Detection

  ♣️ Generalization

         - Revisiting Generalizability in Deepfake Detection: Improving Metrics and Stabilizing Transfer Paper

  ♣️ Empirical Study

         - Beyond Deepfake Images: Detecting AI-Generated Videos Paper

🎵 Audio


Pure Detection

Audio Detection

The artifacts introduced by DM-based neural vocoders (WaveGrad and DiffWave) to a voice signal. The differences in mel-spectrograms between real and generated ones are illustrated in the third and fifth columns.

  ♣️ Vocoder-based

         - AI-Synthesized Voice Detection Using Neural Vocoder Artifacts Paper GitHub

Beyond Detection

  ♣️ Generalization

         - Improving Generalization for AI-Synthesized Voice Detection Paper GitHub

🍯 Multimodal


Pure Detection

Multimodal Detection

Illustrations of pure detection methodologies for LAIM-generated multimodal media.

  ♣️ Prompt-guided

         - Parents and Children: Distinguishing Multimodal DeepFakes from Natural Images Paper

         - On Learning Multi-Modal Forgery Representation for Diffusion Generated Video Detection Paper GitHub

         - Human Action CLIPS: Detecting AI-generated Human Motion Paper

  ♣️ Text-image Inconsistency

         - Detecting Cross-Modal Inconsistency to Defend Against Neural Fake News Paper GitHub

         - Exposing Text-Image Inconsistency Using Diffusion Models Paper

Beyond Detection

Multimodal Detection

Illustrations of beyond detection methodologies for LAIM-generated multimodal media.

  ♣️ Attribution

         - De-fake: Detection and attribution of fake images generated by text-to-image generation models Paper

         - FIDAVL: Fake Image Detection and Attribution using Vision-Language Model Paper GitHub

  ♣️ Generalization

        ▶️ Prompt Tuning

         - AntifakePrompt: Prompt-Tuned Vision-Language Models are Fake Image Detectors Paper GitHub

        ▶️ Contrastive Learning

         - Generalizable Synthetic Image Detection via Language-guided Contrastive Learning Paper GitHub

  ♣️ Interpretability

         - Combating Misinformation in the Era of Generative AI Models Paper

         - FFAA: Multimodal Large Language Model based Explainable Open-World Face Forgery Analysis Assistant Paper GitHub

         - X^2-DFD: A framework for eXplainable and eXtendable deepfake detection Paper

  ♣️ Localization

        ▶️ Spatial-based

         - Detecting and grounding multi-modal media manipulation Paper

         - Exploiting Modality-Specific Features For Multi-Modal Manipulation Detection And Grounding Paper

        ▶️ Frequency-based

         - Unified Frequency-Assisted Transformer Framework for Detecting and Grounding Multi-Modal Manipulation Paper

        ▶️ MLLM-based

         - FakeShield: Explainable Image Forgery Detection and Localization via Multi-modal Large Language Models Paper GitHub

         - ForgeryGPT: Multimodal Large Language Model For Explainable Image Forgery Detection and Localization Paper

  ♣️ Empirical Study

         - Detecting Images Generated by Diffusers Paper GitHub

         - CLIPping the Deception: Adapting Vision-Language Models for Universal Deepfake Detection Paper

         - VERITE: a Robust benchmark for multimodal misinformation detection accounting for unimodal bias Paper GitHub

         - Can ChatGPT Detect DeepFakes? A Study of Using Multimodal Large Language Models for Media Forensics Paper

Detection Tools

ModalityToolCompanyLinkTypeOpen SourceCost
TextAI Content DetectorCopyleaksLinkWebapp & APILimited free usage
AI Content Detector, ChatGPT detectorZeroGPTLinkWebapp & APIFree usage
AI DetectorGPTZeroLinkMulti-platformLimited free usage
AI Content DetectorWinston AILinkWebapp & APILimited free usage
AI Content DetectorCrossplagLinkWebappLimited free usage
Giant Language model Test RoomGLTRLinkWebappFree usage
The AI DetectorBrandwellLinkWebappFree usage
AI CheckerOriginality aiLinkWebapp & APILimited free usage
Advanced AI Detector and HumanizerUndetectable aiLinkWebapp & APILimited free usage
AI Content DetectorWriterLinkWebapp & APILimited free usage
AI Content DetectorConchLinkWebappLimited free usage
Illuminarty TextIlluminartyLinkWebapp & APILimited free usage
AI-Generated Text DetectorIs it AILinkWebapp & APILimited free usage
ImageLiveness Detection, Facial RecognitionIncodeLinkMulti-platformPaid
AI or Not imageAI or NotLinkWebapp & APILimited free usage
AI-Generated Image DetectorIs it AILinkWebapp & APILimited free usage
Illuminarty ImageIlluminartyLinkWebapp & APILimited free usage
AI Image DetectorUndetectable aiLinkWebapp & APILimited free usage
SynthIDGoogleLinkWebappFree usage
The AI image detectorWinstonLinkWebapp & APILimited free usage
Advanced AI Image DetectorBrandwellLinkWebappLimited free usage
VideoDeepware ScannerDeepwareLinkWebapp & APIFree usage
Attestiv Deepfake Video DetectionAttestivLinkWebapp & APILimited free usage
AudioPulse InspectPindropLinkMulti-platformPaid
AI Voice DetectorAI Voice DetectorLinkWebapp & APILimited free usage
AI Speech ClassifierElevenLabsLinkWebapp & APILimited free usage
AI or Not audioAI or NotLinkMulti-platformLimited free usage
Multi-modalVideo, Image, and Audio DetectorDeep MediaLinkMulti-platformLimited free usage
Deepfake DetectionSensity AILinkMulti-platformPaid
Hive AI’s Deepfake Detection APIHive AILinkAPILimited free usage
Resemble DetectResemble AILinkWebapp & APILimited free usage
DuckDuckGoose AI (Phocus)DuckDuckGoose AILinkWebappPaid
SentinelSentinelLinkWebappPaid
Deepfake DetectorDeepfake DetectorLinkMulti-platformFree usage
DeepFake-o-meterU of BuffaloLinkWebappFree usage
BioIDBioIDLinkWebapp & APILimited free usage
Get Real ProtectGet RealLinkMulti-platformPaid
Reality DefenderReality DefenderLinkMulti-platformPaid

Generation Works

WorksTimeModalityLinks
T5Q4 2019TextLink
GPT-3Q2 2020TextLink
Wave-Grad2Q1 2021AudioLink
PanGuQ2 2021TextLink
LDMsQ4 2021ImageLink
GLIDEQ4 2021ImageLink
ImagenQ2 2022ImageLink
PaLMQ2 2022TextLink
OPTQ2 2022TextLink
Make-A-VideoQ3 2022VideoLink
GLMQ3 2022TextLink
HuggingGPTQ3 2022MultimodalLink
WhisperQ3 2022AudioLink
ChatGPTQ4 2022TextLink
DALL-E 2Q4 2022ImageLink
SDQ4 2022ImageLink
mT0Q4 2022TextLink
BLOOMQ4 2022TextLink
Make-An-AudioQ1 2023AudioLink
GPT-4Q1 2023MultimodalLink
BardQ1 2023TextLink
LLaMAQ1 2023TextLink
GEN-1Q1 2023VideoLink
ImageRewardQ2 2023ImageLink
PaLM2Q2 2023TextLink
CodeGen2Q2 2023TextLink
IFQ2 2023ImageLink
VideoGenQ3 2023VideoLink
DALL-E 3Q3 2023ImageLink
LLaMA 2Q3 2023TextLink
GeminiQ4 2023TextLink
Emu EditQ4 2023ImageLink
Emu VideoQ4 2023VideoLink
TitanQ4 2023ImageLink
Stable VideoQ4 2023VideoLink
MidjourneyV6Q4 2023ImageLink
Imagen 2Q4 2023ImageLink
Claude 3.5Q1 2024MultimodalLink
aMUSEdQ1 2024ImageLink
Synthesia 2Q2 2024VideoLink
MultiBoothQ2 2024ImageLink
GPT-4oQ2 2024MultimodalLink
LLaMA 3Q2 2024MultimodalLink
GLM-4Q2 2024MultimodalLink
CustomCrafterQ3 2024VideoLink
MegaFusionQ3 2024ImageLink
Qwen 2Q3 2024MultimodalLink
Tri-ErgonQ4 2024AudioLink
Veo 2Q4 2024VideoLink
SoraQ4 2024VideoLink
AudioXQ1 2025AudioLink
Grok 3Q1 2025MultimodalLink
DeepSeek-V3Q1 2025MultimodalLink
Gemini 2.5 ProQ1 2025MultimodalLink
LLaMA 4Q2 2025MultimodalLink
Qwen 3Q2 2025MultimodalLink

Detection Works

WorksTimeModalityLinks
LinguisticQ4 2020TextLink
XLNet-FTQ1 2021TextLink
Turing-BenchQ4 2021TextLink
UnravelingQ2 2022TextLink
WaveletQ3 2022ImageLink
WhodunitQ3 2022TextLink
De-FakeQ4 2022MultimodalLink
TowardsQ4 2022ImageLink
TruForQ4 2022ImageLink
DIREQ1 2023ImageLink
GPTZeroQ1 2023TextLink
DetectGPTQ1 2023TextLink
HAMMERQ2 2023MultimodalLink
DetectVocoderQ2 2023AudioLink
SeDIDQ3 2023ImageLink
RADARQ3 2023TextLink
OUTFOXQ3 2023ImageLink
SeqXGPTQ4 2023TextLink
RevisitVideoQ4 2023VideoLink
RaisingQ4 2023ImageLink
BinocularsQ1 2024TextLink
AI FaceQ2 2024ImageLink
DuB3DQ2 2024VideoLink
GECScoreQ2 2024TextLink
FFAAQ3 2024MultimodalLink
BreakingQ4 2024ImageLink
B-FreeQ4 2024ImageLink
ForgeryGPTQ4 2024MultimodalLink
FakeShieldQ4 2024MultimodalLink
GenVidBenchQ1 2025VideoLink

Contributors

Purdue-M2

80 commits

ll-cqu

56 commits

gneeraj97

18 commits

discovershu

4 commits