本仓库由 OpenClaw (🦞) 和人类共同维护

🔍 A comprehensive collection of papers, datasets, and tools for AI-Generated Content (AIGC) detection research.
This repository tracks research on detecting AI-generated content across image, video, text, audio, and multimodal settings. Paper, dataset, tool, README, and timeline metadata are generated from structured files under data/.
Recent image/video entries are cross-referenced with ant-research/Awesome-AIGC-Image-Video-Detection.
🕐 Interactive Timeline · Local HTML · 📊 View as Table
git clone https://github.com/MuskAI/Awesome-AIGC-Detection.git
cd Awesome-AIGC-Detection
python3 scripts/validate_data.py
python3 scripts/generate_readme.py --check
| Category | Count |
|---|---|
| Total Papers | 151 |
| Datasets | 44 |
| News Items | 14 |
| Timespan | 2020 - 2026 |
| Image Detection | 116 |
| Video Detection | 30 |
| Text Detection | 3 |
| Audio Detection | 1 |
| Multimodal Detection | 1 |
| Venue | Papers |
|---|---|
| arXiv | 66 |
| CVPR | 26 |
| ICLR | 10 |
| ECCV | 8 |
| ICCV | 5 |
| ICML | 5 |
| NeurIPS | 4 |
| CVPRW | 4 |
| ICCVW | 2 |
| ICASSP | 2 |
| WACV | 2 |
| TMM | 2 |
Fresh signals from platforms, policy, provenance, and industry deployment, with China-specific updates pulled into a dedicated watch lane.
2026-05-18 · Industry / Identity
GetReal Security launches real-time identity verification and deepfake detection platform — PR Newswire / GetReal Security
Signal: GetReal announced general availability of GetReal Protect, combining multimodal deepfake detection with continuous identity verification for voice and video workflows.
Why it matters: AIGC detection is moving from offline forensic review into live enterprise authentication and incident response.
Coverage mix: Platform x3, China Policy x3, Industry x2, Policy x2, Provenance x1, China Industry x1, Research x1, China Platform x1. The current flow spans domestic governance and global deployment: platform-scale likeness search, capture-time provenance, labeling standards, review workflows, and real-time fraud defense.
Domestic updates are especially useful for tracking how AIGC detection becomes platform policy, labeling infrastructure, rights protection, and video governance.
2026-04-30 · China Policy / AI Governance 中央网信办部署开展“清朗·整治AI应用乱象”专项行动 — 中国网信网 Signal: 中央网信办启动为期4个月的专项行动,重点整治生成合成内容标识落实不到位、AI数据投毒、AI换脸拟声滥用等问题。 Why it matters: 国内治理重点已经从单点内容处置扩展到源头备案、安全审核、标识互认和检测能力建设。
2026-04-08 · China Policy / Video Governance 国家广播电视总局通报“AI魔改”视频治理工作取得实效 — 国家广播电视总局 Signal: 广电总局通报专项治理结果,清理大量基于经典电视剧作品进行AI魔改的违规视频,并推动平台建立常态化治理机制。 Why it matters: AIGC视频检测在国内内容平台中正从研究问题变成审核、版权、文化安全和平台责任问题。
2026-04-02 · China Industry / Likeness Rights 中广联演员委员会就AI换脸合成、声纹克隆、影视素材魔改发布声明 — 澎湃新闻 Signal: 中国广电联合会演员委员会针对AI换脸、声纹克隆、影视素材魔改和未经授权采集训练数据等侵权行为发布声明。 Why it matters: 国内演艺行业把检测和溯源需求推向肖像权、声音权、训练数据授权和二创边界。
2026-02-12 · China Platform / Labeling 小红书要求AI生成合成内容主动标识,未标识内容将限制推荐 — 新浪科技 Signal: 小红书公告称将持续加强AI生成合成内容识别检测能力,未主动标识的AI生成合成内容将由平台添加标识并限制分发。 Why it matters: 国内社区平台正在把检测结果直接接入分发、补标、申诉和举报机制。
2025-09-01 · China Policy / Labeling Standard 《人工智能生成合成内容标识办法》正式施行 — 央视网 Signal: 国家网信办等四部门发布的标识办法正式施行,要求AI生成的文字、图片、音频、视频和虚拟场景等内容明确标识。 Why it matters: 国内检测生态有了制度入口,显式标识、隐式标识、元数据和传播端风险提示会影响后续工具与评测。
2026-05-15 · Platform / Likeness YouTube expands AI likeness detection to all adult users — The Verge Signal: The Verge reported that YouTube is expanding its likeness detection program to users aged 18 or older with a YouTube account. Why it matters: Consumer-scale enrollment turns deepfake discovery into a mainstream safety feature and raises new questions about false matches and review workflows.
2026-05-11 · Provenance / Camera Canon introduces C2PA-compliant Authenticity Imaging System for news organizations — Canon Global Signal: Canon announced a C2PA-based system for supported cameras that preserves provenance from capture through newsroom workflows. Why it matters: Secure capture and provenance metadata are becoming a practical complement to detector-only pipelines.
2026-04-30 · Industry / Fraud Sumsub launches Adaptive Deepfake Detector for fraud prevention — PR Newswire / Sumsub Signal: Sumsub released an upgraded deepfake detection system focused on fast model adaptation and multi-signal fraud analysis. Why it matters: The fraud setting emphasizes continual learning, liveness, device signals, and session context beyond media pixels alone.
2026-04-21 · Platform / Likeness YouTube expands likeness detection to the entertainment industry — YouTube Official Blog Signal: YouTube expanded its AI likeness detection access to talent agencies, management companies, and celebrities. Why it matters: Large platforms are operationalizing face-likeness detection as a rights and removal workflow, not only as a research benchmark.
2026-03-26 · Policy / Market UK DSIT publishes analysis of the deepfake detection technology market — GOV.UK Signal: The UK Department for Science, Innovation and Technology published a market and evidence review for deepfake detection technology. Why it matters: Policy interest is shifting toward supply, demand, barriers, and future growth of detection capabilities.
2026-03-10 · Platform / Public Interest YouTube expands likeness detection to civic leaders and journalists — YouTube Official Blog Signal: YouTube began piloting likeness detection access for government officials, journalists, and political candidates. Why it matters: High-risk public figures are a key deployment target for deepfake discovery, identity verification, and takedown workflows.
2026-02-19 · Research / Provenance Microsoft publishes study on media integrity and authentication methods — Microsoft Signal Signal: Microsoft discussed a media integrity report covering provenance, watermarking, fingerprinting, and their limitations. Why it matters: The field is converging on layered authenticity signals instead of relying on a single detector score.
2026-02-05 · Policy / Evaluation UK government announces deepfake detection evaluation framework — GOV.UK Signal: The UK government announced work with technology companies, academics, and experts to evaluate deepfake detection tools against real-world threats. Why it matters: Benchmarking is becoming a public-sector requirement, especially for fraud, abuse, impersonation, and safety use cases.
This repository is curated as a map, not just a shelf: use the routes below to jump from problem type to representative papers, code, and benchmarks.
| Lane | Signal | Representative Entries |
|---|---|---|
| Latest frontier | 34 papers in 2026 | DiffSeg30k: A Multi-Turn Diffusion Editing Benchmark for…; DINO-Detect: A Simple yet Effective Framework for Blur-Ro…; IVY-FAKE: A Unified Explainable Framework and Benchmark f… |
| Reproducible work | 75 papers with code/project links | ForensicHub: A Unified Benchmark & Codebase for All-Domai…; AgentFoX: LLM Agent-Guided Fusion with eXplainability for…; Automated In-the-Wild Data Collection for Continual AI Ge… |
| Benchmark-first reading | 44 datasets and benchmarks | ActivityForensics; AIGI-Now; AIGVDBench |
| Area | 2020 | 2021 | 2022 | 2023 | 2024 | 2025 | 2026 |
|---|---|---|---|---|---|---|---|
| Image | 4 | 1 | 2 | 14 | 31 | 37 | 27 |
| Video | - | 1 | 1 | 3 | 9 | 12 | 4 |
| Text | - | - | - | - | - | 1 | 2 |
| Audio | - | - | - | - | - | 1 | - |
| Multimodal | - | - | - | - | - | - | 1 |
| Year | Dataset | Real | Fake | Scale | Real Sources | Generation Methods |
|---|---|---|---|---|---|---|
| 2020 | CNNSpot | 362,000 | 362,000 | - | LSUN, ImageNet, CelebA, COCO... | ProGAN, StyleGAN, BigGAN, CRN, SITD... |
| 2023 | DiffusionDB | 3,300,000 | 16,000,000 | - | DiscordChatExporter | SD |
| 2023 | DMimage | 200,000 | 200,000 | - | COCO, LSUN | LDM |
| 2023 | Fake2M | - | 2,300,000 | - | CC3M | SD-V1.5, IF, StyleGAN3 |
| 2023 | GenImage | 1,331,167 | 1,350,000 | - | ImageNet | SDMs, Midjourney, BigGAN |
| 2024 | AIGCDetectBenchmark | - | - | 100K | - | - |
| 2024 | DF40 | - | - | 0.1M+ videos, 1M+ images | - | - |
| 2024 | DRCT | - | - | 2M | MSCOCO | LDM, SDv1.4, SDv1.5, SDv2, SDXL, SD-ControlNet |
| 2024 | GenVideo | - | - | 2.3M | - | - |
| 2024 | WildFake | 2,557,278 | 1,013,446 | - | ImageNet, LAION, Wukong, COCO... | BigGAN, StyleGAN, StarGAN, Midjourney, DALL-E... |
| 2024 | WildRF | - | - | - | Reddit, X (Twitter), Facebook (real images) | Reddit, X (Twitter), Facebook (social media deepfakes) |
| 2025 | AEGIS | - | - | 10K+ | Vript (YouTube, TikTok), DVF, YouTube (self-collected) | Stable Video Diffusion, CogVideoX-5B, I2VGen-XL, Pika, KLing, Sora |
| 2025 | AIGIBench | - | - | 200K | FFHQ, CelebA-HQ, Open Images V7 | Common generators & SocialRF, CommunityAI |
| 2025 | ARForensics | - | - | 300k | ImageNet | Infinity, Janus_Pro, RAR, Switti, VAR, LlamaGen, Open_MAGVIT2 |
| 2025 | Chameleon | - | - | 26K | Unsplash | Midjourney, DALLE-3, Stable Diffusion (various LoRA fine-tuned) |
| 2025 | Community Forensics | - | - | 2.7M | LAION, ImageNet, COCO, FFHQ, CelebA, MetFaces, AFHQ, etc. | 4803 generators (Latent Diffusion, GAN, Autoregressive, Pixel Diffusion, Commercial) |
| 2025 | DDL | - | - | 367K | - | - |
| 2025 | DiffSeg30k | - | - | 30K | COCO | SD2, SD3.5, SDXL, Flux.1, Glide, Kolors, HunyuanDiT1.1, Kandinsky 2.2 |
| 2025 | FakeClue | - | - | 100K | - | - |
| 2025 | FakeParts | - | - | 81K | - | - |
| 2025 | ForensicHub | - | - | 23 datasets ; 42 models | - | ProGAN, StyleGAN, LDM, SDv1.4, SDv1.5, SDv2, SDXL, SD-ControlNet, MidJourney, ADM, GLIDE, VQDM, BigGAN |
| 2025 | Forensics-Bench | - | - | 63K | Various public datasets | GAN, Diffusion, VAE, RNN, Encoder-Decoder, Graphics-based |
| 2025 | GenBuster | - | - | 200K | - | - |
| 2025 | GenBuster++ | - | - | 4K | - | - |
| 2025 | Ivy-Fake | - | - | 150K | - | - |
| 2025 | LOKI | - | - | 18K | - | SORA, Keling, Open-Sora, FLUX, Midjourney, Stable Diffusion, Nerf-based, Gaussian-based, GPT-4o, Qwen-Max, Llama 3.1-405B, MusicGen, AudioLDM2... |
| 2025 | NeXT-IMDL | - | - | 558K | Flickr30k, COCO, OpenImages V7 | SD2-Inpainting, SDXL-Inpainting, FLUX-Inpainting, etc. |
| 2025 | OpenFake | - | - | ~4M | LAION-400M | SD 1.5/2.1/XL/3.5, Flux 1.0-dev/1.1-Pro/Schnell, Midjourney v6/v7, DALL·E 3, Imagen 3/4, GPT Image 1, Ideogram 3.0, Grok-2, HiDream-I1, Recraft v3, Chroma, and 10 community LoRA/finetune variants |
| 2025 | OpenSDI | - | - | 300K | Megalith-10M | SD1.5, SD2.1, SDXL, SD3, Flux.1 |
| 2025 | RewardData | - | - | 4.3K | - | - |
| 2025 | So-Fake-Set | - | - | 2M+ | F30k, WIDER, FFHQ, CelebA, OpenImages, COCO, OpenForensics | Qwen-image, GPT-4o, Nano Banana, Seedream3.0, Ideogram3.0, etc. |
| 2025 | Video Reality Test | - | - | 149 real + dynamic fake | YouTube ASMR (social media) | Veo3.1-Fast, Sora2, Wan2.2-A14B, Wan2.2-5B, OpenSora-V2, HunyuanVideo, StepVideo |
| 2025 | XAIGID-RewardBench | - | - | 3K | COCO-2017 | Imagen 4, Flux.1 Dev, Bagel, etc. |
| 2026 | ActivityForensics | - | - | 6K | - | - |
| 2026 | AIGI-Now | - | - | 18K | COCO | Nano Banana, GPT-4o, Jimeng, Kling, Minimax, etc. |
| 2026 | AIGVDBench | - | - | 440k | OpenVid-HD | 31 generation models |
| 2026 | BR-Gen | - | - | 150K | - | - |
| 2026 | GenVidBench | - | - | 6M | - | - |
| 2026 | HiResolution | - | - | 50K | - | - |
| 2026 | HydraFake | - | - | 100K | FFHQ, VFHQ, CelebAHQ, FF++, etc. | GPT-4o, HailuoAI, ICLight, InfiniteYou, etc. |
| 2026 | MintVid | - | - | 4K | OpenVid, VFHQ, HDTF, TikTok | Jimeng3.0-Pro, Seedance, Kling2.5-Turbo, Sora2, TikTok, Youtube, etc. |
| 2026 | RealChain | - | - | 14K | - | - |
| 2026 | SciFigDetect | - | - | 150K | - | Nano Banana Pro, GPT-image-1.5 |
| 2026 | Skyra | - | - | 4K | - | - |
| Tool | Language | Platform | Description |
|---|---|---|---|
| Fake Image Detector | - | Chrome/Firefox | Browser extension for fake image detection |
| Hive AI Detector | - | Chrome | Browser extension for AI content detection |
| Tool | Language | Platform | Description |
|---|---|---|---|
| DeepFake-Official | Python | - | Face swapping detection resources |
| FaceForensics | Python | - | Face manipulation detection dataset and tooling |
| FakerLab | Python | - | AIGC detection benchmark |
| Tool | Language | Platform | Description |
|---|---|---|---|
| AI or Not | - | Web | AI detection service |
| EyeSift | - | Web | Free online AI text/image/video/audio detector with detailed per-model benchmarks |
| Hive AI Detection | - | Web/API | Video and image detection |
| Hive Moderation | - | Web | Website |
| Illuminarty | - | Web | Website |
| Illuminet | - | Web | AI image verification |
| Is it AI? | - | Web | Website |
| Tencent Zhuque AI Detection Assistant | - | Web | Website |
| TruthScan | - | Web | Website |
| Winston AI | - | Web | Website |
| Tool | Language | Platform | Description |
|---|---|---|---|
| Content Credentials | - | Web | C2PA provenance and content credential verification |
| SiliconSignature | - | GitHub | Hardware-bound image authentication and provenance certification using ASIC PoW nonces |
Resources here are useful context for authenticity, misinformation, and provenance work, but they are intentionally kept out of the core AIGC detection paper list.
2026.05 · Multimodal fake news / misinformation detection The AI Slop Intelligence Dashboard Problem (James Sawyer Field Notes) Why adjacent: Related to authenticity and misinformation, but the primary task is critical analysis of AI-generated intelligence dashboard claims rather than direct AIGC content detection. Links: Resource
2025.04 · Multimodal fake news / misinformation detection Exploring Modality Disruption in Multimodal Fake News Detection (arxiv) Why adjacent: Related to authenticity and misinformation, but the primary task is not direct AIGC content detection. Links: Resource
2024.01 · Multimodal fake news / misinformation detection MiRAGeNews: Multimodal Realistic AI-Generated News Detection (ACL) Why adjacent: Related to authenticity and misinformation, but the primary task is not direct AIGC content detection. Links: Resource
Contributions are welcome. To add or update entries:
data/.python3 scripts/validate_data.py.python3 scripts/generate_readme.py.Please keep placeholder links such as https://arxiv.org/abs/ out of data/; leave incomplete candidates in PAPERS_TO_ADD.md until metadata is complete.
If you have questions or suggestions, please open an issue.
Last generated on 2026.05.19
Python
100.0%
本仓库由 OpenClaw (🦞) 和人类共同维护

🔍 A comprehensive collection of papers, datasets, and tools for AI-Generated Content (AIGC) detection research.
This repository tracks research on detecting AI-generated content across image, video, text, audio, and multimodal settings. Paper, dataset, tool, README, and timeline metadata are generated from structured files under data/.
Recent image/video entries are cross-referenced with ant-research/Awesome-AIGC-Image-Video-Detection.
🕐 Interactive Timeline · Local HTML · 📊 View as Table
git clone https://github.com/MuskAI/Awesome-AIGC-Detection.git
cd Awesome-AIGC-Detection
python3 scripts/validate_data.py
python3 scripts/generate_readme.py --check
| Category | Count |
|---|---|
| Total Papers | 151 |
| Datasets | 44 |
| News Items | 14 |
| Timespan | 2020 - 2026 |
| Image Detection | 116 |
| Video Detection | 30 |
| Text Detection | 3 |
| Audio Detection | 1 |
| Multimodal Detection | 1 |
| Venue | Papers |
|---|---|
| arXiv | 66 |
| CVPR | 26 |
| ICLR | 10 |
| ECCV | 8 |
| ICCV | 5 |
| ICML | 5 |
| NeurIPS | 4 |
| CVPRW | 4 |
| ICCVW | 2 |
| ICASSP | 2 |
| WACV | 2 |
| TMM | 2 |
Fresh signals from platforms, policy, provenance, and industry deployment, with China-specific updates pulled into a dedicated watch lane.
2026-05-18 · Industry / Identity
GetReal Security launches real-time identity verification and deepfake detection platform — PR Newswire / GetReal Security
Signal: GetReal announced general availability of GetReal Protect, combining multimodal deepfake detection with continuous identity verification for voice and video workflows.
Why it matters: AIGC detection is moving from offline forensic review into live enterprise authentication and incident response.
Coverage mix: Platform x3, China Policy x3, Industry x2, Policy x2, Provenance x1, China Industry x1, Research x1, China Platform x1. The current flow spans domestic governance and global deployment: platform-scale likeness search, capture-time provenance, labeling standards, review workflows, and real-time fraud defense.
Domestic updates are especially useful for tracking how AIGC detection becomes platform policy, labeling infrastructure, rights protection, and video governance.
2026-04-30 · China Policy / AI Governance 中央网信办部署开展“清朗·整治AI应用乱象”专项行动 — 中国网信网 Signal: 中央网信办启动为期4个月的专项行动,重点整治生成合成内容标识落实不到位、AI数据投毒、AI换脸拟声滥用等问题。 Why it matters: 国内治理重点已经从单点内容处置扩展到源头备案、安全审核、标识互认和检测能力建设。
2026-04-08 · China Policy / Video Governance 国家广播电视总局通报“AI魔改”视频治理工作取得实效 — 国家广播电视总局 Signal: 广电总局通报专项治理结果,清理大量基于经典电视剧作品进行AI魔改的违规视频,并推动平台建立常态化治理机制。 Why it matters: AIGC视频检测在国内内容平台中正从研究问题变成审核、版权、文化安全和平台责任问题。
2026-04-02 · China Industry / Likeness Rights 中广联演员委员会就AI换脸合成、声纹克隆、影视素材魔改发布声明 — 澎湃新闻 Signal: 中国广电联合会演员委员会针对AI换脸、声纹克隆、影视素材魔改和未经授权采集训练数据等侵权行为发布声明。 Why it matters: 国内演艺行业把检测和溯源需求推向肖像权、声音权、训练数据授权和二创边界。
2026-02-12 · China Platform / Labeling 小红书要求AI生成合成内容主动标识,未标识内容将限制推荐 — 新浪科技 Signal: 小红书公告称将持续加强AI生成合成内容识别检测能力,未主动标识的AI生成合成内容将由平台添加标识并限制分发。 Why it matters: 国内社区平台正在把检测结果直接接入分发、补标、申诉和举报机制。
2025-09-01 · China Policy / Labeling Standard 《人工智能生成合成内容标识办法》正式施行 — 央视网 Signal: 国家网信办等四部门发布的标识办法正式施行,要求AI生成的文字、图片、音频、视频和虚拟场景等内容明确标识。 Why it matters: 国内检测生态有了制度入口,显式标识、隐式标识、元数据和传播端风险提示会影响后续工具与评测。
2026-05-15 · Platform / Likeness YouTube expands AI likeness detection to all adult users — The Verge Signal: The Verge reported that YouTube is expanding its likeness detection program to users aged 18 or older with a YouTube account. Why it matters: Consumer-scale enrollment turns deepfake discovery into a mainstream safety feature and raises new questions about false matches and review workflows.
2026-05-11 · Provenance / Camera Canon introduces C2PA-compliant Authenticity Imaging System for news organizations — Canon Global Signal: Canon announced a C2PA-based system for supported cameras that preserves provenance from capture through newsroom workflows. Why it matters: Secure capture and provenance metadata are becoming a practical complement to detector-only pipelines.
2026-04-30 · Industry / Fraud Sumsub launches Adaptive Deepfake Detector for fraud prevention — PR Newswire / Sumsub Signal: Sumsub released an upgraded deepfake detection system focused on fast model adaptation and multi-signal fraud analysis. Why it matters: The fraud setting emphasizes continual learning, liveness, device signals, and session context beyond media pixels alone.
2026-04-21 · Platform / Likeness YouTube expands likeness detection to the entertainment industry — YouTube Official Blog Signal: YouTube expanded its AI likeness detection access to talent agencies, management companies, and celebrities. Why it matters: Large platforms are operationalizing face-likeness detection as a rights and removal workflow, not only as a research benchmark.
2026-03-26 · Policy / Market UK DSIT publishes analysis of the deepfake detection technology market — GOV.UK Signal: The UK Department for Science, Innovation and Technology published a market and evidence review for deepfake detection technology. Why it matters: Policy interest is shifting toward supply, demand, barriers, and future growth of detection capabilities.
2026-03-10 · Platform / Public Interest YouTube expands likeness detection to civic leaders and journalists — YouTube Official Blog Signal: YouTube began piloting likeness detection access for government officials, journalists, and political candidates. Why it matters: High-risk public figures are a key deployment target for deepfake discovery, identity verification, and takedown workflows.
2026-02-19 · Research / Provenance Microsoft publishes study on media integrity and authentication methods — Microsoft Signal Signal: Microsoft discussed a media integrity report covering provenance, watermarking, fingerprinting, and their limitations. Why it matters: The field is converging on layered authenticity signals instead of relying on a single detector score.
2026-02-05 · Policy / Evaluation UK government announces deepfake detection evaluation framework — GOV.UK Signal: The UK government announced work with technology companies, academics, and experts to evaluate deepfake detection tools against real-world threats. Why it matters: Benchmarking is becoming a public-sector requirement, especially for fraud, abuse, impersonation, and safety use cases.
This repository is curated as a map, not just a shelf: use the routes below to jump from problem type to representative papers, code, and benchmarks.
| Lane | Signal | Representative Entries |
|---|---|---|
| Latest frontier | 34 papers in 2026 | DiffSeg30k: A Multi-Turn Diffusion Editing Benchmark for…; DINO-Detect: A Simple yet Effective Framework for Blur-Ro…; IVY-FAKE: A Unified Explainable Framework and Benchmark f… |
| Reproducible work | 75 papers with code/project links | ForensicHub: A Unified Benchmark & Codebase for All-Domai…; AgentFoX: LLM Agent-Guided Fusion with eXplainability for…; Automated In-the-Wild Data Collection for Continual AI Ge… |
| Benchmark-first reading | 44 datasets and benchmarks | ActivityForensics; AIGI-Now; AIGVDBench |
| Area | 2020 | 2021 | 2022 | 2023 | 2024 | 2025 | 2026 |
|---|---|---|---|---|---|---|---|
| Image | 4 | 1 | 2 | 14 | 31 | 37 | 27 |
| Video | - | 1 | 1 | 3 | 9 | 12 | 4 |
| Text | - | - | - | - | - | 1 | 2 |
| Audio | - | - | - | - | - | 1 | - |
| Multimodal | - | - | - | - | - | - | 1 |
| Year | Dataset | Real | Fake | Scale | Real Sources | Generation Methods |
|---|---|---|---|---|---|---|
| 2020 | CNNSpot | 362,000 | 362,000 | - | LSUN, ImageNet, CelebA, COCO... | ProGAN, StyleGAN, BigGAN, CRN, SITD... |
| 2023 | DiffusionDB | 3,300,000 | 16,000,000 | - | DiscordChatExporter | SD |
| 2023 | DMimage | 200,000 | 200,000 | - | COCO, LSUN | LDM |
| 2023 | Fake2M | - | 2,300,000 | - | CC3M | SD-V1.5, IF, StyleGAN3 |
| 2023 | GenImage | 1,331,167 | 1,350,000 | - | ImageNet | SDMs, Midjourney, BigGAN |
| 2024 | AIGCDetectBenchmark | - | - | 100K | - | - |
| 2024 | DF40 | - | - | 0.1M+ videos, 1M+ images | - | - |
| 2024 | DRCT | - | - | 2M | MSCOCO | LDM, SDv1.4, SDv1.5, SDv2, SDXL, SD-ControlNet |
| 2024 | GenVideo | - | - | 2.3M | - | - |
| 2024 | WildFake | 2,557,278 | 1,013,446 | - | ImageNet, LAION, Wukong, COCO... | BigGAN, StyleGAN, StarGAN, Midjourney, DALL-E... |
| 2024 | WildRF | - | - | - | Reddit, X (Twitter), Facebook (real images) | Reddit, X (Twitter), Facebook (social media deepfakes) |
| 2025 | AEGIS | - | - | 10K+ | Vript (YouTube, TikTok), DVF, YouTube (self-collected) | Stable Video Diffusion, CogVideoX-5B, I2VGen-XL, Pika, KLing, Sora |
| 2025 | AIGIBench | - | - | 200K | FFHQ, CelebA-HQ, Open Images V7 | Common generators & SocialRF, CommunityAI |
| 2025 | ARForensics | - | - | 300k | ImageNet | Infinity, Janus_Pro, RAR, Switti, VAR, LlamaGen, Open_MAGVIT2 |
| 2025 | Chameleon | - | - | 26K | Unsplash | Midjourney, DALLE-3, Stable Diffusion (various LoRA fine-tuned) |
| 2025 | Community Forensics | - | - | 2.7M | LAION, ImageNet, COCO, FFHQ, CelebA, MetFaces, AFHQ, etc. | 4803 generators (Latent Diffusion, GAN, Autoregressive, Pixel Diffusion, Commercial) |
| 2025 | DDL | - | - | 367K | - | - |
| 2025 | DiffSeg30k | - | - | 30K | COCO | SD2, SD3.5, SDXL, Flux.1, Glide, Kolors, HunyuanDiT1.1, Kandinsky 2.2 |
| 2025 | FakeClue | - | - | 100K | - | - |
| 2025 | FakeParts | - | - | 81K | - | - |
| 2025 | ForensicHub | - | - | 23 datasets ; 42 models | - | ProGAN, StyleGAN, LDM, SDv1.4, SDv1.5, SDv2, SDXL, SD-ControlNet, MidJourney, ADM, GLIDE, VQDM, BigGAN |
| 2025 | Forensics-Bench | - | - | 63K | Various public datasets | GAN, Diffusion, VAE, RNN, Encoder-Decoder, Graphics-based |
| 2025 | GenBuster | - | - | 200K | - | - |
| 2025 | GenBuster++ | - | - | 4K | - | - |
| 2025 | Ivy-Fake | - | - | 150K | - | - |
| 2025 | LOKI | - | - | 18K | - | SORA, Keling, Open-Sora, FLUX, Midjourney, Stable Diffusion, Nerf-based, Gaussian-based, GPT-4o, Qwen-Max, Llama 3.1-405B, MusicGen, AudioLDM2... |
| 2025 | NeXT-IMDL | - | - | 558K | Flickr30k, COCO, OpenImages V7 | SD2-Inpainting, SDXL-Inpainting, FLUX-Inpainting, etc. |
| 2025 | OpenFake | - | - | ~4M | LAION-400M | SD 1.5/2.1/XL/3.5, Flux 1.0-dev/1.1-Pro/Schnell, Midjourney v6/v7, DALL·E 3, Imagen 3/4, GPT Image 1, Ideogram 3.0, Grok-2, HiDream-I1, Recraft v3, Chroma, and 10 community LoRA/finetune variants |
| 2025 | OpenSDI | - | - | 300K | Megalith-10M | SD1.5, SD2.1, SDXL, SD3, Flux.1 |
| 2025 | RewardData | - | - | 4.3K | - | - |
| 2025 | So-Fake-Set | - | - | 2M+ | F30k, WIDER, FFHQ, CelebA, OpenImages, COCO, OpenForensics | Qwen-image, GPT-4o, Nano Banana, Seedream3.0, Ideogram3.0, etc. |
| 2025 | Video Reality Test | - | - | 149 real + dynamic fake | YouTube ASMR (social media) | Veo3.1-Fast, Sora2, Wan2.2-A14B, Wan2.2-5B, OpenSora-V2, HunyuanVideo, StepVideo |
| 2025 | XAIGID-RewardBench | - | - | 3K | COCO-2017 | Imagen 4, Flux.1 Dev, Bagel, etc. |
| 2026 | ActivityForensics | - | - | 6K | - | - |
| 2026 | AIGI-Now | - | - | 18K | COCO | Nano Banana, GPT-4o, Jimeng, Kling, Minimax, etc. |
| 2026 | AIGVDBench | - | - | 440k | OpenVid-HD | 31 generation models |
| 2026 | BR-Gen | - | - | 150K | - | - |
| 2026 | GenVidBench | - | - | 6M | - | - |
| 2026 | HiResolution | - | - | 50K | - | - |
| 2026 | HydraFake | - | - | 100K | FFHQ, VFHQ, CelebAHQ, FF++, etc. | GPT-4o, HailuoAI, ICLight, InfiniteYou, etc. |
| 2026 | MintVid | - | - | 4K | OpenVid, VFHQ, HDTF, TikTok | Jimeng3.0-Pro, Seedance, Kling2.5-Turbo, Sora2, TikTok, Youtube, etc. |
| 2026 | RealChain | - | - | 14K | - | - |
| 2026 | SciFigDetect | - | - | 150K | - | Nano Banana Pro, GPT-image-1.5 |
| 2026 | Skyra | - | - | 4K | - | - |
| Tool | Language | Platform | Description |
|---|---|---|---|
| Fake Image Detector | - | Chrome/Firefox | Browser extension for fake image detection |
| Hive AI Detector | - | Chrome | Browser extension for AI content detection |
| Tool | Language | Platform | Description |
|---|---|---|---|
| DeepFake-Official | Python | - | Face swapping detection resources |
| FaceForensics | Python | - | Face manipulation detection dataset and tooling |
| FakerLab | Python | - | AIGC detection benchmark |
| Tool | Language | Platform | Description |
|---|---|---|---|
| AI or Not | - | Web | AI detection service |
| EyeSift | - | Web | Free online AI text/image/video/audio detector with detailed per-model benchmarks |
| Hive AI Detection | - | Web/API | Video and image detection |
| Hive Moderation | - | Web | Website |
| Illuminarty | - | Web | Website |
| Illuminet | - | Web | AI image verification |
| Is it AI? | - | Web | Website |
| Tencent Zhuque AI Detection Assistant | - | Web | Website |
| TruthScan | - | Web | Website |
| Winston AI | - | Web | Website |
| Tool | Language | Platform | Description |
|---|---|---|---|
| Content Credentials | - | Web | C2PA provenance and content credential verification |
| SiliconSignature | - | GitHub | Hardware-bound image authentication and provenance certification using ASIC PoW nonces |
Resources here are useful context for authenticity, misinformation, and provenance work, but they are intentionally kept out of the core AIGC detection paper list.
2026.05 · Multimodal fake news / misinformation detection The AI Slop Intelligence Dashboard Problem (James Sawyer Field Notes) Why adjacent: Related to authenticity and misinformation, but the primary task is critical analysis of AI-generated intelligence dashboard claims rather than direct AIGC content detection. Links: Resource
2025.04 · Multimodal fake news / misinformation detection Exploring Modality Disruption in Multimodal Fake News Detection (arxiv) Why adjacent: Related to authenticity and misinformation, but the primary task is not direct AIGC content detection. Links: Resource
2024.01 · Multimodal fake news / misinformation detection MiRAGeNews: Multimodal Realistic AI-Generated News Detection (ACL) Why adjacent: Related to authenticity and misinformation, but the primary task is not direct AIGC content detection. Links: Resource
Contributions are welcome. To add or update entries:
data/.python3 scripts/validate_data.py.python3 scripts/generate_readme.py.Please keep placeholder links such as https://arxiv.org/abs/ out of data/; leave incomplete candidates in PAPERS_TO_ADD.md until metadata is complete.
If you have questions or suggestions, please open an issue.
Last generated on 2026.05.19
Python
100.0%