文献汇总|AI生成图像检测相关数据集汇总

发布时间:2026/10/8 17:16:14

文献汇总|AI生成图像检测相关数据集汇总 前言本博客汇总当前AI生成图像检测领域用到的数据集及相关链接。⚠️ 更新说明由于博客与 Synthetic Image Research Map 的内容存在较多重叠分别维护两套整理结果会带来较高的更新成本因此本博客后续将不再单独更新相关论文与数据集列表。更完整、持续更新的生成图像研究整理请访问https://meiling-fdu.github.io/synthetic-image-research-map/web/?datasetpreview. 如需查看与数据集和基准相关的论文请在页面的筛选栏中选择 Paper Type → Datasets / Benchmarks。Research Map 中的信息将作为后续主要维护版本。 相关文章链接工具分享Synthetic Image Research MapAI 生成图像检测与溯源研究的交互式文献地图目录20202022202320242020CNNSpothttps://github.com/peterwang512/CNNDetectionTestset: The zip file contains images from 13 CNN-based synthesis algorithms, including the 12 testsets from the paper and images downloaded from whichfaceisreal.com. Images from each algorithm are stored in a separate folder. In each category, real images are in the 0_real folder, and synthetic images are in the 1_fake folder.Note: ProGAN, StyleGAN, StyleGAN2, CycleGAN testset contains multiple classes, which are stored in separate subdirectories.Training set: The training set used in the paper can be downloaded here (Try alternative links 1,2 if the previous link does not work). All images are from LSUN or generated by ProGAN, and they are separated in 20 object categories. Similarly, in each category, real images are in the 0_real folder, and synthetic images are in the 1_fake folder.Validation set: The validation set consists of held-out ProGAN real and fake images, and can be downloaded here. The directory structure is identical to that of the training set.2022IEEE VIP Cup2022 IEEE Video and Image Processing Cup | Synthetic Image Detection Challenge)https://grip-unina.github.io/vipcup2022/SAChttps://github.com/JD-P/simulacra-aesthetic-captions数据集中图像命名包含生成所需的提示词如0_An_artwork_of_a_broken_wine_bottle_in_the_medium_of_dry_pigments_1.png43044_…png此外该数据集也被用于美学质量评价。2023DiffusionForensicshttps://github.com/ZhendongWang6/DIREDMimageDetectionhttps://github.com/grip-unina/DMimageDetection/tree/main/training_codehttps://luminohope.org/pub/publication/arxiv_diffusion_detection_2022/GenImagehttps://github.com/GenImage-Dataset/GenImageWe employ eight generative models for image generation, namely BigGAN [2], GLIDE [21], VQDM [8], Stable Diffusion V1.4 [25], Stable Diffusion V1.5 [25], ADM [5], Midjourney [20], and Wukong [35].Fake2Mhttps://arxiv.org/pdf/2304.13023We constructed 3 training fake datasets with about 2M images, named Fake2M, and 11 validation fake datasets with about 257K images using different latest modern generative models, which contain the SOTA Diffusion models (Stable Diffusion [46], IF [4]), the SOTA GAN model (StyleGAN3 [31]), the SOTA autoregressive model (CogView2 [19]), and the SOTA generative model (Midjounrey [6]), as shown in Tab. 2. We describe the details of our datasets in the following subsections.TWIGMAhttps://yiqunchen.github.io/TWIGMA/index.html#datasetArtiFacthttps://github.com/awsaf49/artifactTo include a diverse collection of real images from multiple categories, including Human/Human Faces, Animal/Animal Faces, Places, Vehicles, Art, and many other real-life objects, the proposed dataset utilizes 8 sources [7], [14]–[16] that are carefully chosen. Additionally, to inject diversity in terms of generators, the proposed dataset synthesizes images from 25 distinct methods [7]–[9], [14]–[24]. Specifically, it includes 13 GANs, 7 Diffusion, and 5 other miscellaneous generators. On the other hand, in terms of syntheticity, there are 20 fully manipulating and 5 partially manipulating generators, thus providing a broad spectrum of diversity in terms of generators used. The distribution of real and fake data with different sources is shown in Fig.1 and Fig.2, respectively. The dataset contains a total of 2,496,738 images, comprising 964,989 real images and 1,531,749 fake images. The most frequently occurring categories in the dataset are Human/Human Faces, Animal/Animal Faces, Vehicles, Places, and Art.13GANs: BigGAN, CycleGAN, Denoising Diffusion GAN, Diffusion GAN, FaceSynthetics, GANformer, GauGAN, ProGAN, ProjectedGAN, StarGAN, StyleGAN1, StyleGAN2, StyleGAN37DMs: DDPM, Glide, LaMa, Latent Diffusion, Stable Diffusion, Taming Transformer, VQDiffusion5 Others: CIPS, Generative Inpainting, MAT, Palette, SFHQSynthbusterhttps://zenodo.org/records/10066460UniversarialFakeDetecthttps://github.com/WisconsinAIVision/UniversalFakeDetect11GANs 7 DMs 1 其他DiffusionDBhttps://github.com/poloclub/diffusiondbWe construct DIFFUSIONDB (Fig. 2) by scraping user-generated images from the official Stable Diffusion Discord server. We choose Stable Diffusion as it is currently the only open-source large text-to-image generative model, and all generated images have a CC0 1.0 license that allows uses for any purposeCiFAKEhttps://github.com/jordan-bird/CIFAKE-Real-and-AI-Generated-Synthetic-ImagesCIFAKE is a dataset that contains 60,000 synthetically-generated images and 60,000 real images (collected from CIFAR-10). For the FAKE images, we generated the equivalent of CIFAR-10 with Stable Diffusion version 1.4LASTEDhttps://github.com/HighwayWu/LASTED训练集生成模型ProGANLexicaStable Diffusion测试集DreamBooth, Midjourney, NightCafe, StalbeAI, YiJian蚁鉴DDDB 未公开https://arxiv.org/abs/2302.14475DeepArt 未公开https://export.arxiv.org/pdf/2312.10407DEFAKE 未公开https://github.com/zeyangsha/De-Fake20k real image for training 10k real images for testing2024COCOFakehttps://github.com/aimagelab/COCOFakeCOCOFake, containing about 1.2 million images generated from the original COCO image–caption pairs using two recent text-to-image diffusion models, namely Stable Diffusion v1.4 and v2.0.FOSIDhttps://github.com/mever-team/fosidhttps://zenodo.org/records/13648239D^3https://aimagelab.ing.unimore.it/imagelab/page.asp?IdPage57The Diffusion-generated Deepfake Detection (D3) Dataset is a comprehensive collection designed for large-scale deepfake detection. It includes 9.2 million generated images, created using four state-of-the-art diffusion model generators. Each image is generated based on realistic textual descriptions from the LAION-400M dataset.We generate a comprehensive dataset that focuses on images generated by diffusion models and encompasses a collection of 9.2 million images produced by using four different generators.Generators: Stable Diffusion 1.4, Stable Diffusion 2.1, Stable Diffusion XL, and DeepFloyd IFConsequently, we generate and release the Diffusion-generated Deepfake Detection (D3 ) dataset containing 2.3 million records, each composed of a real image coming from LAION-400M [44] dataset and images from four generators, for a total of 9.2 million generated images. To verify the generation capabilities of deepfake detection methods to unseen generators, we also collect a challenging test set composed of 4.8k real images, each paired with 12 fake images generated by as many diffusion-based generators.With the aim of increasing the variance of the dataset, images have been generated with different aspect ratios, i.e. 256x256, 512x512, 640×480, and 640×360. Moreover, to mimic the distribution of real images, we also employ a variety of encoding and compression methods (BMP, GIF, JPEG, TIFF, PNG). In particular, we closely follow the distribution of encoding methods of LAION itself, therefore favoring the presence of JPEG-encoded images.ImagiNethttps://github.com/delyan-boychev/imaginethttps://huggingface.co/datasets/delyanboychev/imaginetTo support the development of defensive methods, we introduce ImagiNet, a high-resolution and balanced dataset for synthetic image detection, designed to mitigate potential biases in existing resources. It contains 200k examples, spanning four content categories: photos, paintings, faces, and uncategorized. Synthetic images are produced with open-source and proprietary generators, whereas real counterparts of the same content type are collected from public datasets.AntifakePrompthttps://github.com/nctu-eva-lab/AntifakePromptWe conduct full-spectrum experiments on datasets from a diversity of 3 held-in and 20 held-out generative models, covering modern text-to-image generation, image editing and adversarial image attacks.Real datasets. We use Microsoft COCO (COCO) (Lin et al. 2014) dataset and Flickr30k (Young et al. 2014) dataset. In our work, we selected 90K images, with shorter sides greater than 224, from COCO dataset for the real images in the training dataset. Moreover, to assess the generalizability of our method over various real images, we additionally select 3K images from Flickr30k dataset to form a held-out testing dataset, adhering to the same criterion of image size. 93kFake image for training 150kfor testing3k*21 63kFakeBenchhttps://arxiv.org/abs/2404.13306Regarding the genuine images, we sample 3,000 images from ImageNet [76] and DIV2K dataset [77].COCOXGENhttps://github.com/heikeadel/cocoxgenCOCOXGENCOCO Extended With Generated Images, which consists of real photos from the COCO dataset as well as images generated with SDXL and Fooocus using prompts of two standardized lengths.WildRFhttps://github.com/barcavia/RealTime-DeepfakeDetection-in-the-RealWorldWe propose to improve deepfake evaluation and align it with real-world settings by introducing WildRF, a realistic benchmark consisting of images sourced from popular social platforms. Specifically, we manually collected real images and fake images using keywords and hashtags associated with the suitable content. Our protocol is to train on one platform (e.g., Reddit) and test the detector on real and fake images from other unseen platforms (e.g., Twitter and Facebook).WildFakePaper: https://arxiv.org/pdf/2402.11843Dataset: https://modelscope.cn/datasets/hy2628982280/WildFakeLSUNDBhttps://github.com/jonasricker/diffusion-model-deepfake-detectionThe main dataset used in this work is hosted on Zenodo. In total, the dataset contains 50k samples (256x256) for each of the following generators trained on LSUN Bedroom, divided into train, validation, and test set (39k/1k/10k).DIFhttps://sergo2020.github.io/DIF/
延伸阅读

更多相关文章

2026/10/8 2:01:59

精读+全文阅读:华为基于市场需求的IPD集成产品规划和策划

该文档围绕华为基于市场需求的 IPD 集成产品规划和策划展开,从产品开发与技术开发的区别切入,介绍产品研发成功的评价方式、竞争环境下的研发策略,详细阐述 IPD 的定义、核心思想及企业实施 IPD 的价值。其适合企业中与产品研发、市场、管理相关的各类人员。 (本解读资料未…

2026/10/5 8:33:58

IPD流程执行的标准规范化——IPD流程执行检查表

IPD(Integrated Product Development,集成产品开发)流程执行检查表在IPD产品研发中具有重要性,它有助于确保研发过程的规范化、高效化和产品质量的可控性。以下是对其必要性和大概内容的描述: - 必要性 - 保证流程合规性:IPD流程是一套复杂的、经过优化的产品开发流程,…

2026/10/7 6:54:50

深入解读:159页华为IPD流程管理培训

(本解读资料未在绑定资源内) 该文档围绕华为 IPD 流程管理展开,适用于企业中参与产品开发、市场管理、项目管理等相关工作的人员,以及对 IPD 流程感兴趣、希望提升企业产品管理能力的人士。 主要内容涵盖 IPD 流程的多个关键部分:首先是需求管理(OR)流程,旨在统…

2026/10/8 17:12:06

Superpowers增强方案:从设计思路到实操避坑的完整指南

1. 从“superpowers”这个标题说起:它到底是什么第一次看到“superpowers”这个词,很多人脑子里蹦出来的可能是漫威电影里的超能力,或者是某些游戏里的技能系统。但如果你是在技术社区、开源项目或者工具链的语境下看到它,那它大概…

2026/10/8 17:12:06

Agent Skills 实战指南:从安装、开发到调试的完整路径

1. 从"skills"这个热词说起:它到底指什么 最近一段时间,"skills"这个词在技术社区里出现的频率明显高了起来。如果你只是偶尔刷到,可能会以为它说的是"技能"这个泛泛的概念,但实际在当下的语境里&a…

2026/10/8 17:12:06

从零搭建7x24小时无人值守的多Agent集群:架构设计与实践记录

去年冬天我把一台吃灰的 Mini 主机塞进了书架角落,给它配了一堆服务,然后它就真的开始“上班了”。到现在,这套系统已经连续跑了快两个月,期间我只远程重启过两次,还都是因为我自己乱改配置。这台机器上跑的&#xff0…

2026/10/8 17:12:06

AI Agent Skills实战:从npx本地验证到GKE云端部署

1. 从“skills”这个热词说起:它到底是什么,为什么突然火了最近几个月,不管是在技术社区、开发者群聊,还是在做AI应用的朋友圈子里,“skills”这个词出现的频率高得离谱。有人把它当成一个工具包,有人把它当…

2026/10/8 17:12:06

Superpowers:开源自托管的实时协作3D游戏开发环境安装与使用指南

第一次看到“superpowers”这个关键词的人,十有八九会把它当成一本成功学书籍,或者某个能强化浏览器功能的插件。但如果你搜索框里打的是“想要安装 superpowers”,那我猜你要找的,多半是那个开源、可自托管、基于浏览器的实时协作…

2026/10/8 17:07:05

context-mode实战:AI编程中上下文管理的工程化方法

最近这两周,"context-mode"这个词在技术社区里出现的频率明显高了起来。群里有人问它是不是某个编辑器新加的开关,有人说它是一套提示词模板,还有人干脆觉得这是又一轮概念炒作。我自己的态度比较明确:context-mode 背后…

2026/10/8 10:03:18

Jev+Agent接管浏览器:browser-use实战与jev-ultrafast性能优化

1. 从“Jev”说起:为什么我要把Agent接进浏览器“Jev”这个词最近在圈子里出现的频率越来越高,很多人第一次听到会以为是某个新模型的名字,其实它更像是一种思路——把Jev模型的能力当作底座,通过Agent的方式去接管浏览器&#xf…

2026/10/8 10:03:20

多智能体集群实战:DeepAgents编排、MCP与A2A协议及Skills体系

1. 从"单兵作战"到"集群协同":多智能体编排到底在解决什么问题如果你最近在折腾 Agent 相关的东西,大概率会有一种感觉:单个 Agent 能做的事情,其实很快就摸到天花板了。你给它一个提示词,挂几个工…

2026/10/8 6:05:44

无源低通滤波器设计实战:从RC到LC,手把手教你避开那些坑

/* MD / 富文本中的 .toc(含博客园搬家等嵌套结构);.toc-box 在侧栏,不受影响 */#content_views .toc,/* 编辑器常在目录前后插入空 p(:empty 仍占 20px),一并去掉避免顶空隙 */#content_views.markdown_views > p:empty:has(+ .toc),#content_views.markdown_views …

2026/10/8 0:02:17

自然数立方等于连续奇数之和:从证明到编程验证

十几年来我一直游走在数学科普和编程教学这两块内容之间,对“看起来像魔法、拆开全是数学”的结论总是格外敏感。最近翻资料时又撞见一句话:任何一个自然数 m 的立方,都可以写成 m 个连续奇数之和。2 的立方等于 3 加 5,3 的立方等…

2026/10/8 0:02:17

C#上位机SSH连接实战:用SSH.NET补齐超时、批量与密钥认证

简介:这是一份基于 C# 开发的 SSH 连接功能半成品工程,原本作为另一个主项目的子功能模块,现独立打包分享。工程采用 WinForms 界面,包含源码、解决方案、安装部署工程、NuGet 依赖包及说明文档,适合正在做远程连接、网…

2026/10/8 0:02:17

Java SpringBoot一体化智能售后系统设计与实现全解析

毕业设计年年做,Java Web 方向的题目翻来覆去就那么几个,但“一体化智能售后系统”这个题,每次看到我都觉得值得认真聊一聊。它不是一个简单 curd 堆出来的管理系统,而是把客户、工单、派单、处理、回访、统计整条链路串起来的一套…

还想了解更多?直接咨询顾问

免费诊断 + 免费方案 + 透明报价。

全国咨询热线400-8866-253
免费获取方案
☎咨询二维码 ☎ ↑