Share your thoughts, 1 month free Claude Pro on usSee more
WorkDL logo mark

Dual Data Alignment Makes AI-Generated Image Detector Easier Generalizable

About

Existing detectors are often trained on biased datasets, leading to the possibility of overfitting on non-causal image attributes that are spuriously correlated with real/synthetic labels. While these biased features enhance performance on the training data, they result in substantial performance degradation when applied to unbiased datasets. One common solution is to perform dataset alignment through generative reconstruction, matching the semantic content between real and synthetic images. However, we revisit this approach and show that pixel-level alignment alone is insufficient. The reconstructed images still suffer from frequency-level misalignment, which can perpetuate spurious correlations. To illustrate, we observe that reconstruction models tend to restore the high-frequency details lost in real images (possibly due to JPEG compression), inadvertently creating a frequency-level misalignment, where synthetic images appear to have richer high-frequency content than real ones. This misalignment leads to models associating high-frequency features with synthetic labels, further reinforcing biased cues. To resolve this, we propose Dual Data Alignment (DDA), which aligns both the pixel and frequency domains. Moreover, we introduce two new test sets: DDA-COCO, containing DDA-aligned synthetic images for testing detector performance on the most aligned dataset, and EvalGEN, featuring the latest generative models for assessing detectors under new generative architectures such as visual auto-regressive generators. Finally, our extensive evaluations demonstrate that a detector trained exclusively on DDA-aligned MSCOCO could improve across 8 diverse benchmarks by a non-trivial margin, showing a +7.2% on in-the-wild benchmarks, highlighting the improved generalizability of unbiased detectors. Our code is available at: https://github.com/roy-ch/Dual-Data-Alignment.

Ruoxin Chen, Junwei Xi, Zhiyuan Yan, Ke-Yue Zhang, Shuang Wu, Jingyi Xie, Xu Chen, Lei Xu, Isabel Guan, Taiping Yao, Shouhong Ding• 2025

Related benchmarks

TaskDatasetResultRank
AI-generated image detectionGenImage
Midjourney Detection Rate96
154
AI-generated image detectionChameleon
Accuracy84.82
127
Synthetic Image DetectionForenSynths (test)
Mean Accuracy81.4
60
Synthetic Image DetectionDRCT-2M
Average Score98
57
Fake Image DetectionUniversalFakeDetect (test)
Pro-GAN Detection Rate84.7
52
AIGC DetectionAIGCDetectBenchmark--
50
AIGI DetectionBFree Online
B.Acc81.2
47
AIGI DetectionAIGCDetect
B.Acc94.9
46
AIGI DetectionSynthWildx
DALLE3 Performance Score92.3
46
AI-generated image detectionProGAN
mAP97.7
39
Showing 10 of 138 rows
...

Other info

Follow for update