Modelpedia
Work in progress
About
Findings
Models
Concepts
Methods
Datasets
Sources
Light
Dark
Robustness of AI-Image Detectors: Fundamental Limits and Practical Attacks
2024-01-16
· ICLR 2024 poster ·
anchor
Findings
IC-1195
Model substitution adversarial attack reduces TreeRing AUROC to 0.14 at ε=2/255 and StegaStamp AUROC to 0.492 at ε=12/255
IC-1196
Blending a watermarked noise image with a clean image causes watermark detectors to falsely flag clean images as watermarked