InfoResearchPeer-reviewed
Steering the Narrative: Breaking Bots by Forging Robust Adversarial CAPTCHAs With Diffusion Models
- Published
- Record updated
Summary
Researchers propose SteerGuard, a framework that uses diffusion models to generate adversarial CAPTCHAs designed to defeat AI-based solvers. Their approach steers the diffusion generation process to embed adversarial semantics directly into CAPTCHA samples, with the aim of reducing manual dataset curation and enabling dynamic updates. The authors report superior transferability, robustness and adaptability across attacker strategies without added runtime cost, including validation on a large-scale online platform.
Related items
- InfoDoes Target Alignment Mean Target Recovery? An Evidence-Ladder Study of Adversarial Claims on Contrastive EncodersSimilar attack · Arxiv (cs.RO + cs.CV security)
- InfoBRANCH: Bypassing Multi-Scanner AI GuardrailsSimilar attack · Arxiv (cs.CR + cs.CL + cs.LG)
- InfoDetecting Adversarial Images through Response Profiles of Vision-Language ModelsSimilar attack · Arxiv (cs.RO + cs.CV security)
- InfoGraphRectify: Graph-Based Transfer of Adversarial Example Detectors Across Neural NetworksSimilar attack · Arxiv (cs.RO + cs.CV security)
- InfoVCR-Bench: A Modular Open-Source Benchmark for Video Classification RobustnessSimilar attack · Arxiv (cs.RO + cs.CV security)