
USENIX Security ’26 Artifact Evaluated — Available
COGNITION: From Evaluation to Defense against Multimodal LLM CAPTCHA Solvers
COGNITION maps the practical boundary of automated visual CAPTCHA solving and turns that evidence into concrete redesign guidance. The study pairs broad evaluation with a defense case study that reduces state-of-the-art solver success from over 95% to 0%.
- 7
- multimodal models
- 21
- task types
- 458
- instances
