ATRIUMsearch → argument graph
ClaimArticle

Autonomous exploit development by frontier AI agents is no longer a hypothetical capability.

The ExploitGym paper concludes that frontier AI agents can now autonomously turn vulnerabilities into working exploits, a capability previously considered implausible. ✦ AI generated

ExploitGym authors (UC Berkeley, Max Planck Institute, UC Santa Barbara, Arizona State) · Simon Willison's Weblog · 2026-07-22 · original ↗

Our results show that autonomous exploit development by frontier AI agents is no longer a hypothetical capability. While current agents are not yet reliable across all targets, they already exploit a non-trivial fraction of real-world vulnerabilities, including complex targets such as kernel components.

Read full article ↗excerpt · fair-use quotation

Around this claim
This moment responds to