Anthropic red team flags binary exploitation breakthrough across frontier models

Anthropic's red team has documented a critical capability threshold: both GLM-5.3 and Claude Mythos Preview now succeed at binary exploitation tasks at measurable rates, whereas prior generations showed zero success. This marks the first time advanced models have crossed into reliable control-flow hijacking, signaling that cyber offense capabilities are diffusing across the frontier lab ecosystem. The finding underscores an emerging asymmetry in AI safety: defensive benchmarking is outpacing industry consensus on responsible disclosure and containment protocols.
Modelwire context
ExplainerThe detail worth sitting with is the zero-to-nonzero jump itself. In security research, that specific transition matters more than the absolute success rate, because it means the capability is no longer theoretically possible but practically achievable, and optimization pressure will push the number upward from here.
The related coverage on Modelwire this week has been dominated by competitive positioning and political branding, including the Trump administration's executive order mandating the term 'Super Intelligence' across federal communications. That story is largely disconnected from this one in substance, but the contrast is instructive: while federal language policy is moving toward maximalist framing of AI capability, Anthropic's own red team is quietly publishing findings that give that framing concrete and uncomfortable weight. The gap between political rhetoric and technical disclosure norms is widening, and this finding sits squarely in that gap.
Watch whether any of the named models, GLM-5.3 or Claude Mythos Preview, appear in a coordinated responsible disclosure timeline within the next 60 days. If no such timeline materializes, that confirms the asymmetry the summary flags: red team findings are outrunning the industry's ability to agree on what to do with them.
This analysis is generated by Modelwire’s editorial layer from our archive and the summary above. It is not a substitute for the original reporting. How we write it.
MentionsAnthropic · GLM-5.3 · Claude Mythos Preview · Claude Opus 4.6 · GLM-5.2
Modelwire Editorial
This synthesis and analysis was prepared by the Modelwire editorial team. We use advanced language models to read, ground, and connect the day’s most significant AI developments, providing original strategic context that helps practitioners and leaders stay ahead of the frontier.
Modelwire summarizes, we don’t republish. Simon Willison originally reported this story as “Quoting Anthropic Frontier Red Team”. The full content lives on simonwillison.net. If you’re a publisher and want a different summarization policy for your work, see our takedown page.