Our evaluation of OpenAI's GPT-5.5 cyber capabilities
Source published ·Modelwire updated
Original coverage: Simon Willison ↗·How Modelwire adds context

The development
The UK's AI Security Institute has completed a formal evaluation of GPT-5.5's ability to identify security vulnerabilities, finding it matches Claude Mythos in capability but with a critical advantage: immediate public availability. This benchmark matters because it signals that frontier models are now reaching parity on high-stakes cybersecurity tasks, raising both the bar for responsible deployment and the urgency around access controls for dual-use AI capabilities. The comparison to Mythos positions GPT-5.5 as the more accessible threat vector for security teams to monitor.
Modelwire’s AI-generated summary of coverage from Simon Willison.
Modelwire analysis
Analyst takeOur AI-generated reading of the wider context and the next developments to watch.
The evaluation's most consequential finding isn't parity with Claude Mythos on capability scores, it's that GPT-5.5's public availability means the dual-use risk calculus is already live, while Mythos remains gated. Capability equality with asymmetric access is a materially different threat posture than a simple benchmark tie.
This lands directly against the Anthropic valuation story from April 30, where investor appetite north of $900 billion is explicitly tied to confidence in Claude's competitive positioning against OpenAI. A formal government evaluation finding GPT-5.5 at parity with Mythos on cybersecurity tasks complicates that narrative: if the capability gap has closed on one of the highest-stakes dimensions, Anthropic's differentiation increasingly rests on access controls and safety process rather than raw model performance. That's a defensible moat, but a narrower one than investors may be pricing in.
Watch whether Anthropic responds by accelerating Mythos's public release timeline or by leaning harder into restricted access as a selling point for enterprise and government contracts. If Mythos remains gated six months after this evaluation publishes, that tells you Anthropic has made a deliberate strategic choice to cede the accessibility argument entirely.
This interpretation is generated from the summary above and the archive coverage cited below. Our methodology · Report an error
Coverage behind this analysis
These archive entries ground the connection in our analysis. They are ordered by source publication date, with links to our coverage and the original sources.
·TechCrunch - AI
Sources: Anthropic potential $900B+ valuation round could happen within two weeks
Anthropic is accelerating a major capital raise that could value the AI safety-focused lab north of $900 billion, with investor commitments due within 48 hours. The timeline suggests imminent close and signals continued investor appetite for frontier AI infrastructure despite market volatility. A valuation at this level would place Anthropic among the most valuable private…
MentionsOpenAI · GPT-5.5 · Claude Mythos · UK AI Security Institute · Simon Willison
How this coverage is produced
Modelwire uses AI to generate summaries and context from source headlines, snippets, and selected archive coverage. Automated checks do not verify every claim, and items are not routinely reviewed by a person before publication. Zacaria Solis operates the site. Read the linked source for the full evidence and report errors through our corrections process.
Modelwire summarizes, we don’t republish. The full content lives on simonwillison.net. If you’re a publisher and want a different summarization policy for your work, see our takedown page.