Modelwire
Subscribe

Claude used to breach OpenAI systems in security research

Illustration accompanying: Researchers used Anthropic’s Claude to hack into OpenAI

Security researchers weaponized Anthropic's Claude to penetrate OpenAI's infrastructure, compromising employee credentials and accessing internal code repositories in a coordinated vulnerability disclosure. The incident underscores a critical tension in AI safety culture: frontier labs are now attack vectors for each other, and the tools they build can be turned against competitors. This raises uncomfortable questions about responsible disclosure norms when the attacker and target are both safety-focused organizations, and whether Claude's capabilities have crossed a threshold where offensive security applications become routine.

Modelwire context

Analyst take

The detail worth sitting with is not that Claude was capable enough to do this, but that a coordinated vulnerability disclosure happened at all between two organizations that are simultaneously competitors, occasional collaborators, and mutual critics on safety. The norms governing that kind of disclosure between frontier labs simply do not exist yet in any codified form.

Modelwire has no prior coverage to anchor this to directly, so it stands largely on its own. The story belongs to an emerging category we have not yet tracked systematically: offensive security applications of frontier models, and the liability questions that follow when a lab's product is the instrument of a breach. The closest adjacent territory is the broader responsible-disclosure debate in enterprise software, but that field assumed the attacker and the target had no shared safety mission or overlapping governance concerns. That assumption breaks down here in ways that will matter to anyone watching how labs negotiate trust with each other.

Watch whether Anthropic and OpenAI issue any joint statement or updated usage-policy language specifically addressing offensive security research within the next 60 days. If neither lab moves to formalize disclosure norms, that silence will tell you the competitive relationship is too fraught for the cooperation their public safety commitments imply.

This analysis is generated by Modelwire’s editorial layer from our archive and the summary above. It is not a substitute for the original reporting. How we write it.

MentionsAnthropic · Claude · OpenAI · TechCrunch

MW

Modelwire Editorial

This synthesis and analysis was prepared by the Modelwire editorial team. We use advanced language models to read, ground, and connect the day’s most significant AI developments, providing original strategic context that helps practitioners and leaders stay ahead of the frontier.

Modelwire summarizes, we don’t republish. TechCrunch - AI originally reported this story as Researchers used Anthropic’s Claude to hack into OpenAI”. The full content lives on techcrunch.com. If you’re a publisher and want a different summarization policy for your work, see our takedown page.

Claude used to breach OpenAI systems in security research · Modelwire