Three researchers have reportedly used an Anthropic Claude model to successfully hack OpenAI, according to a CBS News investigation that is drawing attention across the AI security community. The incident is being read as evidence that frontier AI labs face a specific and underappreciated vulnerability: their own competitors' tools can be turned against them.

What Happened and Why It Matters

The researchers, described as a small independent group, leveraged Claude to probe and ultimately penetrate OpenAI's systems. Details about exactly which systems were accessed and what data, if any, was exposed remain limited. What is clear is that the attack used one large language model as an instrument against infrastructure built around another. This is not the first time AI models from rival companies have been used to stress-test each other's defenses, but the scale and apparent success of this effort has prompted new scrutiny of lab-wide security postures.

Key Facts

  • Three independent researchers carried out the attack using a Claude model.
  • The target was OpenAI's internal systems, not a public-facing product.
  • The incident reveals gaps in how frontier labs anticipate AI-assisted threats.
  • Neither OpenAI nor Anthropic has publicly confirmed full details of the breach.

Security experts argue the episode exposes a structural issue. AI companies spend considerable resources defending against conventional cyberattacks, but the idea that a rival model could be used as an autonomous or semi-autonomous attack tool is a newer threat vector. Anthropic has consistently positioned safety and security as central to its mission, and the company's models include guardrails intended to prevent misuse. How the researchers circumvented those safeguards, or whether they found a gap that did not require circumvention, is a key open question.

"Frontier labs are building the most powerful software tools in history, and they're doing it while also being targets. That's a difficult position to be in."AI security researcher, CBS News
Claude AI Handboek by Leon Tindemans
Get the Claude AI Handboek
458 pages on getting more out of Claude, by AI expert Leon Tindemans. A printed book, written in Dutch, shipped worldwide with track and trace.
View the book →

A Broader Pattern Across the Industry

The hack fits into a pattern that observers have been tracking for months. As AI models grow more capable, their potential utility as attack instruments grows alongside their utility as productivity tools. Anthropic has published metrics aimed at tracking how quickly AI capabilities are advancing, partly to give policymakers and the public a clearer picture of where the technology stands. That kind of transparency is useful, but it also underscores how rapidly the risk landscape is shifting.

OpenAI has not issued a detailed public statement about the breach. Anthropic, for its part, faces its own complicated position: its model was the instrument used, even if the company was not a target. Questions about whether Claude's capabilities were exploited in ways that fall outside its intended use case will likely prompt internal review, regardless of what either company says publicly.

The incident also arrives at a moment when AI regulation is actively being debated at the highest levels. Anthropic and other frontier lab leaders have been meeting with G7 governments to discuss AI governance, and episodes like this one tend to accelerate those conversations. Governments watching a frontier lab get hacked with a rival frontier model are likely to see this as concrete evidence that voluntary safety measures may not be enough.

For now, the incident serves as a pointed reminder that AI security is not a solved problem. The same properties that make large language models useful for researchers, developers, and businesses also make them potentially useful for adversarial purposes. Frontier labs will need to grapple with that reality in a more structured way, and this breach may be the catalyst that forces the issue onto boardroom agendas across the industry.

Further reading: Learn more about Claude's model family, read our background on Anthropic, or browse the latest Claude AI news.