Three Hackers Used Claude to Break Into OpenAI In Less Than 72 Hours
Three researchers breached OpenAI systems in under 72 hours using Anthropic's Claude AI model. The attack is described by some outlets as an ethical hack, while others highlight it as a significant escalation in AI-driven cybersecurity risks. The breach occurred after Anthropic released its Opus 5 model, which reportedly provided the capabilities necessary to succeed where previous versions of Claude failed. This incident demonstrates the ability of large language models to be weaponized for targeted cyberattacks against other AI developers.
Listen to Live Briefing
Real-time synthesized voice briefing · Live Feeds Desk
- ✓ Researchers used Anthropic's Claude models to breach OpenAI.
- ✓ The breach was carried out by three hackers in less than 72 hours.
What changed
Reports now specify that three hackers completed the breach in less than 72 hours using the Opus 5 model.
Live updates
-
Three hackers use Anthropic Claude to breach OpenAI
Three researchers breached OpenAI systems in under 72 hours using Anthropic's Claude AI model. The attack is described by some outlets as an ethical hack, while others highlight it as a significant escalation in AI-driven cybersecurity risks. The breach occurred after Anthropic released its Opus 5 model, which reportedly provided the capabilities necessary to succeed where previous versions of Claude failed. This incident demonstrates the ability of large language models to be weaponized for targeted cyberattacks against other AI developers.
Why it matters
The breach occurs amid heightened global concerns regarding AI safety and the security of model infrastructure. It highlights a competitive tension where one AI model is used to compromise a rival. This event raises questions about the guardrails implemented by AI labs to prevent their tools from assisting in malicious activities.
What is confirmed
- Researchers used Anthropic's Claude models to breach OpenAI.
- The breach was carried out by three hackers in less than 72 hours.
Still unconfirmed
- The breach was an ethical hack.
- The attack was only possible after the release of the Opus 5 model.
What to watch next
- Official response or post-mortem report from OpenAI regarding the breach
- Statement from Anthropic on the specific guardrails bypassed by the hackers
- Evidence of similar AI-assisted breaches targeting other LLM developers
confidence 80%Sources used for this update (9)
- WSJ — Exclusive | Hackers Used Anthropic’s Claude to Break Into OpenAI
- Financial Times — OpenAI breached by researchers using Anthropic models
- Ars Technica — Researchers used Claude to hack OpenAI
- The Guardian — OpenAI ‘ethically hacked’ with help of Anthropic’s Claude chatbot
- Semafor — AI cybersecurity risks explode as Claude used to break into ChatGPT
- NBC News — Hackers breached OpenAI, adding to fever pitch of security and safety concerns
- Gizmodo — Three Hackers Used Claude to Break Into OpenAI In Less Than 72 Hours
- The New Stack — Claude couldn’t hack OpenAI. Then Anthropic shipped Opus 5.
- www.tweaktown.com — Google Gemini broke out of its test environment and hacked three real companies
Community Sentiment: How do you assess this situation?
Voice your perspective · Real-time aggregated sentiment from the Live Feeds community