OpenAI ‘ethically hacked’ with help of Anthropic’s Claude chatbot
AI security firm Hacktron used Anthropic's Claude Opus 5 to breach OpenAI's internal systems in under 72 hours. The researchers exploited an image-processing flaw to reach employee accounts and the company's internal code repository. OpenAI confirmed the breach, patched the vulnerability within 14 hours, and paid a $6,500 reward to the three individuals involved. The incident demonstrates how rival AI models can be used to identify and exploit vulnerabilities in competing AI infrastructure.
Listen to Live Briefing
Real-time synthesized voice briefing · Live Feeds Desk
- ✓ Hacktron used Anthropic's Claude to breach OpenAI.
- ✓ The breach took less than 72 hours to execute.
- ✓ Researchers accessed OpenAI's internal code repository and employee accounts.
- ✓ OpenAI paid a $6,500 reward for the discovery.
What changed
Hacktron identified as the research firm that used Claude Opus 5 to access OpenAI's source code.
Live updates
-
Hacktron researchers use Claude to breach OpenAI source code
AI security firm Hacktron used Anthropic's Claude Opus 5 to breach OpenAI's internal systems in under 72 hours. The researchers exploited an image-processing flaw to reach employee accounts and the company's internal code repository. OpenAI confirmed the breach, patched the vulnerability within 14 hours, and paid a $6,500 reward to the three individuals involved. The incident demonstrates how rival AI models can be used to identify and exploit vulnerabilities in competing AI infrastructure.
Why it matters
This event highlights the escalating cybersecurity risks as large language models become capable of automating complex hacking tasks. It marks a rare instance where one leading AI model was directly instrumental in compromising another.
What is confirmed
- Hacktron used Anthropic's Claude to breach OpenAI.
- The breach took less than 72 hours to execute.
- Researchers accessed OpenAI's internal code repository and employee accounts.
- OpenAI paid a $6,500 reward for the discovery.
- OpenAI patched the systems within 14 hours of confirmation.
What to watch next
- Anthropic's response to the use of Claude for offensive cyber operations
- OpenAI's updated security protocols for image-processing systems
confidence 90%Sources used for this update (8)
- WSJ — Exclusive | Hackers Used Anthropic’s Claude to Break Into OpenAI
- Ars Technica — Researchers used Claude to hack OpenAI
- The Guardian — OpenAI ‘ethically hacked’ with help of Anthropic’s Claude chatbot
- Semafor — AI cybersecurity risks explode as Claude used to break into ChatGPT
- Fortune — Three guys using Anthropic’s Claude hacked into OpenAI and accessed its source code for $6,500 reward
- NBC News — Hackers breached OpenAI, adding to fever pitch of security and safety concerns
- cryptoslate.com — Anthropic’s Claude helped 3 researchers breach OpenAI in under 72 hours
- news.jkn.co.kr — Rival AI Firm Hacked in 72 Hours by AI Model; Cybersecurity Concerns Mount
Community Sentiment: How do you assess this situation?
Voice your perspective · Real-time aggregated sentiment from the Live Feeds community