Technology 15 sources · today Latest coverage 18 Sept 2026, 0:49 pm UTC

Security Researchers Breach OpenAI Using Anthropic's Claude Model in 72 Hours

A security team used Anthropic's Claude AI to exploit vulnerabilities in OpenAI systems, exposing employee accounts and internal code, raising urgent concerns about AI-driven cybersecurity threats.

By Claire Dubois · First published 18 Sept 2026

In brief

  1. A group of security researchers from Hacktron used Anthropic's Claude AI to breach OpenAI's infrastructure.
  2. The researchers exploited a vulnerability in OpenAI's Discourse forum, gaining unauthorized access to ChatGPT accounts and internal GitHub code.
  3. The hack was completed in less than 72 hours and was part of OpenAI's official bug bounty program.
  4. OpenAI awarded the team a $6,500 bounty for responsibly disclosing the flaws.
  5. The incident highlights growing risks as AI tools become more capable of both launching and defending against cyberattacks.
Security Researchers Breach OpenAI Using Anthropic's Claude Model in 72 Hours
Source: The Wall Street Journal Tech

Timeline · 5 moments

5 moments Open the full timeline →

Hacktron begins probing OpenAI systems using Anthropic's Claude model

Публикации по подписке ↗

Researchers exploit Discourse forum flaw to access employee ChatGPT accounts

Digital Trends ↗

Hackers access internal GitHub repositories linked to OpenAI employee accounts

The Wall Street Journal Tech ↗

OpenAI confirms breach and awards $6,500 bug bounty to researchers

Forbes ↗

Incident prompts renewed debate over AI-driven cybersecurity risks

AI - The Guardian ↗

How it started

The breach began when a team of independent security researchers, operating under the name Hacktron, decided to participate in OpenAI's bug bounty program. Their goal was to find exploitable vulnerabilities in OpenAI's public-facing systems. Instead of relying solely on their expertise, the team leveraged Anthropic's Claude, a rival AI model, to assist in their search for weaknesses.

According to The Wall Street Journal Tech, the researchers targeted OpenAI's Discourse forum, which hosts discussions and support for the company's products. This forum became the entry point for their investigation.

How it unfolded

On July 25, 2026, the Hacktron team began probing OpenAI's systems with the help of Anthropic's Claude model, focusing on the Discourse forum software. They discovered two linked vulnerabilities, one of which allowed them to exploit the forum's authentication process.

Using code generated by Claude, the team managed to access an OpenAI employee's ChatGPT account. This initial breach opened the door to further access, including private GitHub code repositories connected to the compromised accounts, as reported by Digital Trends.

The entire process, from identifying the vulnerability to gaining access, took less than 72 hours, underlining the speed with which modern AI tools can accelerate both attack and defense in cybersecurity, according to TechRadar.

After confirming the extent of their access, the researchers responsibly disclosed the vulnerabilities to OpenAI as part of the company's bug bounty program. OpenAI reviewed their findings and verified the breach, which involved sensitive internal code and employee accounts.

Where it stands

OpenAI responded quickly to the disclosure, patching the vulnerabilities and awarding Hacktron a $6,500 bug bounty for their work. The company has not reported evidence of malicious exploitation beyond this ethical hacking demonstration.

The incident has sparked debate in the tech community about the growing capabilities of AI models in both offensive and defensive cybersecurity roles. It also prompted renewed scrutiny of the security of AI-driven platforms and the importance of robust vulnerability disclosure programs.

What to watch

Observers are watching to see how OpenAI and other AI companies will adapt their security practices in the wake of this breach. The case may influence how future AI bug bounty programs are designed and could accelerate efforts to secure AI-powered platforms against similar attacks.

Written from 15 outlets' coverage of this story. Every timeline entry links to the original report.

More in Technology

All →