Back to feed

Hacktron AI researchers breach OpenAI systems using Anthropic Claude

1 min
Hacktron AI researchers breach OpenAI systems using Anthropic Claude

This digest was compiled by AI from multiple sources — links to the originals are below.

Three Hacktron AI researchers breached OpenAI's internal systems by exploiting a vulnerability in the company's community forum using Anthropic's Claude model. They accessed an employee's ChatGPT account and viewed confidential code, earning a $6,500 bug bounty from OpenAI. OpenAI confirmed the vulnerabilities have been fixed.

Key Facts

  • Hacktron AI received a $6,500 bug bounty from OpenAI for the discovery.
  • The researchers exploited a vulnerability in OpenAI's community forum to access internal credentials and an employee's ChatGPT account.
  • Anthropic's Claude model was used to write exploit code for the vulnerability.
  • OpenAI stated that the vulnerabilities have been fixed.
  • Anthropic reported that Claude assists with 26% of its R&D projects, up from 1% in March.

The Breach

Three researchers from Hacktron AI exploited a vulnerability in OpenAI's community forum to obtain internal credentials. They then accessed an OpenAI employee's ChatGPT account, which had access to internal code via GitHub. The researchers used a specialized version of Anthropic's Claude Opus 4.8 to write the exploit code. OpenAI paid the researchers a $6,500 bug bounty and confirmed the vulnerabilities were fixed.

Anthropic's AI Development

Anthropic disclosed that Claude now assists with 26% of its research and development projects, up from 1% in March. The company shared this data to help the public understand how close the world is to recursive self-improvement in AI. Anthropic noted that in 90% of tasks, AI collaborates with humans rather than acting autonomously.