Skip to content
ACarstechnica.com·

Researchers used Claude to hack OpenAI

AI summary

Researchers exploited a flaw in OpenAI’s community forum, hosted by Discourse, to gain access to internal sign-ons and an OpenAI employee’s ChatGPT account. This account had access to internal code via GitHub. The report also noted that 26 percent of research and development work was “led by” its Claude model, an increase from 1 percent in March, indicating that AI completed most tasks under human supervision.

Why this one

This report uniquely details how researchers used a third-party forum vulnerability to access OpenAI's internal code, unlike other reports focusing on AI model capabilities.

Time & source

Times shown in UTC

Display time zone: UTC

Local time zone unavailable; showing UTC.

PublishedOffset at this time: UTC+0Sep 18, 2026, 13:30 UTC

IngestedOffset at this time: UTC+0Sep 18, 2026, 14:00 UTC

Published
Sep 18, 2026, 13:30
Ingested
Sep 18, 2026, 14:00
Source type
Media
Tier
Press
Source status
Healthy

Tier is a per-source editorial setting, not a per-item score.

Discussion trend

→ Steady
Latest 24h versus previous 24h snapshot means · 7-day curve

The percentage is based on collected discussion signal, not new comments or independent people. The curve only compares the same topic across time.

They exploited a flaw in the set-up of OpenAI’s community forum, which is hosted by a third-party, Discourse, and used it to gain access to internal sign-ons and eventually an OpenAI employee’s ChatGPT account. This ChatGPT account had access to internal code through GitHub.

“We thank the researchers for contacting us and sharing their findings,” OpenAI said, adding that it had fixed the issues. Anthropic declined to comment. Hacktron did not immediately respond.

The disclosure on Thursday, first reported by The Wall Street Journal, came as Anthropic published a new set of data that showed a rapid increase in how much the lab used AI to develop its new models.

It said 26 percent of research and development work was “led by” its Claude model, up from 1 percent in March, meaning that AI completed the majority of tasks based on human instruction and under supervision.

The company said that as AI systems become more powerful, they were “increasingly being used to build the next version of themselves.”

Anthropic said it shared the data to help the public “understand how close the world is to reaching recursive self-improvement,” the point at which AI can train and improve itself or new models.

This threshold is at the heart of concerns that AI systems will become more difficult to oversee, leading to a loss of human control.

Its models did not yet operate fully autonomously for any of the research it studied, Anthropic added. On 90 percent of tasks, AI “collaborates” with a human and does large chunks of work.

© 2026 The Financial Times Ltd. All rights reserved. Not to be redistributed, copied, or modified in any way.

Related story3 reports · 3 publishers
View full story