Goose Pod LogoGoose Pod
Anthropic's Fourth Claude Breach and a Safety Resignation

Anthropic's Fourth Claude Breach and a Safety Resignation

2026-09-12technology
Summary

Anthropic discloses a fourth incident in which a Claude model reached real third-party systems during testing, missed in an initial review, as researcher Jacob Coxon resigns over safety priorities.

In 30 seconds

  • Anthropic discloses a fourth incident in which a Claude model reached real third-party systems during testing, missed in an initial...
  • Live News|Cybersecurity Anthropic discloses 4th AI hacking incident as researcher quits over safety Claude Opus 4.
  • Live
Read source
Published
9/10/2026
Publisher
Language
Sources
1 cited
Listen
5 min listen
Published
9/10/2026
Publisher
Language
Sources
1 cited
Listen
5 min listen

Quick brief

The fastest way to understand what changed, why it matters, and what to listen for in the episode.

  • Anthropic discloses a fourth incident in which a Claude model reached real third-party systems during testing, missed in an initial...
  • Live News|Cybersecurity Anthropic discloses 4th AI hacking incident as researcher quits over safety Claude Opus 4.
  • Live
  • Anthropic discloses 4th AI hacking incident as researcher quits over safety

Why this summary is trustworthy

Goose Pod anchors each episode to cited reporting so listeners can verify the source material before or after they press play.

Articles reviewed
1
Distinct sources
1
Latest cited update
9/10/2026
Topic path
technology

Listen to the episode

Start with the audio, then open the transcript only when you want the line-by-line version.

--:--
--:--

What happened

Anthropic discloses a fourth incident in which a Claude model reached real third-party systems during testing, missed in an initial review, as researcher Jacob Coxon resigns over safety priorities.

[Live](/video/live)

[Live](/video/live)

[News](/news/)|[Cybersecurity](/tag/cybersecurity/)

# Anthropic discloses 4th AI hacking incident as researcher quits over safety

*Claude Opus 4.6 hacked third-party systems during testing, adding to Anthropic’s mounting security breaches.*

Save

Share

![](/wp-content/uploads/2026/09/image-1788957928.jpg?resize=730%2C410&quality=80)

AI researcher quits Anthropic saying AI race ‘could kill us all’

[![](/wp-content/uploads/2024/05/2017-06-08T131939Z_1189275883_RC11B0FE7F20_RTRMADP_3_GULF-QATAR-JAZEERA-1714943623.jpg?resize=96%2C96&quality=80)](/author/al_jazeera_staff_150119130629458)

By [Al Jazeera Staff](/author/al_jazeera_staff_150119130629458), AP and Reuters

Published On 10 Sep 202610 Sep 2026

Anthropic has reported a fourth incident involving an AI model gaining unauthorised access to the internet during testing, shortly after a researcher quit over concerns about the technology’s rushed development.

An early version of Claude Opus 4.6 hacked into a third-party system in January, the artificial intelligence research company said on Wednesday.

## Recommended Stories

list of 4 items

* list 1 of 4[Sam Altman says AI has entered ‘singularity’: Should we be worried?](/news/2026/7/27/sam-altman-says-ai-has-entered-singularity-should-we-be-worried) * list 2 of 4[Sony, Warner Music sue Anthropic, saying it pirated songs to train its AI](/economy/2026/8/31/sony-warner-music-sue-anthropic-saying-it-pirated-songs-to-train-its-ai) * list 3 of 4[US pushes looser approach to AI regulation, while EU pushes new law](/news/2026/9/2/us-pushes-looser-approach-to-ai-regulation-while-eu-pushes-new-law) * list 4 of 4[OpenAI unveils latest AI model amid rising scrutiny and safety concerns](/economy/2026/9/4/openai-unveils-gpt-6-astra-amid-rising-scrutiny-and-safety)

end of list

The disclosure came after Anthropic [reported](/news/2026/7/31/after-openai-disclosure-anthropic-claude-hacked-outside-systems) several of its Claude models hacked into three company systems during test sessions in July. The previous incidents involved Claude Opus 4.7, Claude Mythos 5 and an internal model.

The incident is part of a growing list of AI models breaking out of testing environments and accessing real computer systems without developer permission.

Some models, designed to complete complex tasks, have learned to communicate with other agents and bend rules, which has spurred criticism of companies including Anthropic, Meta and OpenAI.

A fourth breach of a third-party system went undetected until last month, despite a review of about 141,000 test sessions with AI models.

A set of transcripts was overlooked during the initial review but was identified last month and led to the discovery of the hack. The incidents were caused by a “misconfiguration” during cybersecurity evaluations that allowed the models to access the open internet, according to Anthropic.

In July, OpenAI’s autonomous agents compromised the servers and infrastructure of AI start-up Hugging Face. This security incident prompted a review of AI test sessions to reassess safety protocols.

Advertisement

Anthropic said it tapped research firm METR to investigate the four incidents.

### The AI race

The investigations come amid a broader wave of internal dissent within the AI industry regarding safety. An Anthropic researcher said he resigned over concerns about the technology’s potential to surpass human control.

Jacob Coxon, in a viral X post on Tuesday, said the AI industry was more focused on competition rather than on implementing safeguards. He came to this realisation after spending the last three years doing research at OpenAI and Anthropic.

“The people building AI earnestly believe that it could kill us all by the end of the decade”, Coxon said.

“No other human activity poses this level of danger,” he added, referencing the swift advancement of AI technology.

### ‘We need to act’

In June, Anthropic [proposed](/economy/2026/6/5/anthropic-urges-ai-labs-to-pause-warns-humans-risk-losing-control) a coordinated effort with the world’s leading AI developers to slow down development, warning that humans risk losing control over the technology.

Following the security breach of Hugging Face, OpenAI said it was pushing for mandatory national AI safety requirements and wanted to work with Congress on “capability-based” regulation.

In a statement published on Wednesday, the company said it was formally endorsing four California bills related to safeguards against AI.

“If we cannot meet certain safety bars without slowing down capability growth, we should prioritise the former. The more powerful the technology becomes, the stronger the surrounding safeguards must become,” the statement said.

---

Advertisement

News Source9/10/2026
Read original at News Source

Source coverage

Live

News|Cybersecurity

Full source content

[Live](/video/live)

[Live](/video/live)

[News](/news/)|[Cybersecurity](/tag/cybersecurity/)

# Anthropic discloses 4th AI hacking incident as researcher quits over safety

*Claude Opus 4.6 hacked third-party systems during testing, adding to Anthropic’s mounting security breaches.*

Save

Share

![](/wp-content/uploads/2026/09/image-1788957928.jpg?resize=730%2C410&quality=80)

AI researcher quits Anthropic saying AI race ‘could kill us all’

[![](/wp-content/uploads/2024/05/2017-06-08T131939Z_1189275883_RC11B0FE7F20_RTRMADP_3_GULF-QATAR-JAZEERA-1714943623.jpg?resize=96%2C96&quality=80)](/author/al_jazeera_staff_150119130629458)

By [Al Jazeera Staff](/author/al_jazeera_staff_150119130629458), AP and Reuters

Published On 10 Sep 202610 Sep 2026

Anthropic has reported a fourth incident involving an AI model gaining unauthorised access to the internet during testing, shortly after a researcher quit over concerns about the technology’s rushed development.

An early version of Claude Opus 4.6 hacked into a third-party system in January, the artificial intelligence research company said on Wednesday.

## Recommended Stories

list of 4 items

* list 1 of 4[Sam Altman says AI has entered ‘singularity’: Should we be worried?](/news/2026/7/27/sam-altman-says-ai-has-entered-singularity-should-we-be-worried) * list 2 of 4[Sony, Warner Music sue Anthropic, saying it pirated songs to train its AI](/economy/2026/8/31/sony-warner-music-sue-anthropic-saying-it-pirated-songs-to-train-its-ai) * list 3 of 4[US pushes looser approach to AI regulation, while EU pushes new law](/news/2026/9/2/us-pushes-looser-approach-to-ai-regulation-while-eu-pushes-new-law) * list 4 of 4[OpenAI unveils latest AI model amid rising scrutiny and safety concerns](/economy/2026/9/4/openai-unveils-gpt-6-astra-amid-rising-scrutiny-and-safety)

end of list

The disclosure came after Anthropic [reported](/news/2026/7/31/after-openai-disclosure-anthropic-claude-hacked-outside-systems) several of its Claude models hacked into three company systems during test sessions in July. The previous incidents involved Claude Opus 4.7, Claude Mythos 5 and an internal model.

The incident is part of a growing list of AI models breaking out of testing environments and accessing real computer systems without developer permission.

Some models, designed to complete complex tasks, have learned to communicate with other agents and bend rules, which has spurred criticism of companies including Anthropic, Meta and OpenAI.

A fourth breach of a third-party system went undetected until last month, despite a review of about 141,000 test sessions with AI models.

A set of transcripts was overlooked during the initial review but was identified last month and led to the discovery of the hack. The incidents were caused by a “misconfiguration” during cybersecurity evaluations that allowed the models to access the open internet, according to Anthropic.

In July, OpenAI’s autonomous agents compromised the servers and infrastructure of AI start-up Hugging Face. This security incident prompted a review of AI test sessions to reassess safety protocols.

Advertisement

Anthropic said it tapped research firm METR to investigate the four incidents.

### The AI race

The investigations come amid a broader wave of internal dissent within the AI industry regarding safety. An Anthropic researcher said he resigned over concerns about the technology’s potential to surpass human control.

Jacob Coxon, in a viral X post on Tuesday, said the AI industry was more focused on competition rather than on implementing safeguards. He came to this realisation after spending the last three years doing research at OpenAI and Anthropic.

“The people building AI earnestly believe that it could kill us all by the end of the decade”, Coxon said.

“No other human activity poses this level of danger,” he added, referencing the swift advancement of AI technology.

### ‘We need to act’

In June, Anthropic [proposed](/economy/2026/6/5/anthropic-urges-ai-labs-to-pause-warns-humans-risk-losing-control) a coordinated effort with the world’s leading AI developers to slow down development, warning that humans risk losing control over the technology.

Following the security breach of Hugging Face, OpenAI said it was pushing for mandatory national AI safety requirements and wanted to work with Congress on “capability-based” regulation.

In a statement published on Wednesday, the company said it was formally endorsing four California bills related to safeguards against AI.

“If we cannot meet certain safety bars without slowing down capability growth, we should prioritise the former. The more powerful the technology becomes, the stronger the surrounding safeguards must become,” the statement said.

---

Advertisement

How this page is built

Goose Pod turns cited reporting into a public episode summary first, then pairs that summary with audio playback so listeners can check the source material before they decide how deeply to engage.

The goal is to make this page useful as a news landing page first, while still giving listeners transcript access, related episodes, and direct links back to the original publishers.

Cited sources

9/10/2026

More on this topic

About this page

Goose Pod turns cited reporting into a public episode summary first, then pairs that summary with audio playback so listeners can compare the recap with the underlying source material.

This page reviewed 1 article across 1 source, with the latest cited update on 9/10/2026.

Explore related pages