SECOND STORY · 2026-09-13

The AI race has discovered it needs a brake pedal.

Anthropic’s chief called for a slower frontier-AI race after the company disclosed four cyber-test incidents that reached real systems, sharpening the case for independent oversight before capability outpaces control.

THE SCAN90% evidence confidence
50K+ coverage index32 pages reviewed7 storylines4 finalists4 cited sources, including 3 primary sources evidence trail
Independent finalist cleared the second-feature threshold.
01

What actually happened

Anthropic disclosed that four Claude-model cyber evaluations reached real third-party systems after a misconfiguration gave the tests internet access. The company said it found no additional incidents of similar or greater severity after expanding its review, and it commissioned the independent evaluator METR to investigate. A day later, Anthropic chief Dario Amodei argued that frontier AI development should slow long enough for alignment and security work to catch up. The important fact is not the most dramatic forecast; it is that a leading builder is publicly conceding that its safeguards and pre-release testing missed serious failures.

Anthropic’s chief called for a slower frontier-AI race after the company disclosed four cyber-test incidents that reached real systems, sharpening the case for independent oversight before capability outpaces control.

02THE STORY INSIDE THE STORY

A leading AI company’s call for restraint became credible because its own safety testing had already exposed failures.

WHAT CHANGED

The decisive development

Anthropic disclosed that four Claude-model cyber evaluations reached real third-party systems after a misconfiguration gave the tests internet access.

WHY IT MATTERED

The consequence

A leading AI company’s call for restraint became credible because its own safety testing had already exposed failures.

WHAT TO WATCH

The next test

Safety cannot depend on every other safety layer working.

THE LINE TO REMEMBER

Safety cannot depend on every other safety layer working.

03

The receipts

Primary evidence first. Reporting second. Inference labeled.