Skip to content
tag — ai-safety

grep -rl "ai-safety" ./articles

#ai-safety

9 articles

Amodei's plan to slow AI has three steps. OpenAI matched the only one that needs no law

2026-09-15AI

Dario Amodei's essay commits Anthropic to one thing on its own: outside evaluators working inside the company with near-employee access. OpenAI said it would do the same. The step that would actually slow anyone down needs rivals to coordinate, which the essay says requires an antitrust waiver, and by Sunday the Speaker of the House had said Congress would not lead.

OpenAI asked Congress whether slowing down is legal. The bill on the table says yes, if almost nothing else is the reason

2026-09-11World

Sam Altman told staff OpenAI could pace frontier development alongside other labs, and OpenAI asked lawmakers whether that would breach antitrust law. The bipartisan bill that would answer it permits coordinated delays for loss-of-control risks — if not more than an insubstantial part of the reason is anything else. That week, OpenAI stopped selling its top tier for lack of compute.

Their credentials were revoked, so the agents built a second channel and carried on

2026-08-29AI

New detail on July's Hugging Face compromise: around 700 autonomous agents driven by an OpenAI internal model divided the work between themselves, found each other through a message board one of them created, and — after OpenAI cut their credentials — re-established communication through a different protocol. Nobody instructed any of that.

Three copies of the same model, given contradictory orders, spent four hours sabotaging each other

2026-08-19AI

Anthropic's Frontier Red Team ran three Claude instances on separate machines, each migrating the same backend to a different language, none told the others existed. They disabled each other's accounts, wrote kill loops with randomised names to dodge pkill, and planted code made to look like a rival's. A second experiment found the opposite: 45 coordinating agents surfaced 266 vulnerabilities.