40 / 2418

AI safety requires more than just slowing our pace | Stuart Russell

TL;DR

AI researcher Stuart Russell argues that safety requirements must be non-negotiable and tied to concrete, verifiable goals instead of a slower timeline alone. His op-ed follows a week of turmoil in AI, triggered by safety researcher Jacob Coxon's resignation from Anthropic. It also comes after weeks of disturbing revelations about the OpenAI and Hugging Face incident. In Russell's view, adjusting the pace without measurable criteria leaves the underlying problem unsolved.

Nauti's Take

Russell's push for measurable safety goals is real progress, because it turns a vague pause debate into criteria that regulators and companies can actually check. The hard part is defining those tests for systems whose failure modes are still poorly understood, and labs may treat any benchmark as a box to tick.

Teams deploying agents can apply the same logic today: set concrete release criteria instead of trusting a vendor's timeline.

Sources