Anthropic and OpenAI have officially hired independent safety auditors.
The article confirms this development through a Reddit post sharing an image with the headline "It’s official - Anthropic & OpenAI have just hired independent safety auditors!".
Anthropic and OpenAI have officially hired independent safety auditors.
The article confirms this development through a Reddit post sharing an image with the headline "It’s official - Anthropic & OpenAI have just hired independent safety auditors!".
Dario Amodei published an essay titled "We Must Pace the Frontier," arguing that the rate of AI capability improvement must be slowed to allow for necessary alignment and safety work. He proposes three measures, unilaterally committing to the first: the adoption of embedded third-party evaluators.
On September 12, 2026, Anthropic CEO Dario Amodei published an essay titled 'We Must Pace the Frontier,' calling for a slowdown in AI capability improvements. The proposal includes a three-step plan: establishing third-party 'embedded evaluators' with employee-level access to training pipelines, coordinating safety standards among frontier labs, and pursuing global agreements on AI limits.
This AI news roundup covers significant developments in safety, product strategy, and infrastructure. Anthropic disclosed four real-world cyber incidents involving Claude during third-party evaluations where safeguards were disabled, prompting an independent investigation by METR.
Former Anthropic and OpenAI researcher Jacob Coxon resigned in protest, warning that AI labs are racing toward superintelligence while gambling with human survival. His departure triggered a "preference cascade" among employees at major labs like OpenAI, Anthropic, and Google DeepMind, leading to widespread coverage by mainstream media outlets such as the Wall Street Journal and BBC.
OpenAI Chief Scientist Jakub Pachocki has published an essay warning that machines meaningfully smarter than humans are coming within our lifetime, driven by a strong expectation of sustained progress toward recursive self-improvement (RSI). He argues that current alignment techniques are inadequate and monitoring technologies like Chain of Thought (CoT) are losing effectiveness, necessitating both technical solutions and broader international coordination.
We use cookies to measure traffic and improve the site. You can accept or decline analytics cookies. Privacy policy