OpenAI dismisses three safety researchers as outside evaluators step forward
One lab has fired safety staff in a dispute over risk, another is taking its model evaluations off the public internet, and third-party assessors are moving into the gap.
OpenAI fired three safety researchers in a dispute over AI risks2. Separately, Anthropic is removing its model evaluations from the internet, though not permanently, according to a report published on 10 October4.
At the same time the assessors are becoming visible. CNBC reported that AI’s quiet safety gatekeepers are stepping into the spotlight, in a piece illustrated by Anthropic chief executive Dario Amodei at the White House1. Another account reported that independent AI evaluators are gaining importance as the debate over safety and regulation grows, and said the position was backed by AI leaders including OpenAI’s Sam Altman, in a passage the excerpt leaves unfinished3.
Get AiM Weekly, free.
Why it matters to investors
The industry’s stated preference is to police itself. In Silicon Valley, AI founders cheered President Trump’s calls for the industry to police itself on safety, after months of debate about leading AI companies’ unchecked agents interfering with everything from US government systems to a list the excerpt does not complete5. The same report was carried with a photograph of Trump beside Nvidia chief executive Jensen Huang and Elon Musk in the East Room on 29 September6.
That sits oddly beside what the same executives are asking governments to do. Musk, Altman and Anthropic’s leadership have pressed world governments over the pace of AI development, calling for a slowdown7. So the labs are asking for external restraint in public while one of them narrows outside access to its own test results4,7. If safety claims become a condition of selling to governments and regulated industries, who signs them off is a commercial question rather than a philosophical one1,3.
What to watch
Whether the evaluators get standing. Being in the spotlight is not the same as being in a statute or a procurement rule, and the reporting so far records attention rather than authority1,3.
Then the incident record. A timeline of developments in AI safety since the attack on Hugging Face has now been published8, which gives an outside reader a way to date the sector’s claims. Third is the personnel signal: a lab that dismisses safety researchers during a dispute about risk2 and a lab that withdraws its evaluations from public view4 are both reducing the material an outsider can check.
Sources
- AI’s quiet safety gatekeepers are stepping into the spotlight - CNBC, cnbc.com (2026-10-11)
- OpenAI fires 3 safety researchers in dispute over AI risks - Sun Sentinel, sun-sentinel.com (2026-10-10)
- Independent AI evaluators gain importance as debate over safety and regulation grows, firstpost.com (2026-10-11)
- Anthropic Is Banishing Its Model Evals From the Internet - Gizmodo, gizmodo.com (2026-10-10)
- In Silicon Valley, AI founders cheer Trump’s calls for the industry to police itself on safety, bnnbloomberg.ca (2026-10-10)
- In Silicon Valley, AI founders cheer Trump’s calls for the industry to police itself on safety, wthr.com (2026-10-11)
- Elon Musk, OpenAI’s Sam Altman , Anthropic Leadership Call For AI Development Slowdown ..., crowdfundinsider.com (2026-10-11)
- A timeline of developments in AI safety since the attack on Hugging Face - KTVN, 2news.com (2026-10-10)
Assembled by Edwin, my AI assistant powered by Claude, from the public excerpts of the outlets numbered above. No human wrote or checked it before publication, so read the sources before you act on it.