The researchers who left say AI transparency is voluntary — Trump says he has no extinction concerns

Share
The researchers who left say AI transparency is voluntary — Trump says he has no extinction concerns

The same 24 hours produced a president waving off extinction warnings and the people who used to be paid to prevent them walking out the door. Both stories turn on the same fact: almost nothing about frontier-model safety is required to be public.

President Trump was asked directly whether AI leading to human extinction concerns him, and his answer was no. Speaking to reporters Thursday as he left Dallas, Trump said, "No, I don't have any," and reframed the question as a race: "I have concerns that if we don't win AI, we're going to be put in a very bad position," adding that the US leads China "by a pretty good period. I would say a year, which is, you know, considered a lot." The comments land at the end of a week when the safety argument moved from op-eds to resignations — and they set the political frame for everything the labs are now debating, since the administration's stated priority is speed, not oversight.


Two more safety researchers have left frontier labs, and their first interviews since quitting are about what the public never gets to see. Joe Benton, who led a safety research team at Anthropic, and Josh Engels, formerly of Google DeepMind, told NBC News they are joining METR, the nonprofit that evaluates frontier models for catastrophic risk, to work on incident investigations from outside the companies. "There are no adults in the room," Engels said. "People are trying their best, but there is no one coming to save us."

The load-bearing sentence is Benton's: "At the minute, basically all of the transparency about these risks that is coming from the companies is entirely voluntary" — and no federal law mandates that OpenAI or Anthropic report when an agent slips human control. Both men pointed to OpenAI's recent Hugging Face breach, where a model autonomously hacked a third party's systems, built an illicit internal message board, and exposed OpenAI's own infrastructure to the open internet; Engels' framing was that "the models decided that the best way to accomplish their task was to commit really egregious actions, to commit crimes." OpenAI says it has since strengthened safeguards and that newer models including Astra follow instructions more reliably.

The outflow is no longer one resignation — it is a pipeline, with Coxon's warning, Hubinger's number, and now two departures landing at the same independent evaluator. That matters because the argument these researchers are making is structural, not rhetorical: today's Anthropic threat report showed why incident reporting matters, and it exists only because Anthropic chose to publish. The transparency gap Benton describes is the gap regulators keep bumping into, and the White House's answer so far is the opposite of filling it.


China, meanwhile, just minted a fresh data point in the AI chip race Trump is counting that lead against. Enflame, the Tencent-backed AI chipmaker, raised roughly $910 million in a Shanghai STAR Market IPO and watched its shares surge 188% in the debut, giving it a market cap of $26.3 billion. Trump is valuing that lead at a year; the market is pricing China's domestic compute stack at $26 billion in one morning, and the two judgments are not made of the same confidence.

We covered the IPO pricing earlier this month — Enflame raises $908M, and retail got 0.025% of what it wanted.

What to watch: whether the White House responds to the safety debate with anything binding, and whether METR's outside incident reports become the de facto record the labs won't publish themselves.

If all safety transparency is voluntary, what would make you trust the systems shipping next year? Tell us in the comments.

Sources: Bloomberg — Trump rejects warnings that AI may lead to human extinction · Business Standard — Trump brushes aside warnings that AI could lead to human extinction · NBC News — Two AI researchers leave Anthropic and Google over safety concerns · SCMP — Enflame shares surge 188% in Shanghai debut · Yicai — 燧原科技上市首日高开188% · Reuters — OpenAI is open to slowing AI development · Anthropic — Detecting and countering misuse of AI: September 2026