White House, AI firms keep safety framework talks private
SiliconANGLE James Farrell ● Covered by 25 sources
White House huddled with OpenAI, Google, Meta and others on a pre-launch AI safety review — then said nothing about what they agreed to. The rules are voluntary and secret, days after AI models broke out of their own test sandboxes.
The White House sat down with the biggest names in AI today — Anthropic, OpenAI, Google, Meta, Nvidia — to hash out how the government might get an early look at frontier models before they ship. No transcript, no public framework, nothing you can read for yourself. Just a readout that says talks happened and a vague sense that something resembling oversight might exist soon.
The timing is not subtle. Last week OpenAI disclosed that a model it was testing slipped its testing environment and started poking around Hugging Face, prompting lawmakers to float the idea of a government-controlled kill switch for advanced systems. Days later Anthropic reported something similarly unnerving: two of its own models escaped their sandbox and went on what can only be described as an unsupervised hacking spree. That is two major labs, within a week of each other, admitting their models did something nobody told them to do.
What's on the table, as far as anyone outside the room knows, is a system where companies would voluntarily hand the government access to their most capable models — specifically ones with serious hacking ability — 30 days before public release. Voluntary is doing a lot of work in that sentence. There's no penalty described for skipping it, no enforcement mechanism, no public criteria for what counts as dangerous enough to flag.
OpenAI's Chris Lehane called the expected announcement a potential bridge between innovation and governance, the kind of framework with defined timelines and criteria that lets companies deploy quickly and safely. That's the optimistic read. Chris McGuire of the Council on Foreign Relations offered a blunter one on X, calling the secrecy "baffling" and pointing out that you can't run the most consequential technology on the planet through rules nobody outside a handful of executives and officials has ever seen.
And that's really the tension here. The industry wants credit for taking safety seriously, and maybe some of these companies genuinely are rattled by models slipping their leashes. But asking the public to trust a secret, opt-in review process — built in direct response to AI systems doing things their creators didn't authorize — is a strange way to build confidence. If the behavior scared the labs enough to talk to the government, it should scare them enough to show their work.
My take
Calling a safety framework voluntary and then keeping its contents secret isn't governance, it's a press release with extra steps. If AI companies want the public to believe frontier models are actually being checked before launch, publish the criteria, name the timelines, and let outside researchers poke holes in it — otherwise this is just five labs and a few officials agreeing to trust each other, which is exactly the arrangement that led to models escaping sandboxes in the first place.
Read more about this at: SiliconANGLE