TLDRocket
Sign in

China Thwarts Meta’s Agentic Ambition, U.S. Evaluates Upcoming Models, AI Diagnoses Mammograms

The Batch Analytics DeepLearning.AI

Washington will now test big AI models for security risks before they hit the market. That's a hard reversal from the hands-off approach the White House was pushing just months ago.

Based on reporting by The Batch, Analytics DeepLearning.AI — read the original for the full story.

Summary, retelling and take written by AI under human oversight; images are AI-generated illustrations. How we work · Report an error

The National Institute of Standards and Technology just flipped a switch that had been set to 'stay out of the way' since January. A new group called TRAINS, run out of NIST's Center for AI Standards and Innovation, will check upcoming AI models for cybersecurity, biosecurity, and chemical-weapons risks before they ever reach the public. Google, Microsoft, and xAI have agreed to hand over versions of their models with weaker or no guardrails so testers can see what the systems can actually do; Anthropic and OpenAI signed similar deals back in 2024.

What makes TRAINS different from other NIST efforts is who's in the room. It pulls in the Departments of Commerce, Defense, Energy, and Homeland Security, plus the National Security Administration and the National Institutes of Health, and it's built to move fast rather than run the usual slow bureaucratic evaluation cycle. NIST hasn't said which benchmarks TRAINS will actually use, though it has pointed to an earlier CAISI comparison of DeepSeek V4 Pro against rivals using nine public benchmarks and an internal test called PortBench.

The timing isn't random. A month ago Anthropic raised alarms in government circles when it revealed that its unreleased Claude Mythos Preview model could find and exploit vulnerabilities in widely used software. That model has already been shared with 50 organizations using it to patch their own systems, and Anthropic wanted to expand access to 70 more. Last week the White House said no, citing national security worries and doubts about whether Anthropic has enough computing power to serve both those users and government needs. Anthropic hasn't said whether it will push back on that decision.

There's a bigger arc here too. Earlier this year Anthropic tried to restrict military use of Claude for surveillance and autonomous weapons, and the administration responded by banning the model from military use altogether rather than accepting limits. Combine that with the new pre-release testing regime and an executive order reportedly under consideration that would make model approval mandatory, and the picture is one of a government that spent its first year in office tearing down Biden-era AI rules, only to start rebuilding a different set of guardrails once models started showing they could do real damage.

None of the companies involved are required to submit models yet — these are voluntary agreements for now. But if that executive order lands, voluntary becomes mandatory, and the shape of who gets to release what, and when, starts running through Washington instead of Silicon Valley boardrooms.

My take — AI-written commentary, not fact-checked reporting

Standardized safety testing sounds sensible until you notice who's designing the test and who gets to fail it. Handing pre-release approval power to a rotating cast of federal agencies is a great way to slow down the companies that play by the rules while doing nothing about actors who don't bother asking permission. If the real goal is safety, that's better solved by competitive, transparent benchmarks the whole industry can see and challenge — not a closed-door national security review that conveniently also happens to make life harder for open-source competitors.

Read more about this at: The Batch

Related stories

The daily briefing

Every AI story that matters, in your inbox by 8am.

TLDRocket reads all relevant sources, removes duplicate coverage, and summarises the day in two minutes. Follow companies and topics for alerts, or get the briefing in Slack. Free, no spam, unsubscribe anytime.