TLDRocket
Sign in

Anthropic and Meta update AI safety governance frameworks with enhanced capability thresholds and risk assessment processes

Policy change Provisional 78% confidence first seen

Anthropic updated its Responsible Scaling Policy to introduce more flexible capability thresholds and refined safeguard assessment processes for frontier AI systems, particularly around autonomous AI R&D and CBRN weapon assistance. Meta simultaneously published an Advanced AI Scaling Framework that broadens safety evaluations for its most capable models, including assessments of chemical, biological, and cybersecurity risks, with plans to publish Safety & Preparedness Reports for each advanced model.

Source coverage

The daily briefing

Every AI story that matters, in your inbox by 8am.

TLDRocket reads 60+ sources, removes duplicate coverage, and summarises the day in two minutes. Free, no spam, unsubscribe anytime.