Anthropic and Meta update AI safety governance frameworks with enhanced capability thresholds and risk assessment processes
Policy change Provisional 78% confidence first seen
Anthropic updated its Responsible Scaling Policy to introduce more flexible capability thresholds and refined safeguard assessment processes for frontier AI systems, particularly around autonomous AI R&D and CBRN weapon assistance. Meta simultaneously published an Advanced AI Scaling Framework that broadens safety evaluations for its most capable models, including assessments of chemical, biological, and cybersecurity risks, with plans to publish Safety & Preparedness Reports for each advanced model.