Deep research System Card
OpenAI Blog
Anthropic released a system card documenting the safety evaluation and testing conducted before launching its deep research tool. The evaluation included external red teaming exercises and assessments using the company's Preparedness Framework to identify frontier risks. The document details the specific mitigations built into the system to address identified risk areas.
Why it matters
This report outlines the safety work carried out prior to releasing deep research including external red teaming, frontier risk evaluations according to our Preparedness Framework, and an overview of the mitigations we built in to address key risk areas.