Claude Sonnet 4
Model ● Covered in 6 stories + Follow
Claude Sonnet 4 is a large language model from Anthropic that has been used as a benchmark in recent AI safety and capability research. The model achieved a 67% win rate in simulated nuclear crisis scenarios and demonstrated competitive performance on coding tasks with a 70.4% score on SWE-bench Verified, though recent studies showed that specialized fine-tuned smaller models can outperform it on specific domain tasks like clinical documentation.
Updated 5 August 2026
Specifications
No specifications recorded yet.
Latest developments
Q2 2026
Q1 2026
Q3 2025
Alibaba releases Qwen3-Coder, a 480-billion parameter open-source coding model with agentic capabilities Model release
Relationships
Products & technology
- Anthropic develops this model · 2 sources