DeepSeek-R1
DeepSeek-R1 is an open-weight reasoning model released by DeepSeek that matches OpenAI's o1 performance while costing approximately $2.19 per million output tokens, nearly 30 times cheaper than o1. The model has been widely deployed across cloud infrastructure and benchmarked for serving optimization, but research has also identified security vulnerabilities including susceptibility to denial-of-service attacks through adversarial prompting and inconsistent instruction-following during reasoning processes compared to final responses.
Updated 3 August 2026
Specifications
No specifications recorded yet.
Latest developments
Validating Distributed LLM Serving Benchmarks with NVIDIA srt-slurm, SLURM Recipes, Parameter Sweeps, and Pareto Analysis
MarkTechPost · 1 week ago ·
28
NVIDIA Vera Rubin Driving Performance Per Watt, Lowest Token Cost for Partners Worldwide
NVIDIA · 1 week ago ·
37
Controlling Reasoning Effort in LLMs
Ahead of AI · 2 weeks ago ·
26
Reinforcement Learning Heats Up, White House Orders Muscular AI Policy, and more...
The Batch ·
28
AI Models Overthink Problems—and It’s a Security Risk
IEEE Spectrum AI · 3 weeks ago ·
26
How catastrophic is your LLM?
Amazon Science · 3 months ago ·
49
Large Reasoning Models Fail to Follow Instructions During Reasoning: A Benchmark Study
Together AI · 9 months ago ·
37
Together AI Delivers Top Speeds for DeepSeek-R1-0528 Inference on NVIDIA Blackwell
Together AI · 1 year ago ·
6
2026
- Validating Distributed LLM Serving Benchmarks with NVIDIA srt-slurm, SLURM Recipes, Parameter Sweeps, and Pareto Analysis
- NVIDIA Vera Rubin Driving Performance Per Watt, Lowest Token Cost for Partners Worldwide
- Controlling Reasoning Effort in LLMs
- Reinforcement Learning Heats Up, White House Orders Muscular AI Policy, and more...
- AI Models Overthink Problems—and It’s a Security Risk
- How catastrophic is your LLM?
2025
- Large Reasoning Models Fail to Follow Instructions During Reasoning: A Benchmark Study
- Together AI Delivers Top Speeds for DeepSeek-R1-0528 Inference on NVIDIA Blackwell
- Qwen3: Think Deeper, Act Faster
- Open R1: Update #3
- Introducing Three New Serverless Inference Providers: Hyperbolic, Nebius AI Studio, and Novita 🔥
- Welcome Fireworks.ai on the Hub 🎆
- How to deploy and fine-tune DeepSeek models on AWS
- Open-R1: a fully open reproduction of DeepSeek-R1
- Welcome to Inference Providers on the Hub 🔥
Relationships
Products & technology
- DeepSeek develops this model · 3 sources
- NVIDIA deploys this model · 1 source
- Zhejiang University deploys this model · 1 source
- Open-R1 derived from this model · 1 source
- Hugging Face deploys this model · 1 source
- AWS deploys this model · 1 source
- OlympicCoder derived from this model · 1 source