DeepSeek
DeepSeek is a Chinese AI company that develops large language models and provides API access to them. Recently, the company released V4-Flash-0731, a 284-billion-parameter model with improved agentic and coding capabilities, priced at $0.14 per million input tokens and $0.28 per million output tokens, while its founder disclosed the company operates with approximately one-twentieth of the compute resources available to major U.S. competitors.
Updated 3 August 2026
Signals
27 stories (+2600%)
Media momentum
As of 3 Aug 2026 · Visible stories in the last 30 days, compared with the 30 days before.
26 sources
Source diversity
As of 3 Aug 2026 · Distinct publications behind this entity's visible coverage.
6 events
Release activity
As of 3 Aug 2026 · Model, product and open-source release events in the last 90 days whose coverage involves this entity.
—
Funding signals
As of 3 Aug 2026 · Funding and acquisition events in the last 90 days whose coverage involves this entity.
Mar 2024 → Aug 2026
Coverage span
As of 3 Aug 2026 · First to most recent month of TLDRocket coverage of this entity.
Latest developments
DeepSeek’s smaller model just outperformed its own flagship
The New Stack · 4 hours ago ·
25
DeepSeek V4-Flash brings frontier agent work to bargain pricing
The Neuron · 22 hours ago ·
9
[AINews] not much happened today
Latent Space · 2 days ago ·
48
deepseek-ai/DeepSeek-V4-Flash-0731
Simon Willison · 2 days ago ·
43
DeepSeek Upgrades DeepSeek-V4-Flash-0731 with Major Agentic and Coding Gains
MarkTechPost · 2 days ago ·
48
The Sequence AI of the Week #903: Laguna, the 118 Billion Parameters that Walks Into a Trillion-Parameter Bar
TheSequence · 5 days ago ·
26
Aymo AI Bundles Multiple AI Models Into One Workspace
The Neuron · 1 week ago ·
4
🔮 Copy that: The curious case of AI distillation #594
Exponential View · 1 week ago ·
41
Q3 2026
DeepSeek releases V4-Flash-0731 with improved agentic capabilities through post-training optimization Model release
Poolside AI releases Laguna 118B, an open-weight model demonstrating exceptional efficiency with significant performance gains over much larger competitors Open source release
China regulates anthropomorphic AI services and emergence of facial data licensing marketplace Policy change
Multiple AI model releases and regulatory developments across OpenAI, Anthropic, and competitors Other
Moonshot AI releases Kimi K3, a 2.8 trillion parameter open-weight model achieving frontier-class performance Open source release
Moonshot AI releases Kimi K3, a 2.8 trillion parameter open-weight model with API pricing and delayed weights availability Model release
Moonshot AI releases Kimi K3, a 2.8 trillion-parameter open-weight model competing with frontier American systems Model release
Apple Intelligence receives regulatory approval to launch in China through partnerships with Alibaba and Baidu Feature update
Anthropic alleges Alibaba used fraudulent accounts to extract Claude model data through API Legal action
Analysis reveals Claude Sonnet 5 has higher effective costs and lower quality than predecessor despite lower per-token pricing Feature update
- DeepSeek’s smaller model just outperformed its own flagship
- DeepSeek V4-Flash brings frontier agent work to bargain pricing
- [AINews] not much happened today
- deepseek-ai/DeepSeek-V4-Flash-0731
- DeepSeek Upgrades DeepSeek-V4-Flash-0731 with Major Agentic and Coding Gains
- The Sequence AI of the Week #903: Laguna, the 118 Billion Parameters that Walks Into a Trillion-Parameter Bar
- Aymo AI Bundles Multiple AI Models Into One Workspace
- 🔮 Copy that: The curious case of AI distillation #594
- Chinese Labs' Latest Product? Roleplay
- DeepSeek Founder Liang Wenfeng Resurfaces with Long-term Research Insights
- The U.S. wants to contain China’s AI. Silicon Valley keeps using it
- Last Week in AI #251 - Mythos Back, Sonnet 5, Etched, LongCat
- The Sequence Knowledge #898: The Trace Is the Teacher: Distilling Reasoning Into Small Models
- 7 Consequences of America Finally Losing Its AI Edge to China
- 📈 Data to start your week
- Best Local LLMs You Can Run on a Single 24GB GPU in 2026: Qwen, Gemma, Mistral, DeepSeek Compared
- Inference-Time Scaling and Collective Intelligence for Frontier AI
- 🔮 Kimi K3 surprise & AI economics; the solar paradox; AI's right to learn, cancer vaccine & junior jobs++
- Kimi K3 vs DeepSeek V4 Pro vs GLM-5.2: Open Trillion-Scale MoE Models Compared on Benchmarks, License, and Serving Cost
- Developing Post-Training Technology to Adapt the Largest Open Foundation Model to Country-Specific Specifications
- Kimi K3, and what we can still learn from the pelican benchmark
- Agentic Misalignment in Summer 2026
- Reinforcement Learning Heats Up, White House Orders Muscular AI Policy, and more...
- Apple Intelligence approved for launch in China with Alibaba’s Qwen AI
- OpenAI and Anthropic warn Washington about Chinese distillation of U.S. AI models
- Price per 1M tokens is meaningless
- China’s AI boom is creating a different kind of entrepreneur
Q2 2026
Multiple AI companies release open-source models as pricing competition intensifies Open source release
Amazon researchers present framework for optimizing LLM architecture to improve inference speed without sacrificing accuracy Research publication
- The Unbearable Cheapness of Open Weight Models
- Latest open artifacts (#21): Open model bonanza! Gemma 4, DeepSeek V4, Kimi K2.6, MiMo 2.5, GLM-5.1 & others. On CAISI's V4 assessment.
- Recent Developments in LLM Architectures: KV Sharing, mHC, and Compressed Attention
- Serving DeepSeek-V4: why million-token context is an inference systems problem
- Last Week in AI #340 - OpenAI vs Musk + Microsoft, DeepSeek v4, Vision Banana
- Import AI 455: AI systems are about to start building themselves.
- LWiAI Podcast #243 - GPT 5.5, DeepSeek V4, AI safety sabotage
- DeepSeek-V4 Pro now available on Together AI
- How catastrophic is your LLM?
- Accelerate RL rollouts by up to 50% with distribution-aware speculative decoding
- DeepSeek-V4: a million-token context that agents can actually use
Q1 2026
Q4 2025
Researcher compiles curated list of LLM research papers from July-December 2025 Research publication
- The State Of LLMs 2025: Progress, Problems, and Predictions
- From DeepSeek V3 to V3.2: Architecture, Sparse Attention, and RL Updates
- How to evaluate and benchmark Large Language Models (LLMs)
- Large Reasoning Models Fail to Follow Instructions During Reasoning: A Benchmark Study
Q3 2025
- Fine-Tuning Platform Upgrades: Larger Models, Longer Contexts, Enhanced Hugging Face Integrations
- DeepSeek-V3.1: Hybrid Thinking Model Now Available on Together AI
- The Big LLM Architecture Comparison
- Back to The Future: Evaluating AI Agents on Predicting Future Events
- Meta Superintelligence – Leadership Compute, Talent, and Data
Q2 2025
Relationships
Products & technology
- Develops DeepSeek R1 · 4 sources
- Develops DeepSeek V4 · 4 sources
- Develops Deepseek v4 Pro · 3 sources
- Develops DeepSeek-R1 · 3 sources
- Develops DeepSeek-V4-Flash-0731 · 2 sources
- Develops DeepSeek V4-Flash 0731 · 1 source
- Integrated with Claude · 1 source
- Supplies Huawei · 1 source
- Develops DeepSeek-V3.1 · 1 source
- Develops DeepSeek-V3 · 1 source
- Develops DeepSeek-R1-Distill-Qwen-32B · 1 source
- Develops DeepSeek-V4 · 1 source