Artificial Analysis
Company artificialanalysis.ai ● Covered in 26 stories + Follow ✴ AI Graph
Artificial Analysis is a benchmarking and evaluation organization that publishes an “Intelligence Index” and related model-specific indexes. In recent coverage, it updated the Intelligence Index to versions 4.2 and 4.3—shifting weighting toward private held-out test data—and recalculated rankings, including changes at the top between OpenAI’s GPT-6 Astra and Anthropic’s Claude Fable 5.1 based on the revised test evaluations. The company also reported model performance and cost-related results from its benchmarks, including token/cost comparisons and error-rate changes for GPT-6 Astra.
Updated 14 September 2026
Signals
13 stories (+160%)
Media momentum
As of 17 Sep 2026 · Visible stories in the last 30 days, compared with the 30 days before.
12 sources
Source diversity
As of 17 Sep 2026 · Distinct publications behind this entity's visible coverage.
10 events
Release activity
As of 17 Sep 2026 · Model, product and open-source release events in the last 90 days whose coverage involves this entity.
—
Funding signals
As of 17 Sep 2026 · Funding and acquisition events in the last 90 days whose coverage involves this entity.
May 2024 → Sep 2026
Coverage span
As of 17 Sep 2026 · First to most recent month of TLDRocket coverage of this entity.
Management
updated 2 Sep 2026- Micah Smith Co-Founder and CEO
-
Ben Bayliss
Chief Of Staff
-
George Cameron
Co-founder And Product Lead
Latest developments
Google Releases Gemini 3.8 Live and 3.8 Live Extended Thinking for Production Grade Voice Agents
MarkTechPost · 1 day ago ·
8
Introducing Gemini 3.8 Live and 3.8 Live Extended Thinking
Google · 1 day ago ·
4
How GPT-6 Became No. 1: Artificial Analysis CEO Explains The Much-Debated Index Updates
Trending Topics · 3 days ago ·
11
GPT-6 Astra Now on Par With Claude Fable 5.1 in Updated Artificial Analysis Index
Trending Topics · 1 week ago ·
30
GPT-6 Still Behind Fable 5.1 As Artificial Analysis Overhauls Intelligence Index
Trending Topics · 1 week ago ·
28
Artificial Analysis benchmarks GPT-6 Astra vs other agent models
Artificial Analysis · 1 week ago ·
17
GPT-6 Astra Trails Top Models From Anthropic and Meta in Benchmarks
Trending Topics · 1 week ago ·
24
2026
Meta releases Muse Spark 1.3, rolling it out via its Model API and to Meta AI platforms Model release
Google releases Gemini 3.5 Transcribe, a new speech-to-text model for live and non-streaming transcription Model release
DeepSeek released DeepSeek-V4-Flash (0731) and updated API pricing with peak/off-peak rates for its V4 models Pricing change
DeepSeek releases V4-Flash-0731 model with improved agentic and coding performance at reduced pricing Model release
Anthropic releases Claude Opus 5, matching Fable 5 capabilities at half the price Model release
Moonshot AI releases Kimi K3, a 2.8 trillion parameter open-weight model achieving frontier-class performance Open source release
Moonshot AI releases Kimi K3 open model with 2.8 trillion parameters and 1-million-token context window Open source release
- Gemini 3.8 Live adds faster multilingual voice + ‘Live Extended Thinking’
- Google Releases Gemini 3.8 Live and 3.8 Live Extended Thinking for Production Grade Voice Agents
- Introducing Gemini 3.8 Live and 3.8 Live Extended Thinking
- How GPT-6 Became No. 1: Artificial Analysis CEO Explains The Much-Debated Index Updates
- GPT-6 Astra Now on Par With Claude Fable 5.1 in Updated Artificial Analysis Index
- GPT-6 Still Behind Fable 5.1 As Artificial Analysis Overhauls Intelligence Index
- Artificial Analysis benchmarks GPT-6 Astra vs other agent models
- GPT-6 Astra Trails Top Models From Anthropic and Meta in Benchmarks
- Meta says it has caught up with Anthropic and OpenAI with Muse Spark 1.3, its most powerful AI model yet
- Multiverse says its 438B model is fast enough for AI agents. The benchmarks tell a more complicated story.
- Fable 5.1: The New Best AI Model Is Once Again More Expensive
- Intelligent transcription with Gemini 3.5 Transcribe
- Liquid AI Open-Sources Pipette: A Reproducible Benchmarking Suite That Measures On-Device Models, Quantization, Runtime and Hardware Together
- DeepSeek V4 Pro: Rock-bottom Cost Per Task, But Trailing Kimi K3
- DeepSeek-V4-Flash Outshines Pro, The Biggest GitHub Crawl Yet, Engineering System Prompts for Safer Code
- Opus Outshines Even Fable, Inside the Hugging Face Hack, AI Companies Spend Big for Compute
- deepseek-ai/DeepSeek-V4-Flash-0731
- Introducing Claude Opus 5
- [AINews] not much happened today
- Artificial Analysis Reports Kimi K3 Token Efficiency
- [AINews] Kimi K3 2.8T-A50B: the largest open model ever released; Opus 4.8-class at Sonnet 5 pricing
- How Together AI built the world’s fastest speech-to-text stack
- Why Artificial Analysis uses Ai2's IFBench instruction-following eval
- Gemini 3.1 Flash TTS: the next generation of expressive AI speech
2025
2024
Relationships
Products & technology
- Integrated with GPT-6 Astra · 2 sources
- Integrated with Claude Fable 5.1 · 2 sources
- Integrated with IFBench · 1 source
- Derived from DeepSeek-V4-Pro-0813 · 1 source
- Integrated with Intelligence Index · 1 source
- Integrated with Anthropic · 1 source
- Deploys Artificial Analysis Intelligence Index · 1 source
- Integrated with Quasar 438B · 1 source
- Develops Intelligence Index · 1 source
- Derived from GPQA Diamond · 1 source
- Integrated with Terminal-Bench · 1 source
Partnerships
- Partnered with Hugging Face · 1 source