2026 in LLMs (so far)
Simon Willison’s Weblog Simon Willison
A keynote recap outlined how LLM progress and agent tools in 2026 made coding agents noticeably more reliable and drove new “Claw” software adoption. The speaker said AI token spending went from under $50 in the past year to $1,000 in a day after coding agents became broadly useful. The result was product-market fit centered on coding agents, followed by companies tightening token/spend rules to manage costs.
Why it matters
On Friday I gave the closing keynote at the WeAreDevelopers World Congress North America in San Jose. I tied together the key trends from the past year into a chronological exploration of everything that happened in 2026. The video is on YouTube; here are my annotated slides and notes to accompany the talk. # I'm going to give a lightning tour of everything that has happened so far in 2026. The year isn't over yet! # For me, 2026 started a couple of months earlier in November 2025. # November saw the release of two important models: Claude Opus 4.5 and GPT-5.1. As is usually the case with new models, these were incremental improvements on the models that came before them. But every now and then when a model improves, it crosses an invisible line where something that didn't really work starts working. In this case, the thing that started working was their coding agents. Claude Code had been around since February 2025, Codex was a little younger. These two new models, when paired with th