Simon Willison's Weblog·1 month ago·
40
● 3 sources
Mira Murati's Thinking Machines Lab released Inkling, an open-weights multimodal transformer with 975 billion total parameters and 41 billion active parameters, licensed under Apache 2.0 and trained on 45 trillion tokens. A smaller version with 276 billion parameters is in testing. The model is positioned as a strong base for fine-tuning rather than a frontier model and competes with other open-weights alternatives like NVIDIA Nemotron and Gemma 4.
Thinking Machines launched Inkling, its first open-weights model with a 1M-token context window supporting text, images and audio, available on the Tinker fine-tuning platform. The model is positioned for custom fine-tuning as startups increasingly shift workloads from frontier models to self-hosted versions, with alternatives like GLM-5.2 gaining adoption despite lacking vision capabilities. The release reflects a growing market trend toward open-source and customized AI models rather than reliance on leading proprietary systems.
Every AI story that matters,
in your inbox by 8am.
TLDRocket reads all relevant sources, removes duplicate coverage, and summarises the
day in two minutes. Follow companies and topics for alerts, or get the
briefing in Slack. Free, no spam, unsubscribe anytime.
Reading TLDRocket needs no cookies, and the readership counts we rely on come from
our own cookieless analytics. Google Analytics is the exception: it sets cookies and
reports to Google, so it stays switched off until you allow it. You can change your
mind any time from “Cookie settings” in the footer.
Strictly necessary
Session security and form protection (tldrocket-session,
XSRF-TOKEN, 2 hours). The site cannot work without them,
so they need no consent.
Always on
Google Analytics 4 (_ga,
_ga_<id>, up to 2 years). Measures which
stories and sections readers use. Google acts as a third-party processor and may
store the data outside the EU. No advertising, no profiling, no data sold.