EmbeddingGemma 2
Simon Willison’s Weblog Simon Willison ● Covered by 3 sources
Opinion — commentary, not a factual news event.
EmbeddingGemma 2 is out under Apache 2.0. That matters because embedding models can leave you stuck paying to redo millions of vectors if a vendor pulls the plug.
Based on reporting by Simon Willison’s Weblog, Simon Willison — read the original for the full story.
Summary, retelling and take written by AI under human oversight; images are AI-generated illustrations. How we work · Report an error
EmbeddingGemma 2 comes with an Apache 2.0 license, and that is the part that stands out here. For embedding models, open weights are not some abstract ideology; they are insurance against getting trapped by a provider’s future product plans.
That matters because these models are often used at huge scale. A real deployment can mean thousands or even millions of embedding vectors, all stored for later comparison. If the model is closed and hosted-only, then a provider can eventually stop offering it, swap in something newer, and leave customers with a bill for recalculating everything they already stored.
The source points to OpenAI’s April 2024 promise to cover the financial cost of re-embedding with new models, but treats that as an exception rather than a rule. The broader point is simpler: don’t build on the hope that every vendor will be generous when the model changes.
There is a nice middle ground here. A team can still pay someone to host the model, and still avoid lock-in if the weights are open. If that host disappears, the open version can be run elsewhere, which is exactly the kind of boring freedom infrastructure should have.
My take — AI-written commentary, not fact-checked reporting
Closed embedding models are a neat way to turn a routine infrastructure choice into a future hostage situation. The funny part is that the expensive bit is often not using the model, but getting forced to use it again later. Open weights are just common sense with fewer surprise invoices.
Read more about this at: Simon Willison’s Weblog
Related stories
Cohere Releases Embed 5: How It Compares to Voyage 4 Large, Gemini Embedding 2, and OpenAI
MarkTechPost · 5 days ago ·
46
Gemma 4: Byte for byte, the most capable open models
Google DeepMind · 6 months ago ·
26
Gemma Scope 2: helping the AI safety community deepen understanding of complex language model behavior
Google DeepMind · 9 months ago ·
10