Microsoft Tells Engineers ‘Tokenmaxxing Is Not What We Are Optimizing For’
404 Media Emanuel Maiberg
Microsoft just told its own engineers to stop maxing out AI token usage at work. Turns out even the company selling AI infrastructure can't stomach the bill for using it internally.
Microsoft has drawn a line on how freely its engineers can burn through AI tokens, and the message from leadership is blunt: unlimited AI usage was never the point. In an internal email, EVP Jay Parikh told staff that "tokenmaxxing is not what we are optimizing for," a phrase that will likely follow the company around for a while. He wants employees chasing outcomes, not consumption, and he framed token spend as just another resource that needs the same discipline Microsoft applies to headcount or cloud costs.
The practical changes are modest but telling. GPT-5.6, a cheaper model than some of the alternatives, becomes the default for internal Copilot use starting in July 2026. Divisions will get token budget targets, and individual engineers can now track their own spending, some of which reportedly runs from a few hundred to several thousand dollars a month. No hard caps yet, but the guidelines hint that restrictions could tighten depending on what the usage data shows.
What makes this notable isn't the policy itself but who's writing it. Microsoft just posted a strong quarter, revenue and profit both up, beating Wall Street. This isn't a company pinching pennies out of necessity. It's a company that has poured billions into AI infrastructure and subsidized inference costs for customers, now turning around and telling its own engineers to cool it. Microsoft joins Amazon, Adobe, Atlassian, and Citi, all of whom have quietly started reining in the same kind of maximalist AI use they spent the last two years encouraging.
An anonymous Microsoft employee put it more sharply than any press release would: if the company hosting the AI infrastructure can't afford to let its own people use it freely, what does that say about everyone else buying the product? Parikh insists this isn't about slowing down Microsoft's push to be "AI-first," and that the goal is more impact per token rather than fewer tokens overall. But the optics are hard to ignore. The company selling the shovels is now rationing them for its own miners.
My take
This is the tell everyone should have been watching for. Every AI vendor talks like tokens are basically free once you're locked into their platform, and then the moment the internal bill lands on a real budget line, suddenly discipline matters again. If Microsoft, sitting on record profits and its own AI supply chain, still needs to throttle engineers, the pitch to smaller companies that AI tools pay for themselves deserves a lot more skepticism than it's getting.
Read more about this at: 404 Media