TLDRocket
Sign in

Microsoft AI Releases Microsoft-Decision-1: A Qwen3.5-9B Decision-Scoring Model

MarkTechPost Asif Razzaq ● Covered by 4 sources

Microsoft launched Decision-1, a model that picks from fixed answers instead of writing text. It’s fast, calibrated, and sold as a new kind of AI for software decisions.

Based on reporting by MarkTechPost, Asif Razzaq — read the original for the full story.

Summary, retelling and take written by AI under human oversight; images are AI-generated illustrations. How we work · Report an error

Microsoft has a new model that doesn’t try to sound smart. It tries to choose. Microsoft-Decision-1 is built for routing, classification, verification and agent control, and it returns calibrated probabilities for fixed options instead of generating prose. The company says that’s the point: software can act on a score much more cleanly than on a paragraph.

The model is post-trained from Alibaba’s Qwen3.5-9B and is available now in Microsoft Foundry and through OpenRouter. Microsoft says future versions will be based on its own MAI models and OpenAI models. The supported use cases in the Foundry card are narrow and practical: yes/no calls, multiple choice, ratings, classification, rubric-style grading of AI responses, groundedness checks against evidence, and explicit abstentions like “cannot tell.” Output is JSON. No explanations. No text generation.

On Microsoft’s own benchmark run, Decision-1 came out on top on average accuracy across 36 benchmarks, reaching 83.5%. The company says it tested 9 systems on 147,137 questions, with the benchmark set kept blind from training. It also claims a p50 latency of 85 ms and p95 of 125 ms. Microsoft says that’s 4.5 times faster than Quyet-1.0-Large and 35 times faster than GPT-6 Sol, which it says took 3.01 seconds.

The catch is that this is a closed API model, not open weights. Microsoft lists no explanations, no images, no audio, no video, and no use for open-ended chat, translation or summarization. It also says the model should not be the sole automated decider for credit, employment, housing, healthcare or legal rights. So yes, this is a very specific product. That specificity is the whole pitch.

My take — AI-written commentary, not fact-checked reporting

This is Microsoft being brutally honest about where a lot of AI should have started: not with rambling chat, but with a score and a threshold. The closed weights and vendor-run benchmarks are the usual cloud-shaped asterisk, of course. Still, the industry could use more tools that know when to shut up and pick one option.

Read more about this at: MarkTechPost

Related stories

The daily briefing

Every AI story that matters, in your inbox by 8am.

TLDRocket reads all relevant sources, removes duplicate coverage, and summarises the day in two minutes. Follow companies and topics for alerts, or get the briefing in Slack. Free, no spam, unsubscribe anytime.