Microsoft AI Releases Microsoft-Decision-1: A Qwen3.5-9B Decision-Scoring Model
MarkTechPost Asif Razzaq ● Covered by 4 sources
Microsoft launched Decision-1, a model that picks from fixed answers instead of writing text. It’s fast, calibrated, and sold as a new kind of AI for software decisions.
Based on reporting by MarkTechPost, Asif Razzaq — read the original for the full story.
Summary, retelling and take written by AI under human oversight; images are AI-generated illustrations. How we work · Report an error
Microsoft has a new model that doesn’t try to sound smart. It tries to choose. Microsoft-Decision-1 is built for routing, classification, verification and agent control, and it returns calibrated probabilities for fixed options instead of generating prose. The company says that’s the point: software can act on a score much more cleanly than on a paragraph.
The model is post-trained from Alibaba’s Qwen3.5-9B and is available now in Microsoft Foundry and through OpenRouter. Microsoft says future versions will be based on its own MAI models and OpenAI models. The supported use cases in the Foundry card are narrow and practical: yes/no calls, multiple choice, ratings, classification, rubric-style grading of AI responses, groundedness checks against evidence, and explicit abstentions like “cannot tell.” Output is JSON. No explanations. No text generation.
On Microsoft’s own benchmark run, Decision-1 came out on top on average accuracy across 36 benchmarks, reaching 83.5%. The company says it tested 9 systems on 147,137 questions, with the benchmark set kept blind from training. It also claims a p50 latency of 85 ms and p95 of 125 ms. Microsoft says that’s 4.5 times faster than Quyet-1.0-Large and 35 times faster than GPT-6 Sol, which it says took 3.01 seconds.
The catch is that this is a closed API model, not open weights. Microsoft lists no explanations, no images, no audio, no video, and no use for open-ended chat, translation or summarization. It also says the model should not be the sole automated decider for credit, employment, housing, healthcare or legal rights. So yes, this is a very specific product. That specificity is the whole pitch.
My take — AI-written commentary, not fact-checked reporting
This is Microsoft being brutally honest about where a lot of AI should have started: not with rambling chat, but with a score and a threshold. The closed weights and vendor-run benchmarks are the usual cloud-shaped asterisk, of course. Still, the industry could use more tools that know when to shut up and pick one option.
Read more about this at: MarkTechPost
Related stories
Liquid AI Releases d1: A Decision Model That Returns Calibrated Probabilities With Zero Output Tokens
MarkTechPost · 1 week ago ·
5
Milliseconds.ai
Product Hunt · 2 weeks ago ·
43
AWS Strands Labs Releases Strands Decider 2B: An Open Source Decision Model That Picks Options in About 115 ms
MarkTechPost · 1 week ago ·
48