TLDRocket
Sign in

Qwen2.5-Math: The world's leading open-sourced mathematical LLMs

Qwen

Alibaba released Qwen2.5-Math, an updated series of open-source mathematical language models in sizes 1.5B, 7B, and 72B parameters. The 72B-Instruct model achieved a score of 92.9 on the MATH benchmark using tool-integrated reasoning and solved 12 problems on the AIME 2024 exam compared to 1-2 problems for GPT-4 and Gemini models. The models now support both English and Chinese math problems using chain-of-thought and tool-integrated reasoning approaches, with the 7B model matching previous 72B model performance.

Why it matters

GITHUB HUGGING FACE MODELSCOPE DISCORD 🚨 Qwen2.5-Math mainly supports solving English and Chinese math problems through CoT and TIR. We do not recommend using this series of models for other tasks. Introduction A month ago, we released the first series of mathematical LLMs - Qwen2-Math - of our Qwen family. Today, we have upgraded it and open-sourced Qwen2.5-Math series, including base models Qwen2.5-Math-1.5B/7B/72B, instruction-tuned models Qwen2.5-Math-1.5B/7B/72B-Instruct, and mathematical reward model Qwen2.

Related stories

The daily briefing

Every AI story that matters, in your inbox by 8am.

TLDRocket reads 60+ sources, removes duplicate coverage, and summarises the day in two minutes. Free, no spam, unsubscribe anytime.