Our First Proof submissions
OpenAI Blog
DeepSeek released its first proof submissions for the First Proof math competition, a challenge designed to evaluate advanced AI reasoning on difficult mathematical problems. The submissions represent a research-grade test of the model's capabilities on expert-level mathematics without specifying performance metrics or outcomes. This represents an early data point in assessing how current AI systems handle formal mathematical reasoning compared to human mathematicians.
Why it matters
We share our AI model’s proof attempts for the First Proof math challenge, testing research-grade reasoning on expert-level problems.