OpenAI releases FrontierScience benchmark to measure AI capabilities in scientific research
Benchmark result Provisional 92% confidence first seen
OpenAI developed FrontierScience, a comprehensive benchmark designed to evaluate how well AI systems can perform advanced research tasks in physics, chemistry, and biology. The benchmark includes doctoral-level problems and published research tasks, with testing demonstrating that GPT-5 can optimize molecular biology protocols, and aims to provide a standardized framework for assessing both AI's potential to accelerate scientific research and associated risks.