6 articles
GPT-5.2 Pro just scored 31% on the brutal FrontierMath Tier 4 test—a 12-point jump over anything we've seen before. It's the clearest sign yet that AI is getting genuinely better at advanced math, not just memorizing patterns.
AI has hit a major milestone by solving a longstanding Erdős problem in number theory, showing that machine intelligence can now tackle complex mathematical proofs independently.
OpenAI's GPT-5.2 Pro has independently solved Erdős problem #729 using formal verification tools, marking a breakthrough moment in AI's ability to tackle unsolved mathematical challenges without human guidance.
Advanced AI systems have independently produced formally verified solutions to two previously unsolved Erdős problems, demonstrating a significant leap in machine-driven mathematical discovery.
GPT-5.2 Pro grabbed first place on the FrontierMath benchmark with a 29.2% accuracy score, beating out other cutting-edge AI models on one of the toughest math reasoning tests available.
OpenAI's GPT-5.2 Pro achieved the highest score on the Mensa Norway benchmark, leading a comparative ranking of advanced AI models by reasoning performance.
We use cookies to improve your experience on our site and to show you relevant advertising. To find our more, read our privacy policy and cookie policy