2 articles
Google's Gemini 3 Deep Think achieved 84.6% on the ARC-AGI-2 reasoning benchmark, marking a significant leap in AI capabilities and intensifying the race among tech giants.
Google's Gemini Deep Think models achieve breakthrough performance in mathematical proof generation, significantly outpacing competitors like GPT-5 and Claude on IMO-level reasoning tasks.
We use cookies to improve your experience on our site and to show you relevant advertising. To find our more, read our privacy policy and cookie policy