2 articles
Gemini 3 Pro scored 54 percent on the ARC-AGI-2 reasoning test once Poetiq added it to their system. It became the first model to beat the 50 percent mark, which shows that AI reasoning ability is now advancing fast.
Poetiq sets a new benchmark record on ARC-AGI-2 with a verified 54% score, surpassing Alphabet's Gemini 3 Deep Think. The system uses an open-source scaffold combining Gemini 3 Pro and GPT-5.1.
We use cookies to improve your experience on our site and to show you relevant advertising. To find our more, read our privacy policy and cookie policy