2 articles
A new explanation of ARC-AGI-3 reveals a shift toward efficiency-based scoring using a squared metric, making results from earlier ARC-AGI benchmarks not directly comparable.
Fresh benchmark results from the ARC-AGI-3 challenge at MIT show humans still have the edge over top AI systems when it comes to adaptive reasoning—though that gap is getting smaller as AI learns to learn continuously.
We use cookies to improve your experience on our site and to show you relevant advertising. To find our more, read our privacy policy and cookie policy