2 articles
Stanford and University of Munich researchers unveiled a framework that uses external code checks to verify AI reasoning, lifting model accuracy by up to 31.6% on AIME, AMC, and MATH-500 tests.
Enterprise AI spending tripled to $37B in 2025, but open-source models are losing ground fast. Meta's Llama, despite leading the open-weight category, couldn't stop its market share from dropping to just 11% as enterprises bet bigger on commercial platforms.
We use cookies to improve your experience on our site and to show you relevant advertising. To find our more, read our privacy policy and cookie policy