1 article
A new synthetic dataset called Chimera shows that compact language models can rival far larger systems in complex reasoning tasks. Training on just 9,225 curated problems allowed a 4B model to approach the performance of models up to 50 times its size.
We use cookies to improve your experience on our site and to show you relevant advertising. To find our more, read our privacy policy and cookie policy