1 article
New testing across GPQA Diamond and MMLU-Pro benchmarks reveals that giving AI systems expert personas doesn't improve factual accuracy. Performance stayed virtually the same across all domains and prompt variations.
We use cookies to improve your experience on our site and to show you relevant advertising. To find our more, read our privacy policy and cookie policy