1 article
MARSHAL, a new framework from Tsinghua University and partners, improves LLM performance by 28.7% through self-play. It strengthens multi-agent reasoning and boosts results on key benchmarks.
We use cookies to improve your experience on our site and to show you relevant advertising. To find our more, read our privacy policy and cookie policy