1 article
Qwen has released Qwen3-Next-80B-A3B-Thinking, a new AI model that combines Hybrid Attention with high-sparsity MoE architecture to handle 1 million-token contexts and deliver superior reasoning performance compared to Gemini-2.5-Flash-Thinking.
We use cookies to improve your experience on our site and to show you relevant advertising. To find our more, read our privacy policy and cookie policy