Articles Tagged "MoE"

Qwen3.8-Max-Preview

Qwen3.8-Max-Preview

Alibaba's 2.4 trillion parameter multimodal MoE model claims to trail only Claude Fable 5, but ships with no model card, no benchmark table, and no confirmed pricing.

Kimi K3

Kimi K3

Moonshot AI's Kimi K3 is a 2.8 trillion parameter MoE model that tops LMArena's Frontend Code Arena and nears Claude Fable 5 on intelligence benchmarks, but at roughly triple Kimi K2.6's price and a higher hallucination rate.

Inkling

Inkling

Thinking Machines Lab's first open-weight model - a 975B-parameter MoE with native text, image, and audio reasoning, released under Apache 2.0 and tuned for customization on the Tinker platform.

Gemini 3 Pro

Gemini 3 Pro

Google DeepMind's Gemini 3 Pro debuted at 1501 Elo on LMArena with 91.9% on GPQA Diamond and a 1M-token context window, before Google retired it for Gemini 3.1 Pro.

LongCat-2.0 Review: China's Stealth Coder

LongCat-2.0 Review: China's Stealth Coder

Meituan's 1.6T open-source coding model secretly topped OpenRouter for two months before revealing itself - and the price-to-performance math is hard to argue with.

LongCat-2.0

LongCat-2.0

Meituan's 1.6T-parameter open-source MoE coding model, trained end-to-end on 50,000 domestic Chinese ASICs, with native 1M token context and a 59.5 SWE-bench Pro score.