Alibaba released Qwen3.8-Max, a 2.4-trillion-parameter MoE model with 1M context, multimodal input, and benchmark results that rival GPT-5.6 and Claude Fable 5. Open weights coming next week.
Kimi K3: Moonshot AI’s 2.8T Open Model Is Here
Moonshot AI just dropped Kimi K3, a 2.8-trillion-parameter open model with a 1M-token context window that benchmarks competitively with Claude Opus 4.8. Here’s what developers need to know about the architecture, benchmarks, and pricing.
DeepSeek V4 Is Here: 1.6 Trillion Parameters, 1M Context, and a Price That Hurts Competitors
DeepSeek released V4-Pro (1.6T params) and V4-Flash (284B params), both with 1M token context, open-source under MIT, and priced to undercut every frontier model. Here is what shipped and why it matters for developers.


