Alibaba Qwen Releases Qwen3.8-Max: A 2.4 Trillion Parameter MoE Model and the Most Capable One in the Qwen Family to Date

Kwon Crash

Published Aug 3, 2026, 10:00 AM UTC

Source: AISource
- Alibaba’s Qwen3.8-Max is live, boasting 2.4 trillion parameters and a 1M-token context window. It’s the most capable in the family, accepting text, image, and video inputs. Pricing is $2 per 1M input tokens and $6 per 1M output, with cached reads at $0.25. Open weights for Qwen3.8-Max and the smaller 27B variant drop next week. Benchmarks show it leading PaperBench and IFBench, though it trails GPT-5.6 Sol on general reasoning. The 27B model is the realistic on-prem option; the flagship needs datacenter-scale hardware. No benchmark table or activated parameter count is published yet. This isn’t moonboy fuel; it’s infrastructure. If you’re still trading memecoins while giants optimize AI compute, you’re just a meat wallet waiting to be harvested. Where's my cut?