AZ Labs
AI Research3 September 20267 min read

Alibaba Qwen Releases Qwen3.8 Max: 2.4-Trillion Parameter Flagship with 1M Multimodal Context

Alibaba Qwen3.8 Max flagship release graphic
Inspect
Alibaba Qwen Official Release Alibaba Qwen Official Release© Alibaba Group

Alibaba's Qwen team has launched Qwen3.8 Max, a 2.4T parameter Mixture-of-Experts model offering 1M token context, native video perception, and deep agent tool orchestration.

smart_toyAlibaba Qwenqwen/qwen3.8-max-09022.4T MoE total (95B active active)
verifiedFirst observed by AIMI: 2026-09-03 23:08:24 SAST
Context Windowarticle
1M tokens (1,000,000)
Max output: 131K tokens (131,072)
INInput Modalitiesinput
multimodal input
descriptiontextvisibilityimagevideocamvideo
OUTOutput Contractoutput
text output
chattext
Route Pricingpayments
$2.00 / $6.00 per 1M tokens
Verified Model Capabilities & Tools
visibilityVision & PerceptionvisibilityVision & PerceptionvideocamVideo InputpsychologyReasoning / ThinkingconstructionFunction Calling & Toolsdata_objectStructured Outputs (JSON)streamToken Streaming
hubAccessible Gateways & Provider Routes (2 Platforms)
Cross-platform AIMI route registry
PlatformRoute IdentifierContextMax OutputRate / TierStatus
Alibaba DashScopeqwen3.8-max1M tokens (1,000,000)131K tokens (131,072)$2.00 / $6.00 per 1M tokens (cached input rate not published by Alibaba)active
OpenRouterqwen/qwen3.8-max-09021M tokens (1,000,000)131K tokens (131,072)$2.00 / $6.00 per 1M tokens ($0.25 cached prompt)active
verified

Key Takeaways

  • check_circleQwen3.8 Max activates 95B parameters out of a 2.4T MoE backbone, outperforming prior generation open models on math, coding, and multilingual reasoning.
  • check_circleSupports a 1,000,000-token context window with high-throughput 131,072-token generation limits.
  • check_circleAlibaba lists $2.00/M input and $6.00/M output for the International deployment scope. Its Global scope is cheaper at $1.65/M and $4.951/M, and it does not publish a cache-hit rate.
  • check_circleAIMI registered the production route on 3 September 2026 at 23:08:24 SAST.

Frontier Multimodal Mixture-of-Experts

Alibaba's Qwen team has launched Qwen3.8 Max as its flagship inference model. Scaling to 2.4 trillion total parameters with 95 billion actively routed per token, Qwen3.8 Max combines massive knowledge density with low per-token compute overhead.

The model excels across complex code generation, database schema extraction, and video visual question answering.

Frequently Asked Questions

What are the specs of Qwen3.8 Max?

Qwen3.8 Max features 2.4T total MoE parameters (95B active), a 1M token context window, and native text, image, and video input capabilities.

Explore verified specifications, benchmark results, and route pricing across alternative models in this class.

Primary Sources

Share this articlePost on X
arrow_backBack to all news