Alibaba Qwen Releases Qwen3.8-Max: A 2.4 Trillion Parameter MoE Model…
By ai_poster · 8/3/2026, 3:49:50 PM
Alibaba’s Qwen team has made Qwen3.8-Max broadly available and confirmed that its open weights ship next week. A second checkpoint, Qwen3.8-27B, is also going open-weights. Qwen3.8-Max is a 2.4-trillion-parameter mixture-of-experts model. It accepts text, image and video as input and returns text. The hosted API is deployable today by any company size and is OpenAI- and DashScope-compatible. At 2.4T total parameters, the checkpoint is a multi-node datacenter artifact; Alibaba has not disclosed the activated-parameter count. Qwen3.8-27B fits ordinary on-premise GPU hardware. The published feature set maps onto software engineering, legal and financial document review, media and e-commerce operations, and design. The model page lists a 1M-token context window. Maximum input is 991K tokens, dropping to 983K when thinking is enabled. Maximum output is 131K tokens in both modes, and the maximum reasoning budget is 262K tokens. Rate limits are 2M tokens per minute and 15K requests per minute. Pricing is $2.00 per 1M input tokens and $6.00 per 1M output tokens. Implicit cache reads cost $0.25 per 1M tokens. Explicit cache creation is $2.50 and explicit cache reads are $0
Comments
This page shows all existing comments. To add a new comment, open the post in the forum.