Qwen3.8 Max
Released Jul 19, 2026 · knowledge cutoff unpublished
- Status
- Superseded
- Location
- China
- Modality
- Multimodal
- Context window
- 984K
- Max output
- 131K
- Speed
- —
- Price ($/MTok in / out)
- $2 / $6
- Cost per task
- —
- Open weights
- No
Benchmarks4 of 8 reported
GPQA Diamond
Terminal-Bench 2.1
Humanity's Last Exam
LMArena Elo
MMLU-Pro
SWE-bench Verified
AIME (latest)
ARC-AGI-2
Strengths & weaknesses
Strengths
- Takes text, images, video and documents in one system — the first Qwen flagship to do all four
- Available same-day inside Alibaba's own coding and agent platforms
Weaknesses
- A preview described as continuously evolving — behavior is a moving target, not a fixed release
- No model card or active-parameter count published, so its real speed and cost profile are unknown
- Reachable only through subscription plans rather than a per-token endpoint
Announced 2026-07-19 at WAIC Shanghai as a preview, then given a general-availability launch with standard API pricing on 2026-08-03 (Alibaba Cloud Model Studio blog: "Qwen3.8-Max: A New Bar for Coding and Cowork"). Alibaba's own release benchmark table (self-reported, published as an image and not independently re-derivable) claims GPQA Diamond 92.6, Terminal-Bench 2.1 86.6, HLE 43.6, and SWE-bench Pro 67.7 (not the SWE-bench Verified variant this site tracks, so left unset here) — relayed via MarkTechPost and OfficeChai coverage of the announcement. Artificial Analysis's own independent evaluation (artificialanalysis.ai/models/qwen3-8-max and its per-benchmark leaderboards) measured GPQA Diamond 92.7 (close to Alibaba's claim) but Terminal-Bench v2.1 81.3 and HLE 41.4 — both notably below Alibaba's self-reported figures — so the AA-measured numbers are recorded here as the more independently verifiable source, with the discrepancy noted. Pricing ($2.00/$6.00 per 1M input/output tokens) is Alibaba Cloud Model Studio's Singapore/international rate; the China Beijing, Frankfurt, Virginia, Tokyo and Hong Kong regions price lower at $1.65/$4.951. Open weights, promised "next week" as of the Aug 3 announcement, arrived on 2026-08-13 — but as `qwen3-8-2-4t-a95b`, a reduced text-only variant with a 262K native window rather than this multimodal 1M-context Max build, which stays closed. Context window (983,616) and max output (131,072) match Alibaba Cloud Model Studio's official model-info page (context 1,000,000, max input 991,808 with 1,000,000 minus max-output overhead, max output 131,072); some coverage rounds context down to "1M tokens."
For developers
API model strings
- Alibaba Cloud Model Studio
qwen3.8-max
Licence
Proprietary — weights not released
Retirement
No retirement announced
Lineage
No recorded predecessor
News
In the news
- Aug 14, 2026
Alibaba open-sources Qwen3.8-27B, the dense checkpoint promised alongside Qwen3.8-Max
- Aug 13, 2026
Alibaba open-sources Qwen3.8-2.4T-A95B, its first Max-class model with downloadable weights
- Aug 3, 2026
Alibaba ships Qwen3.8-Max, moving its 2.4-trillion-parameter flagship out of preview
- Jul 19, 2026
Alibaba previews Qwen3.8 Max, a 2.4-trillion-parameter multimodal model