Q

Qwen3.8-Max

Alibaba · Released Jul 2026

pending

Qwen3.8-Max is Alibaba's 2.4-trillion-parameter flagship, previewed July 2026 as a multimodal MoE model with a 984K context window. It is not yet GA — no per-token API pricing, no benchmark table, no open weights, and no model card have been published, and the "second only to Fable 5" claim is Alibaba's own unverified self-assessment. The preview is worth testing for coding and long-context agent workloads at $6/month entry cost, but wait for published benchmarks, per-token pricing, and the promised open-weight release before committing.

Is it right for you?

Good for

  • Massive 984K-token context window for long-document and agent workloads
  • Native multimodal input — text, images, video, and documents in one model
  • OpenAI and Anthropic API-compatible — works with Cursor, Claude Code, OpenCode, and Codex
  • Low preview entry cost ($6/month Token Plan Lite) for early evaluation

Not good for

  • Production deployments — no stable model ID, no per-token pricing, no benchmarks, no SLA
  • Self-hosting — 2.4T parameters requires datacenter-scale infrastructure; open weights not released
  • Regulated environments requiring auditable model cards and verified benchmark data

Pricing

Input

Credit-based only

Output

Credit-based only

Context

984K tokens

View full pricing

No verdict changes yet

The clock starts day one — changes land here as our verdict evolves.

Verification log

  • Pricing— No changes

    Automated agent

How we evaluate