Qwen 3.8 Max
Alibaba's 2.4T-parameter MoE frontier model, 95B active, built for agentic coding and long-horizon tool use at $2/$6 per million tokens.
Updated 2026-08-03
Yes, within limits. Qwen 3.8 Max runs a free tier with paid plans above it.
Listed pricing: $2 / $6 per 1M tokens (input/output); open weights announced.
Overview
Is Qwen 3.8 Max free?
Yes, within limits. Qwen 3.8 Max runs a free tier with paid plans above it.
What the free tier covers: None announced for the API at launch. If the open weights ship as promised, self-hosting would be free of licence cost — but a 2.4T-parameter model carries serious compute requirements.
Listed pricing: $2 / $6 per 1M tokens (input/output); open weights announced.
Pricing on this page has not been re-verified. The entry was last edited on , and no separate pricing check has been run since. Treat the figures as a record of what was published then and confirm on the official site.
What the free tier leaves out
Read straight off the plan list below. Vendors move features between tiers, so check the current split before you pay.
- API (pay-per-token) $2 per 1M input tokens / $6 per 1M output tokens Full model access from launch day, including agentic tool use and multimodal input. No separate tiers announced at launch — check the site for current rates and any volume or cached-input discounts.
- Open weights Announced for the week following the 3 August 2026 launch Self-hosting rights per whatever licence Alibaba publishes. Licence terms, exact release date and hardware requirements were not specified in the launch announcement.
Qwen 3.8 Max pricing
| Plan | Price | What's included |
|---|---|---|
| API (pay-per-token) | $2 per 1M input tokens / $6 per 1M output tokens | Full model access from launch day, including agentic tool use and multimodal input. No separate tiers announced at launch — check the site for current rates and any volume or cached-input discounts. |
| Open weights | Announced for the week following the 3 August 2026 launch | Self-hosting rights per whatever licence Alibaba publishes. Licence terms, exact release date and hardware requirements were not specified in the launch announcement. |
Full model access from launch day, including agentic tool use and multimodal input. No separate tiers announced at launch — check the site for current rates and any volume or cached-input discounts.
Self-hosting rights per whatever licence Alibaba publishes. Licence terms, exact release date and hardware requirements were not specified in the launch announcement.
Pricing on this page has not been re-verified. The entry was last edited on , and no separate pricing check has been run since. Treat the figures as a record of what was published then and confirm on the official site.
Is Qwen 3.8 Max worth it?
You can test that on the free tier before paying anything. The recorded trade-offs are listed below, and any one of them can settle the question on its own.
The 8.7/10 AI Score is an editorial read of published capability, price and shipping pace. Nobody here has hands-on hours with Qwen 3.8 Max. How we verify.
Worth it if
The strengths recorded against this entry.
- $2/$6 per million tokens undercuts closed frontier list pricing substantially, which compounds in token-hungry agentic loops
- API was live on announcement day rather than gated behind a waitlist
- Open weights announced for the following week — unprecedented at this parameter scale if it lands
- Native multimodal input in the same checkpoint, no separate vision model to route to
- Explicitly tuned for terminal and repository work rather than retrofitted for it
Not worth it if
Any one of these blocks your use case.
- TerminalBench 86.6 and the rest of the benchmark table are vendor-reported with no independent replication at time of writing
- Open weights are a dated promise, not a shipped artefact — judge it when the download exists
- 2.4T total parameters puts practical self-hosting out of reach for individuals and small teams even once weights are public
- China-hosted inference raises data-residency questions that regulated buyers will need to resolve before adoption
Key features
Sparse MoE architecture
2.4 trillion total parameters with about 95 billion active per token. The sparsity ratio is what makes the low per-token price plausible — you pay for a fraction of the network on each forward pass.
Agentic coding focus
Alibaba leads the announcement with TerminalBench (86.6, vendor-reported) rather than a general knowledge benchmark, signalling that the model is tuned for multi-step terminal and repository work rather than single-turn answers.
Long-horizon tool use
Built to hold state across extended tool-calling chains, which is the failure mode that separates usable coding agents from demos. Independent evidence on how far it holds up is not available yet.
Open weights commitment
Alibaba announced a weights release for the week following the 3 August 2026 API launch. If it ships, this is the only model at this parameter scale with published weights.
How it compares
| Tool | Best for | Pricing | Score |
|---|---|---|---|
| Qwen 3.8 Max | — | $2 / $6 per 1M tokens (input/output); open weights announced | 8.7/10 |
| ChatGPT vs ChatGPT → | Users who want one subscription covering reasoning, voice, vision, image and video generation, and agentic browsing in a single app. | Free tier + Plus $20/mo + Pro $200/mo | 9.5/10 |
| Claude vs Claude → | Developers and professionals who need agentic coding, computer control, and large-document or codebase analysis in one assistant. | Free tier + Pro $20/mo + Team $30/mo/user | 9.5/10 |
| Gemini vs Gemini → | Google Workspace users who want an assistant that can read their Gmail, Drive, and Calendar while reasoning across huge documents or videos in one pass. | Free tier + Advanced $19.99/mo | 9.2/10 |
Compare head-to-head
Related reading
Qwen3.8-Max Pricing Lands, the Benchmarks Don't
Alibaba priced Qwen3.8-Max at $2/$6 per million tokens on August 3, two weeks after selling a preview on a ranking it never showed.
Qwen3.8-Max vs GPT-5.6 Sol: Terminal-Bench and Price
Qwen3.8-Max posts Terminal-Bench 2.1 of 86.6 against GPT-5.6 Sol's 88.8 at $2/$6 per million tokens. A four-step way to decide whether to switch.
Claude Eval Escapes: An 8-Day Disclosure Clock
Anthropic suspended the cyber evals July 23, notified affected organizations July 27, published July 31. What those dates settle and what they don't.
Ready to try Qwen 3.8 Max?
Head to the official site to start with Qwen 3.8 Max — pricing and plans are listed above.
Visit Qwen 3.8 Max

