Qwen3.8 Max Preview: Release, Access, and Open Weights
Qwen3.8 Max Preview is live through the Alibaba Token Plan and Qoder. Here is what is confirmed about access, benchmarks, open weights, and the 2.4T claim.
By Capital & Compute
Qwen3.8 Max Preview is real and usable, but it is not yet a fully documented model release. Alibaba added Qwen3.8-Max-Preview to its AI Token Plan, and the Qwen team says the preview is also available in Qoder and QoderWork. What Alibaba has not published matters just as much: there is no benchmark table, model card, stable standalone API price, context-window specification, license, or downloadable weight file as of July 20, 2026.
Alibaba calls the 2.4-trillion-parameter preview its strongest Qwen and says it is “second only to Claude Fable 5.” Both are Alibaba claims, not independent findings. The honest launch-day answer is therefore narrower than the headline: Qwen3.8 is a testable preview with an unusually large claimed architecture and an open-weight release promised for later, not yet a measured replacement for Claude, Kimi, or Qwen3.7 Max.
Is Qwen3.8 out yet?
Yes, as a preview inside specific Alibaba products. No, if “out” means a normal model launch with a stable API page, rate card, technical documentation, model card, and weights that anyone can download.
The cleanest primary evidence is Alibaba Cloud’s live Token Plan page. It names Qwen3.8-Max-Preview in both the page introduction and the supported-model list, alongside Qwen3.7 Max, GLM-5.2, and DeepSeek V4 Pro. That is stronger evidence than a leaked model identifier or an anonymous pre-release claim: Alibaba is selling access to a product that names the model.
The rest of the access picture comes from Alibaba’s official Qwen announcement on X, which says the preview debuted in Token Plan, Qoder, and QoderWork. The post is independently described by both the South China Morning Post and SiliconANGLE. Neither report points to a normal Model Studio API documentation page for Qwen3.8, and Alibaba’s public model-pricing documentation still stops at the earlier Qwen line. So access is confirmed; general API availability and per-token pricing are not.
This distinction matters because “available” gets used for three different events in AI launches:
- A model appears in a first-party app or subscription.
- A stable API model ID and rate card let developers build against it.
- A weight repository and license let anyone run it on their own infrastructure.
Qwen3.8 has reached the first stage. Alibaba says the third is coming. The second is not documented clearly enough to budget a production workload around it.
What is actually confirmed about Qwen3.8?
The launch has one unusually solid fact—product availability—and a large ring of claims waiting for documentation. Here is the evidence state as of July 20.
| Detail | Current evidence | What it means |
|---|---|---|
| Model name | Qwen3.8-Max-Preview on Alibaba’s Token Plan page |
Confirmed first-party preview |
| Launch timing | Announced July 19 during the World Artificial Intelligence Conference | Confirmed by two launch reports citing Alibaba |
| Access | Token Plan confirmed by Alibaba; Qoder and QoderWork named in its announcement | Testable in selected Alibaba products |
| Total parameters | 2.4 trillion | Alibaba-reported; no model card yet |
| Active parameters | Not published | Inference compute cannot be estimated |
| Context window | Not published | Do not assume Qwen3.7’s 1M window carries over |
| Modalities | Multimodal capability reported by launch coverage | Awaiting first-party model documentation |
| Benchmark scores | None published | “Second only” cannot be reproduced |
| API price | No standalone rate card found | Token Plan price is not a per-token API price |
| Open weights | Promised “soon” | No repository, files, date, or license yet |
Calling this an evidence gap is not pedantry. A mixture-of-experts model can have trillions of total parameters while activating only a small fraction for each token. Total size helps describe storage and training scale; the active count is closer to the compute required at inference. Without both, “2.4T” says very little about speed, serving cost, or the hardware needed to self-host it.
The same caution applies to context length and modalities. Qwen3.7 Max has a documented 1-million-token context window, and Alibaba has shipped vision-language Qwen models before. That history makes a long-context multimodal Qwen3.8 plausible. It does not make those specifications confirmed. Product generations frequently change limits between preview and general availability.
Is Qwen3.8 really second only to Claude Fable 5?
There is no public evidence strong enough to answer yes. Alibaba says Qwen3.8 is comparable to leading frontier systems and “second only to Fable 5,” but it did not publish the leaderboard, benchmark suite, scores, harness, competitor settings, or number of runs behind that ranking.
That makes the sentence positioning, not measurement. It may later prove directionally correct, but readers cannot reproduce it today. A benchmark claim becomes useful only when it names the test, discloses the conditions, and gives rivals the same opportunity to be measured.
The omission is especially noticeable because the prior model arrived with far more evidence. Qwen3.7 Max shipped with named coding and agent evaluations, and it now has independent results tracked in our Qwen3.7 Max versus Claude cost analysis and AI model value leaderboard. Those numbers carry their own caveats, but at least they can be inspected and challenged. Qwen3.8 currently offers only the conclusion.
The first useful independent signals will likely come from evaluations that already cover the current frontier: Artificial Analysis for broad capability and cost, Arena for human preference, and agentic coding tests such as Terminal-Bench or SWE-bench. Until one of those publishes a Qwen3.8 result, the current model tracker should not replace measured models with a launch claim.
How to access Qwen3.8 Max Preview
The verified route is Alibaba Cloud’s AI Token Plan. The individual plan currently has three tiers: Lite at a promotional $6 per month with about 10,000 credits, Standard at $18 with about 40,000 credits, and Pro at $68 with about 160,000 credits. Alibaba lists Qwen3.8 Max Preview among the models available in the plan.
| Item | Value |
|---|---|
| Lite · $6/month | 10K credits |
| Standard · $18/month | 40K credits |
| Pro · $68/month | 160K credits |
The credit system prevents a clean cost-per-million-token calculation. Alibaba does not state on that page how many Qwen3.8 input or output tokens one credit buys, and the same balance also covers image, video, and audio tools. Treat the plan price as the cost of preview access, not the model’s API rate.
Qoder and QoderWork are the other announced surfaces. They are agentic development products, so this route is most relevant to people evaluating Qwen3.8 for code generation and software tasks. The announcement does not establish that every Qoder plan, region, or account has identical availability. Check the model picker before buying access solely for Qwen3.8.
SiliconANGLE reports that preview usage is priced at 10% of the standard rate during the trial period. Alibaba’s public Token Plan page does not expose a Qwen3.8 standard rate from which to reproduce that discount, so the safer budget is the visible subscription price. A percentage with no published denominator is not a rate card.
Qwen3.8 versus Qwen3.7 Max and Kimi K3
The nearest comparisons show why Qwen3.8 is interesting and why it is too early to rank.
| Model | Release state | Claimed total size | Public API price | Public benchmarks | Open weights |
|---|---|---|---|---|---|
| Qwen3.8 Max Preview | Selected-product preview | 2.4T | Not published | None yet | Promised, no date |
| Qwen3.7 Max | Stable API model | More than 1T | $2.50 input / $7.50 output per Mtok | Vendor and independent results | No |
| Kimi K3 | API release | About 2.8T | $3 input / $15 output per Mtok | Vendor and independent results | Announced for July 27 |
Qwen3.7 Max is the practical baseline. It has documented rates, a stable endpoint, a 1M-token context window, and enough evaluation data to estimate the trade between capability and cost. If you need to deploy a Qwen model this week, those boring details make Qwen3.7 the safer choice even if Qwen3.8 eventually proves much stronger.
Kimi K3 is the scale comparison. Moonshot’s new flagship is reported at about 2.8 trillion parameters and arrived with API pricing, a specification table, vendor benchmarks, and several independent reads. Our Kimi K3 pricing and benchmark explainer separates those measurements from Moonshot’s claims. The contrast is useful: both launches entered the multi-trillion-parameter tier, but Kimi currently gives buyers much more information.
None of this means Qwen3.8 is weaker. It means the evidence is weaker. Ranking models by the completeness of their launch material would be silly; choosing a production dependency based on missing material would be worse.
Are Qwen3.8 open weights available?
No. Alibaba says they are coming soon, but there is no official Qwen3.8 model repository on the Qwen Hugging Face organization, no downloadable checkpoint, no model card, no file sizes, and no license as of July 20.
“Open weights soon” is a commitment about a future release, not the current preview’s license. The distinction determines what developers can actually do:
- A hosted preview lets you send requests through Alibaba’s products under their service terms.
- An open-weight release lets you download the parameters and run the model elsewhere, subject to its license.
- Open source would additionally describe the training code and broader development artifacts, which model launches rarely provide in full.
Qwen has a strong history of publishing downloadable models, so the promise is credible. The exact terms still matter. A permissive Apache 2.0 release would support broad commercial use; a custom community license could add attribution, scale, or use restrictions. Our guide to the best open-weight models treats a model as available only when the files and license exist, and Qwen3.8 has not crossed that line yet.
Should you use Qwen3.8 now?
Use the preview if your goal is evaluation. A short trial can answer questions the announcement cannot: whether the model follows a coding harness reliably, how often it loops, whether its visual understanding works on your documents, and how it compares with Qwen3.7 on the tasks you actually run.
Do not migrate a production system yet. The preview label means behavior can change, and the missing rate card makes unit economics impossible to forecast. There is also no public documentation for context limits, throughput, regional routing, retention, or deprecation policy specific to Qwen3.8. Those are operational requirements, not paperwork.
A sensible evaluation looks like this:
- Run a fixed, non-sensitive task set through Qwen3.8 and the model you use today.
- Score completed work, not how impressive the first answer sounds.
- Record latency, retries, failures, and credits consumed.
- Keep the preview out of automated production routing.
- Repeat the test after Alibaba publishes the final model, because preview behavior may not carry over.
For coding teams, include repository-scale tasks and review the resulting diff rather than relying on a chat impression. The wider AI coding-agent landscape explains why the harness and tool loop can change results as much as the model itself.
Bottom line
Qwen3.8 Max Preview is a real Alibaba preview with confirmed Token Plan access, a claimed 2.4-trillion-parameter architecture, and a promise of open weights. It is not yet a documented frontier-model release. Alibaba’s “second only to Claude Fable 5” line has no public benchmark table behind it, and the missing API price, active parameter count, context window, model card, and license prevent a serious deployment decision.
The right move is to test it without treating it as settled. Qwen3.8 becomes a production candidate when Alibaba publishes a stable endpoint and rate card. It becomes an open-weight model when the repository and license land. And it becomes a proven frontier contender when independent results support the launch claim. Until then, the most important word in its name is not Max. It is Preview.
Frequently asked questions
- Is Qwen3.8 released?
- Qwen3.8 Max Preview is available through Alibaba Cloud Token Plan and, according to Alibaba, through Qoder and QoderWork. A fully documented general API and open-weight release have not arrived yet.
- Is Qwen3.8 open source?
- Not yet. Alibaba says open weights are coming soon, but there is no official Qwen3.8 repository, downloadable checkpoint, release date, model card, or license as of July 20, 2026.
- How can I access Qwen3.8 Max Preview?
- The verified route is Alibaba Cloud's AI Token Plan, which lists Qwen3.8 Max Preview as a supported model. Alibaba also names its Qoder and QoderWork agentic-development products as preview surfaces.
- Does Qwen3.8 have 2.4 trillion parameters?
- Alibaba says Qwen3.8 has 2.4 trillion total parameters. No technical report or model card has independently confirmed that figure or disclosed how many parameters are active for each token.
- Is Qwen3.8 better than Claude Fable 5?
- There is not enough public evidence to know. Alibaba calls Qwen3.8 second only to Claude Fable 5, but it has published no benchmark scores or methodology and no independent evaluator has ranked the preview yet.
Sources and verification
- Alibaba Cloud. AI Token Plan. Preview availability, supported-model list, subscription prices, credits, and concurrency. Verified July 20, 2026.
- Qwen Team. Qwen3.8 Max Preview announcement. Availability, parameter count, positioning, and open-weight intent. Published July 19, 2026.
- South China Morning Post. Alibaba says newest Qwen AI model is second only to Anthropic’s Claude Fable 5. Published July 19, 2026.
- SiliconANGLE. Alibaba previews Qwen3.8, claims it is second only to Claude Fable 5. Published July 19, 2026.