Monday, August 3, 2026
HomeArtificial IntelligenceAlibaba Qwen Releases Qwen3.8-Max: A 2.4 Trillion Parameter MoE Mannequin and the...

Alibaba Qwen Releases Qwen3.8-Max: A 2.4 Trillion Parameter MoE Mannequin and the Most Succesful One within the Qwen Household to Date

Alibaba’s Qwen group has made Qwen3.8-Max broadly accessible and confirmed that its open weights ship subsequent week. A second checkpoint, Qwen3.8-27B, can be going open-weights. Qwen3.8-Max is a 2.4-trillion-parameter mixture-of-experts mannequin. It accepts textual content, picture and video as enter and returns textual content.

Is it deployable

Sure, however the deployable floor depends upon which artifact you might be making use of.

The hosted API is deployable in the present day by any firm dimension. It’s OpenAI- and DashScope-compatible, so integration is a base-URL and model-ID change. The open weights are a distinct matter. At 2.4T complete parameters, the checkpoint is a multi-node datacenter artifact. Alibaba has not disclosed the activated-parameter depend. Serving value subsequently can not but be modeled. Qwen3.8-27B is the checkpoint that matches atypical on-premise GPU {hardware}.

The printed characteristic set maps cleanly onto 4 industries. These are software program engineering, authorized and monetary doc overview, media and e-commerce operations, and design.

Functions embrace repository-scale coding brokers and long-document information bases. Lengthy-video indexing, structured information extraction and multi-step analysis assistants additionally match.

Interactive explainer

What’s Technically Obtainable

The mannequin web page lists a 1M-token context window. Most enter is 991K tokens, dropping to 983K when considering is enabled. Most output is 131K tokens in each modes, and the utmost reasoning finances is 262K tokens. Charge limits are 2M tokens per minute and 15K requests per minute.

Pricing is $2.00 per 1M enter tokens and $6.00 per 1M output tokens. Implicit cache reads value $0.25 per 1M tokens. Specific cache creation is $2.50 and specific cache reads are $0.17 per 1M tokens. Cached enter is eight instances cheaper than contemporary enter. Prefix stability subsequently drives value greater than immediate size does.

Supported capabilities embrace perform calling, structured outputs, batches, prefix completion and fine-tuning. 5 built-in instruments ship on the Responses API: code_interpreter, web_search, web_extractor, t2i_search and i2i_search.

https://qwen.ai/weblog?id=qwen3.8

Efficiency

Alibaba printed a full benchmark desk with this launch. Qwen3.8-Max scores 86.6 on Terminal-Bench 2.1, forward of Claude Opus 4.8 and Claude Fable 5 at 84.6, behind GPT-5.6 Sol (max) at 88.8. It stories 67.7 on SWE-bench Professional towards Fable 5’s 80.0, and 73.5 on FrontierSWE towards Fable 5’s 88.8. It leads PaperBench at 93.0 and IFBench at 82.8. GPQA Diamond lands at 92.6, up marginally from Qwen3.7-Max’s 92.4. The clearest positive factors are multimodal and agentic, not reasoning. It tops most imaginative and prescient rows, together with OSWorld-Verified 86.1, Parametric CAD Bench 91.5, and OmniDocBench 1.5 at 92.1. Towards its personal predecessor the leap is giant: DeepSWE 1.1 strikes from 21.6 to 56.6, FrontierSWE from 40.7 to 73.5, JobBench from 31.3 to 53.4. Two caveats belong in any sincere learn. The multimodal desk benchmarks towards Qwen3.7-Plus, not Qwen3.7-Max, which flatters the generational delta. And Alibaba’s personal RL scaling curve peaks at 0.725 close to 4,000 coaching environments, then declines to 0.719 and 0.689.

Key Takeaways

  • Qwen3.8-Max is a 2.4T-parameter MoE mannequin with 1M context, now typically accessible.
  • Pricing is $2 enter, $6 output and $0.25 cached enter per 1M tokens.
  • Open weights for Qwen3.8-Max and Qwen3.8-27B are promised subsequent week.
  • No benchmark desk, license, or activated-parameter depend has been printed.
  • The 27B checkpoint, not the flagship, is the practical on-premise deployment path.

Take a look at the Technical particulars, API and Qwen StudioAdditionally, be happy to comply with us on Twitter and don’t neglect to affix our 150k+ML SubReddit and Subscribe to our Publication. Wait! are you on telegram? now you’ll be able to be part of us on telegram as effectively.

Must companion with us for selling your GitHub Repo OR Hugging Face Web page OR Product Launch OR Webinar and many others.? Join with us


Asif Razzaq is the CEO of Marktechpost Media Inc.. As a visionary entrepreneur and engineer, Asif is dedicated to harnessing the potential of Synthetic Intelligence for social good. His most up-to-date endeavor is the launch of an Synthetic Intelligence Media Platform, Marktechpost, which stands out for its in-depth protection of machine studying and deep studying information that’s each technically sound and simply comprehensible by a large viewers. The platform boasts of over 2 million month-to-month views, illustrating its recognition amongst audiences.

RELATED ARTICLES

LEAVE A REPLY

Please enter your comment!
Please enter your name here

Most Popular

Recent Comments