Runtime as featured inForbesRead the article
All open source models
Qwen (Alibaba) logo

Qwen3.8 2.4T-A95B

Qwen (Alibaba) · #7 in our ranking · Best for multilingual

Qwen3.8 2.4T-A95B is an open-weight mixture-of-experts model from Qwen (Alibaba), released Aug 2026, with 2.4T parameters (95B active per token), a 256K-token context window, and the Qwen3.8-Max License.

Qwen3.8 2.4T-A95B is the open-weight model behind Qwen3.8-Max. Alibaba reports large gains in coding, professional work, and long-horizon agent tasks over Qwen3.5 and 3.6, with tunable reasoning effort.

Qwen3.8 2.4T-A95B specs

Context256K tokens
Parameters2.4T
Provider
Qwen (Alibaba)
Architecture
Mixture-of-experts
Parameters
2.4T total · 95B active
Context
256K tokens | 197K words
License
Qwen3.8-Max License
Input
Text
Released
Aug 2026

Sourced from the model card on Hugging Face. Last checked 2026-10-01.

Highlights

  • 2.4T total, 95B active
  • Qwen-Max class, open weights
  • Tunable reasoning effort

Where it fits for payment teams

Harnesses in Runtime

Run Qwen3.8 2.4T-A95B through any of these harnesses on Runtime.

OpenCodeCline

How to run Qwen3.8 2.4T-A95B

01

Download the weights

Pull Qwen/Qwen3.8-2.4T-A95B from Hugging Face. Check the Qwen3.8-Max License before you deploy.

02

Serve it in your cloud

Run it with vLLM or SGLang on your own GPUs, or use a managed provider that hosts the model.

03

Put it to work in Runtime

Point an agent or a single skill at the model through OpenCode, then compare it to your current model with evals.

Qwen3.8 2.4T-A95B FAQ

What is the context window of Qwen3.8 2.4T-A95B?

+

Qwen3.8 2.4T-A95B supports a 256K-token context window (262,144 tokens), roughly 197K words of English text.

How many parameters does Qwen3.8 2.4T-A95B have?

+

Qwen3.8 2.4T-A95B is a mixture-of-experts model with 2.4T total parameters, of which 95B are active per token.

Is Qwen3.8 2.4T-A95B free for commercial use?

+

Qwen3.8 2.4T-A95B is released under the Qwen3.8-Max License, a custom license. Commercial use may come with conditions, so read the license on Hugging Face before you deploy or fine-tune it.

What can Qwen3.8 2.4T-A95B take as input?

+

Qwen3.8 2.4T-A95B accepts text input and generates text.

Is Qwen3.8 2.4T-A95B good for AML and compliance investigations?

+

Yes. Qwen3.8 2.4T-A95B is our #3 pick for AML and compliance investigations. Qwen-Max-class quality with broad multilingual coverage for cross-border alerts.

Is Qwen3.8 2.4T-A95B good for payments customer support?

+

Yes. Qwen3.8 2.4T-A95B is our #3 pick for payments customer support. Broad multilingual coverage for global merchant bases.

Can I run Qwen3.8 2.4T-A95B in my own cloud with Runtime?

+

Yes. Download the weights from Hugging Face (Qwen/Qwen3.8-2.4T-A95B) and serve them with vLLM or SGLang in your cloud, or use a managed provider that hosts the model. Runtime agents can then use Qwen3.8 2.4T-A95B through a harness like OpenCode, with your data staying in your environment.