Qwen3.8-Max
Qwen3.8-Max is Alibaba Qwen's hosted 2.4-trillion-parameter mixture-of-experts flagship for coding, professional work, multimodal analysis, and long-horizon agents.
Provider
Alibaba Qwen
Model family
Alibaba Qwen
Hosted multimodal MoE flagship LLM
Cost tier
Max
Status
Current
Release Aug 3, 2026
Why teams choose it
The stable hosted model ID is qwen3.8-max; it supersedes the earlier Qwen3.8-Max-Preview catalog entry.
The stable hosted model ID is qwen3.8-max; it supersedes the earlier Qwen3.8-Max-Preview catalog entry.
QwenCloud documents a 1M context
131K maximum output, multimodal input, and built-in production features.
The hosted Max service differs from the downloadable 2.4T-A95B checkpoint in modality
thinking control, context defaults, and tools.
Tradeoffs to know
- Hosted Max behavior and features should not be assumed for the open 2.4T-A95B weights.
- A 1M context and large reasoning budget can create substantial latency and cost; test retrieval and compaction alternatives.
- Pricing, quotas, regional endpoints, and data handling depend on the selected QwenCloud or DashScope service.
When not to use this
- Not ideal for simple tasks where cheaper models in the same lineup are good enough.
- Avoid for regulated or high-stakes outputs without evaluations that mimic your tooling, data, and review process.
- Pair catalog notes with comparisons and your own benchmarks before declaring a routing winner.
Technical specs
- Inputs
- text, image, video
- Outputs
- text
- Capabilities
- coding, professional work, long-horizon agents, multimodal reasoning, function calling, web search, structured outputs, fine-tuning, context caching
- License
- Proprietary hosted API
- Model string
qwen-3-8-max
Benchmarks
{
"source": "https://www.qwencloud.com/models/qwen3.8-max",
"vendor_reported": true
}Alibaba Qwen family lineup
Current models
Compare with
Qwen3.8-Max FAQ
What is Qwen3.8-Max?
Qwen3.8-Max is Alibaba Qwen's hosted 2.4-trillion-parameter mixture-of-experts flagship for coding, professional work, multimodal analysis, and long-horizon agents. QwenCloud documents text, image, and video input, text output, a 1M-token context, up to 131K output, built-in tools, and the API model ID qwen3.8-max.
When does Qwen3.8-Max fit best?
Long-horizon software projects
What should teams watch out for with Qwen3.8-Max?
Hosted Max behavior and features should not be assumed for the open 2.4T-A95B weights.
Explore next
Models, tools, and comparisons that connect to this reference.