Alibaba
Pro
Qwen3.8 2.4T
qwen3.8-2.4t-a95b
Massive MoE — reasoning and code on very large context.
Technical Specifications
| API Identifier | qwen3.8-2.4t-a95b |
|---|---|
| Plans | Pro |
| Maximum Context | 256K tokens |
| Vision (images) | No |
| Native Web Access | No |
| Input Cost | $1.80 / M tokens |
| Output Cost | $7.20 / M tokens |
Overview
Qwen3.8 2.4T is a massive mixture-of-experts (MoE) model from the Qwen family, optimized for reasoning and code on very long contexts. On Mingo, it complements the Pro pool for the most demanding tasks.
Strengths
Massive MoEVery large contextDeep reasoning
Ideal Use Cases
Very large document analysis
Contracts, reports or large codebases.
Complex reasoning
Long, precise chains of deduction.
Pricing
Input Tokens
$1.80
per million tokens
Output Tokens
$7.20
per million tokens
Related Models
Frequently Asked Questions
What is a Mixture of Experts (MoE) model?
It's an architecture where only part of the parameters activate per request, giving huge capacity without the cost of an equivalent dense model.