Alibaba
Pro

Qwen3.8 2.4T

qwen3.8-2.4t-a95b

Massive MoE — reasoning and code on very large context.

ReasoningCode

Technical Specifications

API Identifierqwen3.8-2.4t-a95b
PlansPro
Maximum Context256K tokens
Vision (images)No
Native Web AccessNo
Input Cost$1.80 / M tokens
Output Cost$7.20 / M tokens

Overview

Qwen3.8 2.4T is a massive mixture-of-experts (MoE) model from the Qwen family, optimized for reasoning and code on very long contexts. On Mingo, it complements the Pro pool for the most demanding tasks.


Strengths

Massive MoEVery large contextDeep reasoning

Ideal Use Cases

Very large document analysis

Contracts, reports or large codebases.

Complex reasoning

Long, precise chains of deduction.

Pricing

Input Tokens
$1.80
per million tokens
Output Tokens
$7.20
per million tokens

Frequently Asked Questions

What is a Mixture of Experts (MoE) model?
It's an architecture where only part of the parameters activate per request, giving huge capacity without the cost of an equivalent dense model.

Use Qwen3.8 2.4T on Mingo

Available on Pro.

Try Mingo for free