Model specification · Qwen

Qwen3.8 2.4T A95B

qwen/qwen3.8-2.4t-a95b

Qwen3.8 2.4T A95B is an open-weight sparse mixture-of-experts model from Qwen and the open-weight variant of [Qwen3.8 Max](/qwen/qwen3.8-max), with 95 billion active parameters out of 2.4 trillion total. It is...

SpecificationsReported values
Context window
1.0M
Maximum output
131K
Architecture
text->text
Tokenizer
Qwen
Knowledge cutoff
Moderated
No
Input modalities
text
Output modalities
text
Full pricingUSD where reported
PriceStandard
Input$2.00 / M
Output$6.00 / M
Benchmark indexesReported indexes
Intelligence
40
Coding
71.9
Agentic
50.4
Design Arena Elo
Example workload costsStandard pricing only
WorkloadInput / outputCache shareEstimated cost
Quick check10K / 1K × 10%$0.026
Long analysis100K / 4K × 150%$0.137
Workday20K / 2K × 1,00030%$41.50
More from QwenAll Qwen models
ModelContextInput / MOutput / MIntel.
Qwen3.8 Max (0902)1M$2.00$6.0045.4
Qwen3.8 27B1M$0.214$2.5533.9
Qwen3.7 Max1M$1.48$4.4229.9
Qwen3.7 Plus1M$0.320$1.2825.8
Qwen3.6 27B262K$0.300$2.0021.9
Qwen3.5 397B A17B262K$0.550$3.5019.1
Qwen3.6 35B A3B262K$0.100$0.90018.8
Qwen3.5-122B-A10B262K$0.260$2.0816.2
Qwen3 235B A22B Thinking 2507131K$0.230$2.3012.7
Qwen3 Coder Next262K$0.120$0.80010.1
Qwen3 30B A3B Thinking 250782K$0.200$2.409.8
Qwen3 32B131K$0.080$0.2807.2