Model specification · Qwen

Qwen3 8B

qwen/qwen3-8b

Qwen3-8B is a dense 8.2B parameter causal language model from the Qwen3 series, designed for both reasoning-heavy tasks and efficient dialogue. It supports seamless switching between "thinking" mode for math,...

SpecificationsReported values
Context window
131K
Maximum output
8.2K
Architecture
text->text
Tokenizer
Qwen3
Knowledge cutoff
2025-03-31
Moderated
No
Input modalities
text
Output modalities
text
Full pricingUSD where reported
PriceStandard
Input$0.117 / M
Output$0.455 / M
Benchmark indexesReported indexes
Intelligence
8.3
Coding
9
Agentic
1.5
Design Arena Elo
Example workload costsStandard pricing only
WorkloadInput / outputCache shareEstimated cost
Quick check10K / 1K × 10%$0.00162
Long analysis100K / 4K × 150%$0.014
Workday20K / 2K × 1,00030%$3.25
More from QwenAll Qwen models
ModelContextInput / MOutput / MIntel.
Qwen3.7 Max1M$1.48$4.4246
Qwen3.6 Plus1M$0.325$1.9539.6
Qwen3.7 Plus1M$0.320$1.2839
Qwen3.6 27B262K$0.289$2.4037.1
Qwen3.5 397B A17B262K$0.390$2.3433.7
Qwen3.5-122B-A10B262K$0.260$2.0832.3
Qwen3.6 35B A3B262K$0.140$1.0031.6
Qwen3.5-35B-A3B262K$0.140$1.0024
Qwen3.5-9B262K$0.100$0.15021.4
Qwen3 Coder Next262K$0.120$0.80021.1
Qwen3 235B A22B Thinking 2507262K$0.230$2.3019.6
Qwen3 Next 80B A3B Thinking262K$0.150$1.2016.7