Model specification · Google
Gemma 4 31B
google/gemma-4-31b-it
Gemma 4 31B Instruct is Google DeepMind's 30.7B dense multimodal model supporting text and image input with text output. Features a 256K token context window, configurable thinking/reasoning mode, native function...
SpecificationsReported values
- Context window
- 262K
- Maximum output
- 16K
- Architecture
- text+image+video->text
- Tokenizer
- Gemma
- Knowledge cutoff
- —
- Moderated
- No
- Input modalities
- image text video
- Output modalities
- text
Full pricingUSD where reported
| Price | Standard | free |
|---|---|---|
| Input | $0.090 / M | free / M |
| Output | $0.340 / M | free / M |
Benchmark indexesReported indexes
- Intelligence
- 15.4
- Coding
- 43.4
- Agentic
- 6.7
- Design Arena Elo
- —
Example workload costsStandard pricing only
| Workload | Input / output | Cache share | Estimated cost |
|---|---|---|---|
| Quick check | 10K / 1K × 1 | 0% | $0.00124 |
| Long analysis | 100K / 4K × 1 | 50% | $0.00836 |
| Workday | 20K / 2K × 1,000 | 30% | $2.24 |
More from GoogleAll Google models
| Model | Context | Input / M | Output / M | Intel. |
|---|---|---|---|---|
| Gemini 3.8 Flash | 1.0M | $0.750 | $3.75 | 41.2 |
| Gemini 3.7 Flash | 1.0M | $0.750 | $3.75 | 39.4 |
| Gemini 3.6 Flash | 1.0M | $0.750 | $3.75 | 34.3 |
| Gemini 3.5 Flash | 1.0M | $1.50 | $9.00 | 33 |
| Gemini 3.1 Pro Preview | 1.0M | $2.00 | $12.00 | 30.4 |
| Gemini 3.5 Flash Lite | 1.0M | $0.300 | $2.50 | 22.7 |
| Gemini 2.5 Pro | 1.0M | $1.25 | $10.00 | 16.7 |
| Gemini 3.1 Flash Lite Preview | 1.0M | $0.250 | $1.50 | 16 |
| Gemma 3 27B | 131K | $0.080 | $0.450 | 4.9 |
| Gemma 3 12B | 131K | $0.050 | $0.150 | 3.8 |