Model specification · Meta
Llama 3.1 8B Instruct
meta-llama/llama-3.1-8b-instruct
Meta's latest class of model (Llama 3.1) launched with a variety of sizes & flavors. This 8B instruct-tuned version is fast and efficient. It has demonstrated strong performance compared to...
SpecificationsReported values
- Context window
- 131K
- Maximum output
- 131K
- Architecture
- text->text
- Tokenizer
- Llama3
- Knowledge cutoff
- 2023-12-31
- Moderated
- No
- Input modalities
- text
- Output modalities
- text
Full pricingUSD where reported
| Price | Standard |
|---|---|
| Input | $0.050 / M |
| Output | $0.080 / M |
Benchmark indexesReported indexes
- Intelligence
- 7.6
- Coding
- 5.4
- Agentic
- 0.5
- Design Arena Elo
- —
Example workload costsStandard pricing only
| Workload | Input / output | Cache share | Estimated cost |
|---|---|---|---|
| Quick check | 10K / 1K × 1 | 0% | $0.00058 |
| Long analysis | 100K / 4K × 1 | 50% | $0.00407 |
| Workday | 20K / 2K × 1,000 | 30% | $1.01 |
More from MetaAll Meta models
| Model | Context | Input / M | Output / M | Intel. |
|---|---|---|---|---|
| Llama 4 Maverick | 1.0M | $0.200 | $0.800 | 14.3 |
| Llama 4 Scout | 1.3M | $0.100 | $0.300 | 10 |
| Llama 3.3 70B Instruct | 131K | $0.130 | $0.400 | 9.4 |