Model specification · InclusionAI
Ling-2.6-flash
inclusionai/ling-2.6-flash
Ling-2.6-flash is an instant (instruct) model from inclusionAI with 104B total parameters and 7.4B active parameters, designed for real-world agents that require fast responses, strong execution, and high token efficiency....
SpecificationsReported values
- Context window
- 262K
- Maximum output
- 33K
- Architecture
- text->text
- Tokenizer
- Other
- Knowledge cutoff
- —
- Moderated
- No
- Input modalities
- text
- Output modalities
- text
Full pricingUSD where reported
| Price | Standard |
|---|---|
| Input | $0.010 / M |
| Output | $0.030 / M |
Benchmark indexesReported indexes
- Intelligence
- 14.1
- Coding
- 25.3
- Agentic
- 2.3
- Design Arena Elo
- —
Example workload costsStandard pricing only
| Workload | Input / output | Cache share | Estimated cost |
|---|---|---|---|
| Quick check | 10K / 1K × 1 | 0% | $0.00013 |
| Long analysis | 100K / 4K × 1 | 50% | $0.00072 |
| Workday | 20K / 2K × 1,000 | 30% | $0.212 |
More from InclusionAIAll InclusionAI models
| Model | Context | Input / M | Output / M | Intel. |
|---|---|---|---|---|
| Ring-2.6-1T | 262K | $0.075 | $0.625 | 30.6 |