Model specification · InclusionAI

Ling-2.6-flash

inclusionai/ling-2.6-flash

Ling-2.6-flash is an instant (instruct) model from inclusionAI with 104B total parameters and 7.4B active parameters, designed for real-world agents that require fast responses, strong execution, and high token efficiency....

SpecificationsReported values
Context window
262K
Maximum output
33K
Architecture
text->text
Tokenizer
Other
Knowledge cutoff
Moderated
No
Input modalities
text
Output modalities
text
Full pricingUSD where reported
PriceStandard
Input$0.010 / M
Output$0.030 / M
Benchmark indexesReported indexes
Intelligence
14.1
Coding
25.3
Agentic
2.3
Design Arena Elo
Example workload costsStandard pricing only
WorkloadInput / outputCache shareEstimated cost
Quick check10K / 1K × 10%$0.00013
Long analysis100K / 4K × 150%$0.00072
Workday20K / 2K × 1,00030%$0.212
More from InclusionAIAll InclusionAI models
ModelContextInput / MOutput / MIntel.
Ring-2.6-1T262K$0.075$0.62530.6