Model specification · InclusionAI

Ling 3.0 Flash

inclusionai/ling-3.0-flash

*Ling-3.0-flash* is a *124B-parameter Mixture-of-Experts (MoE) model*, with approximately *5.1B parameters activated per token*. The model is designed with *token efficiency and production-scale agentic inference* as key priorities, enabling developers...

SpecificationsReported values
Context window
262K
Maximum output
33K
Architecture
text->text
Tokenizer
Other
Knowledge cutoff
—
Moderated
No
Input modalities
text
Output modalities
text
Full pricingUSD where reported
PriceStandard
Input$0.021 / M
Output$0.063 / M
Benchmark indexesReported indexes
Intelligence
20.6
Coding
50.6
Agentic
21
Design Arena Elo
—
Example workload costsStandard pricing only
WorkloadInput / outputCache shareEstimated cost
Quick check10K / 1K × 10%$0.00027
Long analysis100K / 4K × 150%$0.00151
Workday20K / 2K × 1,00030%$0.445
More from InclusionAIAll InclusionAI models
ModelContextInput / MOutput / MIntel.
Ling 3.0 Flash VL131K$0.060$0.18025
Ling 3.0 Flash Fin262K$0.060$0.18023