InclusionAI API Pricing & Model Comparison
We track 1 InclusionAI models, with input pricing from $0.07 to $0.07 per million tokens. The line-up's blended rate runs 96% below the catalog average. Below: price, context, speed and value compared across the range.
For cost, InclusionAI's cheapest option is Ling-3.0-flash ($0.07 input / $0.22 output); the quality ceiling is Ling-3.0-flash at a quality score of 80. On a 50,000-turn-per-month support workload that is $13 versus $13 — the spread inside a single provider is often wider than the spread between providers.
On value (quality score ÷ output price) the pick of the InclusionAI range is Ling-3.0-flash. The largest context window is 262K, offered only by Ling-3.0-flash, and the fastest output is Ling-3.0-flash at 180 tok/s. Modalities across the line-up: Text.
Models and pricing
Sorted by blended rate, cheapest first (3:1 input-to-output mix). Follow a model name for full specs and monthly cost estimates.
| Model | Input /1M | Output /1M | Blended /1M | Quality | Context | Speed | Value | Use cases |
|---|---|---|---|---|---|---|---|---|
| Ling-3.0-flash | $0.07 | $0.22 | $0.11 | 80 | 262K | 180 tok/s | 363.6 | ChatbotAffordableFastLong Context |
Frequently asked questions
Which InclusionAI model is cheapest?
Ling-3.0-flash, at $0.07 input and $0.22 output per million tokens, with a quality score of 80.
What is InclusionAI's flagship model?
By quality score it is Ling-3.0-flash (80), priced at $0.07 input / $0.22 output with a 262K context window.
Does InclusionAI pricing change?
Yes. llmprice.app runs a daily collection job and refreshes these figures, but there can be a lag after a provider re-prices — confirm against InclusionAI's official pricing page before committing.
Prices and specs are collected daily by llmprice.app. Confirm against the official pricing page before production use.