inclusionai
inclusionAI: Ling-2.6-flash API Pricing
inclusionAI: Ling-2.6-flash costs $0.010 per 1M input tokens and $0.030 per 1M output tokens (cache reads: $0.0020). Context window: 262K. Prices verified 2026-08-23 · source
Input / 1M tokens
$0.010
Output / 1M tokens
$0.030
Context window
262K
Max output
33K
What would inclusionAI: Ling-2.6-flash cost you per month?
| Workload | Requests/mo | Avg tokens (in/out) | Est. monthly cost |
|---|---|---|---|
| Light chatbot | 10,000 | 800 / 400 | $0.20 |
| Coding agent | 5,000 | 8,000 / 2,000 | $0.70 |
| RAG pipeline | 30,000 | 4,000 / 500 | $1.65 |
| Doc summarizer | 2,000 | 20,000 / 1,000 | $0.46 |
Estimates use list prices without prompt caching. Caching can cut input cost by up to 90% on repeated context.
About inclusionAI: Ling-2.6-flash
Ling-2.6-flash is an instant (instruct) model from inclusionAI with 104B total parameters and 7.4B active parameters, designed for real-world agents that require fast responses, strong execution, and high token efficiency....