LLMCost
inclusionai

Ling-3.0-flash API Pricing

Ling-3.0-flash costs $0.021 per 1M input tokens and $0.063 per 1M output tokens (cache reads: $0.0042). Context window: 262K. Prices verified 2026-08-23 · source

Input / 1M tokens
$0.021
Output / 1M tokens
$0.063
Context window
262K
Max output
33K

What would Ling-3.0-flash cost you per month?

WorkloadRequests/moAvg tokens (in/out)Est. monthly cost
Light chatbot10,000800 / 400$0.42
Coding agent5,0008,000 / 2,000$1.47
RAG pipeline30,0004,000 / 500$3.46
Doc summarizer2,00020,000 / 1,000$0.97

Estimates use list prices without prompt caching. Caching can cut input cost by up to 90% on repeated context.

About Ling-3.0-flash

*Ling-3.0-flash* is a *124B-parameter Mixture-of-Experts (MoE) model*, with approximately *5.1B parameters activated per token*. The model is designed with *token efficiency and production-scale agentic inference* as key priorities, enabling developers...

Compare Ling-3.0-flash with