LLMCost
qwen

Qwen: Qwen3.8 2.4T A95B API Pricing

Qwen: Qwen3.8 2.4T A95B costs $2.00 per 1M input tokens and $6.00 per 1M output tokens (cache reads: $0.250). Context window: 1.0M. Prices verified 2026-08-23 · source

Input / 1M tokens
$2.00
Output / 1M tokens
$6.00
Context window
1.0M
Max output
131K

What would Qwen: Qwen3.8 2.4T A95B cost you per month?

WorkloadRequests/moAvg tokens (in/out)Est. monthly cost
Light chatbot10,000800 / 400$40.00
Coding agent5,0008,000 / 2,000$140.00
RAG pipeline30,0004,000 / 500$330.00
Doc summarizer2,00020,000 / 1,000$92.00

Estimates use list prices without prompt caching. Caching can cut input cost by up to 90% on repeated context.

About Qwen: Qwen3.8 2.4T A95B

Qwen3.8 2.4T A95B is an open-weight sparse mixture-of-experts model from Qwen and the open-weight variant of [Qwen3.8 Max](/qwen/qwen3.8-max), with 95 billion active parameters out of 2.4 trillion total. It is...

Compare Qwen: Qwen3.8 2.4T A95B with