LLMCost
qwen

Qwen: Qwen3 Next 80B A3B Thinking API Pricing

Qwen: Qwen3 Next 80B A3B Thinking costs $0.150 per 1M input tokens and $1.20 per 1M output tokens. Context window: 262K. Prices verified 2026-08-23 · source

Input / 1M tokens
$0.150
Output / 1M tokens
$1.20
Context window
262K
Max output
33K

What would Qwen: Qwen3 Next 80B A3B Thinking cost you per month?

WorkloadRequests/moAvg tokens (in/out)Est. monthly cost
Light chatbot10,000800 / 400$6.00
Coding agent5,0008,000 / 2,000$18.00
RAG pipeline30,0004,000 / 500$36.00
Doc summarizer2,00020,000 / 1,000$8.40

Estimates use list prices without prompt caching. Caching can cut input cost by up to 90% on repeated context.

About Qwen: Qwen3 Next 80B A3B Thinking

Qwen3-Next-80B-A3B-Thinking is a reasoning-first chat model in the Qwen3-Next line that outputs structured “thinking” traces by default. It’s designed for hard multi-step problems; math proofs, code synthesis/debugging, logic, and agentic...

Compare Qwen: Qwen3 Next 80B A3B Thinking with

Official resources