qwen
Qwen: Qwen3 Next 80B A3B Thinking API Pricing
Qwen: Qwen3 Next 80B A3B Thinking costs $0.150 per 1M input tokens and $1.20 per 1M output tokens. Context window: 262K. Prices verified 2026-08-23 · source
Input / 1M tokens
$0.150
Output / 1M tokens
$1.20
Context window
262K
Max output
33K
What would Qwen: Qwen3 Next 80B A3B Thinking cost you per month?
| Workload | Requests/mo | Avg tokens (in/out) | Est. monthly cost |
|---|---|---|---|
| Light chatbot | 10,000 | 800 / 400 | $6.00 |
| Coding agent | 5,000 | 8,000 / 2,000 | $18.00 |
| RAG pipeline | 30,000 | 4,000 / 500 | $36.00 |
| Doc summarizer | 2,000 | 20,000 / 1,000 | $8.40 |
Estimates use list prices without prompt caching. Caching can cut input cost by up to 90% on repeated context.
About Qwen: Qwen3 Next 80B A3B Thinking
Qwen3-Next-80B-A3B-Thinking is a reasoning-first chat model in the Qwen3-Next line that outputs structured “thinking” traces by default. It’s designed for hard multi-step problems; math proofs, code synthesis/debugging, logic, and agentic...
Compare Qwen: Qwen3 Next 80B A3B Thinking with
Official resources
- Qwen: Qwen3 Next 80B A3B Thinking on OpenRouter — live endpoint, pricing source of record.
- Get an OpenRouter API key → — one key, every model including Thinking. Free credits on signup.
- Shred your current bill — paste a CSV and see if switching saves money.