qwen
Qwen: Qwen3 VL 8B Thinking API Pricing
Qwen: Qwen3 VL 8B Thinking costs $0.180 per 1M input tokens and $2.10 per 1M output tokens. Context window: 131K. Prices verified 2026-08-23 Β· source
Input / 1M tokens
$0.180
Output / 1M tokens
$2.10
Context window
131K
Max output
33K
What would Qwen: Qwen3 VL 8B Thinking cost you per month?
| Workload | Requests/mo | Avg tokens (in/out) | Est. monthly cost |
|---|---|---|---|
| Light chatbot | 10,000 | 800 / 400 | $9.84 |
| Coding agent | 5,000 | 8,000 / 2,000 | $28.20 |
| RAG pipeline | 30,000 | 4,000 / 500 | $53.10 |
| Doc summarizer | 2,000 | 20,000 / 1,000 | $11.40 |
Estimates use list prices without prompt caching. Caching can cut input cost by up to 90% on repeated context.
About Qwen: Qwen3 VL 8B Thinking
Qwen3-VL-8B-Thinking is the reasoning-optimized variant of the Qwen3-VL-8B multimodal model, designed for advanced visual and textual reasoning across complex scenes, documents, and temporal sequences. It integrates enhanced multimodal alignment and...
Compare Qwen: Qwen3 VL 8B Thinking with
Official resources
- Qwen: Qwen3 VL 8B Thinking on OpenRouter β live endpoint, pricing source of record.
- Get an OpenRouter API key β β one key, every model including Thinking. Free credits on signup.
- Shred your current bill β paste a CSV and see if switching saves money.