LLMCost
qwen

Qwen: Qwen3 VL 8B Thinking API Pricing

Qwen: Qwen3 VL 8B Thinking costs $0.180 per 1M input tokens and $2.10 per 1M output tokens. Context window: 131K. Prices verified 2026-08-23 Β· source

Input / 1M tokens
$0.180
Output / 1M tokens
$2.10
Context window
131K
Max output
33K

What would Qwen: Qwen3 VL 8B Thinking cost you per month?

WorkloadRequests/moAvg tokens (in/out)Est. monthly cost
Light chatbot10,000800 / 400$9.84
Coding agent5,0008,000 / 2,000$28.20
RAG pipeline30,0004,000 / 500$53.10
Doc summarizer2,00020,000 / 1,000$11.40

Estimates use list prices without prompt caching. Caching can cut input cost by up to 90% on repeated context.

About Qwen: Qwen3 VL 8B Thinking

Qwen3-VL-8B-Thinking is the reasoning-optimized variant of the Qwen3-VL-8B multimodal model, designed for advanced visual and textual reasoning across complex scenes, documents, and temporal sequences. It integrates enhanced multimodal alignment and...

Compare Qwen: Qwen3 VL 8B Thinking with

Official resources