LLMCost
qwen

Qwen: Qwen2.5 VL 72B Instruct API Pricing

Qwen: Qwen2.5 VL 72B Instruct costs $0.800 per 1M input tokens and $1.00 per 1M output tokens (cache reads: $0.400). Context window: 128K. Prices verified 2026-08-23 Β· source

Input / 1M tokens
$0.800
Output / 1M tokens
$1.00
Context window
128K
Max output
128K

What would Qwen: Qwen2.5 VL 72B Instruct cost you per month?

WorkloadRequests/moAvg tokens (in/out)Est. monthly cost
Light chatbot10,000800 / 400$10.40
Coding agent5,0008,000 / 2,000$42.00
RAG pipeline30,0004,000 / 500$111.00
Doc summarizer2,00020,000 / 1,000$34.00

Estimates use list prices without prompt caching. Caching can cut input cost by up to 90% on repeated context.

About Qwen: Qwen2.5 VL 72B Instruct

Qwen2.5-VL is proficient in recognizing common objects such as flowers, birds, fish, and insects. It is also highly capable of analyzing texts, charts, icons, graphics, and layouts within images.

Compare Qwen: Qwen2.5 VL 72B Instruct with

Official resources