Skip to content
AI Dev Toolkit.
Esc
  • AI Token CounterCount tokens for GPT, Claude, Gemini, DeepSeek, Qwen and more.Tool
  • LLM API Cost CalculatorEstimate per-request, daily and monthly API costs.Tool
  • AI Model ComparisonCompare prices, context windows and features across models.Tool
  • AI Model Pricing PagesSpecs, real costs and cheaper alternatives for popular models.Tool
  • Context Window CheckerSee whether your text fits each model's context window.Tool
  • Subscription vs API CalculatorFind out whether a chat plan or the API is cheaper for you.Tool
  • GPU / VRAM CalculatorCheck how much VRAM a local model needs and which GPUs fit.Tool
  • Claude Code Error DatabaseExact Claude Code error messages with tested fixes.Tool

Provider

Qwen API pricing: every model compared

Qwen offers 53 models through its API in our data, priced from $0.03 to $4 per million input tokens. The newest is Qwen3.8 Max Prime, first listed on 2026-09-23.

Models
53
Cheapest input
$0.03
Priciest input
$4
Largest context
1,048,576
With caching
19 of 53
Open weights
35 of 53

Filter and sort Qwen models in the comparison table

Cheapest

Qwen3.7 FlashQwen$0.03 in · $0.13 out

Largest context

Qwen3.8 2.4T A95BQwen1,048,576 tokens

Newest

Qwen3.8 Max PrimeQwen2026-09-23

Line-up

Every Qwen model and price

ModelInput / 1MOutput / 1MContextReleased
Qwen3.8 Max Prime$4$121,000,0002026-09-23
Qwen3.8 Omni Flash$0.15$0.471,000,0002026-09-21
Qwen3.8 Max (0902)$2$61,000,0002026-09-03
Qwen3.8 Flash$0.15$0.471,000,0002026-08-26
Qwen3.8 27B$0.425$2.551,000,0002026-08-14
Qwen3.8 2.4T A95B$2$61,048,5762026-08-12
Qwen3.7 Flash$0.03$0.131,000,0002026-07-27
Qwen3.7 Plus$0.32$1.281,000,0002026-06-03
Qwen3.7 Max$1.475$4.4251,000,0002026-05-21
Qwen3.5 Plus 2026-04-20$0.30$1.801,000,0002026-04-27
Qwen3.6 27B$0.45$2.70262,1442026-04-27
Qwen3.6 35B A3B$0.15$1262,1442026-04-27
Qwen3.6 Flash$0.1875$1.1251,000,0002026-04-27
Qwen3.6 Max Preview$1.027$6.162262,1442026-04-27
Qwen3.6 Plus$0.325$1.951,000,0002026-04-02
Qwen3.5-9B$0.10$0.15256,0002026-03-10
Qwen3.5-122B-A10B$0.26$2.08262,1442026-02-25
Qwen3.5-27B$0.26$2.60262,1442026-02-25
Qwen3.5-35B-A3B$0.1625$1.30262,1442026-02-25
Qwen3.5-Flash$0.065$0.261,000,0002026-02-25
Qwen3.5 397B A17B$0.55$3.50262,1442026-02-16
Qwen3.5 Plus 2026-02-15$0.26$1.561,000,0002026-02-16
Qwen3 Max Thinking$0.78$3.90262,1442026-02-09
Qwen3 Coder Next$0.12$0.80262,1442026-02-04
Qwen3 VL 32B Instruct$0.104$0.416131,0722025-10-23
Qwen3 VL 8B Instruct$0.117$0.455131,0722025-10-14
Qwen3 VL 8B Thinking$0.18$2.10131,0722025-10-14
Qwen3 VL 30B A3B Instruct$0.15$0.60262,1442025-10-06
Qwen3 VL 30B A3B Thinking$0.20$2.40131,0722025-10-06
Qwen3 Coder Plus$0.65$3.251,000,0002025-09-23
Qwen3 Max$0.78$3.90262,1442025-09-23
Qwen3 VL 235B A22B Instruct$0.21$1.90131,0722025-09-23
Qwen3 VL 235B A22B Thinking$0.40$4131,0722025-09-23
Qwen3 Coder Flash$0.195$0.9751,000,0002025-09-17
Qwen3 Next 80B A3B Instruct$0.09$1.10262,1442025-09-11
Qwen3 Next 80B A3B Thinking$0.15$1.20131,0722025-09-11
Qwen Plus 0728$0.26$0.781,000,0002025-09-08
Qwen3 30B A3B Thinking 2507$0.20$2.4081,9202025-08-28
Qwen3 Coder 30B A3B Instruct$0.07$0.28262,1442025-07-31
Qwen3 30B A3B Instruct 2507$0.10$0.30262,1442025-07-29
Qwen3 235B A22B Thinking 2507$0.23$2.30131,0722025-07-25
Qwen3 Coder 480B A35B$0.30$1262,1442025-07-23
Qwen3 235B A22B Instruct 2507$0.09$0.55262,1442025-07-21
Qwen3 14B$0.12$0.2440,9602025-04-28
Qwen3 235B A22B$0.455$1.82131,0722025-04-28
Qwen3 30B A3B$0.12$0.5040,9602025-04-28
Qwen3 32B$0.08$0.2840,9602025-04-28
Qwen3 8B$0.117$0.455131,0722025-04-28
Qwen-Plus$0.26$0.781,000,0002025-02-01
Qwen2.5 VL 72B Instruct$0.80$1128,0002025-02-01
Qwen2.5 Coder 32B Instruct$0.66$132,7682024-11-11
Qwen2.5 7B Instruct$0.10$0.2032,7682024-10-16
Qwen2.5 72B Instruct$0.36$0.4032,7682024-09-19

Prices in US dollars per 1M tokens, newest first. Release dates are when our source first listed the model.

Overview

Qwen’s pricing at a glance

The spread between Qwen’s cheapest and most expensive model is large: Qwen3.8 Max Prime costs 133× more per input token than Qwen3.7 Flash. Picking the smallest model that does the job well is usually the biggest saving available, ahead of any discount.

19 of 53 of the line-up has a cached-input price, which helps chatbots and agents that resend the same instructions. No batch prices are published in our data. 27 of 53 accept images as input.

FAQ

Qwen API questions

How much does the Qwen API cost?

Qwen's 53 models range from $0.03 to $4 per million input tokens, and output costs more than input on almost every model. The exact bill depends on your token volumes, which you can estimate in the LLM cost calculator.

What is Qwen's cheapest model?

By blended price (three parts input to one part output) it is Qwen3.7 Flash, at $0.03 input and $0.13 output per million tokens. The most expensive is Qwen3.8 Max Prime.

Which Qwen model has the largest context window?

Qwen3.8 2.4T A95B, with 1,048,576 tokens per request.

Does Qwen offer prompt caching or batch discounts?

In our data, 19 of 53 of Qwen's models have a published cached-input price and none have a published batch price. Both can cut costs substantially for repeated prompts or work that can wait.

Updated

Sources: OpenRouter models API and LiteLLM model prices and context windows, checked daily. Not affiliated with Qwen. Confirm critical numbers on the provider’s pricing page. See our methodology.