Skip to content

Qwen API pricing

54 Qwen models, ranked by what your workload would cost per month. Prices as listed on OpenRouter at 3 Oct 2026, 08:40 UTC.

Your workload

90M input and 18M output tokens a month (30-day month; token counts are per request).

Assumes: 1,000 conversations a day, each sending about 3,000 tokens of history and instructions and receiving about 600 tokens of replies.

16 of 16 models, sorted by your monthly cost — loading the complete list…

Each row: input / output price, then your monthly cost. Token prices per 1M tokens. Tick up to 4 models to compare.

Freea month
$5.04a month
$7.81a month
$10.53a month
$11.34a month
$11.70a month
$12.24a month
$12.60a month
$14.17a month
$15.12a month
$16.85a month
$21.96a month
$21.96a month
$91.80a month
$288a month
$576a month

“Your cost” applies the cache-read price only where the source lists one; a model with no price data shows a dash and is never ranked as cheapest.

Collected from the OpenRouter models API (466 entries). 99 models are cross-checked against provider list prices in the LiteLLM price table.

Qwen prices at a glance

  • 54 models listed, 53 of them with a paid price. Batch variants are folded into their base model.
  • Lowest input price: Qwen3.7 Flash, at $0.03 per million input tokens and $0.13 per million output tokens.
  • Highest input price: Qwen3.8 Max Prime, at $4.00 input and $12.00 output per million tokens.
  • 22 models publish a cache-read price; the median discount on input is 80% (from 30% to 89%).

These are OpenRouter prices. Open a model in the table to see the provider list price recorded by LiteLLM, where there is one, and whether the two differ. To understand the discounts, read the guide to prompt caching and batch pricing.

Qwen models with their own page

Questions and answers

How many Qwen models are listed?

54 at the collection of 3 Oct 2026, 08:40 UTC, of which 53 have a paid price on OpenRouter. Batch variants are folded into their base model.

What is the cheapest Qwen model?

By input price, Qwen3.7 Flash: $0.03 per million input tokens and $0.13 per million output tokens. Whether it is the cheapest for you depends on your mix of input and output, which is what the table on this page calculates.

What is the most expensive Qwen model?

By input price, Qwen3.8 Max Prime: $4.00 per million input tokens and $12.00 per million output tokens.

Do Qwen models have cached-input prices?

22 of the 54 listed models publish a cache-read price. The table shows it in the “Cached in” column.

Are these the prices Qwen charges directly?

They are the prices on OpenRouter. Where LiteLLM records the provider list price for a model, opening the model shows both, and the row is flagged when they differ by more than 1%.