Qwen API pricing per 1M tokens

Qwen API pricing starts at $0.035 per 1M input tokens with Qwen3-ASR Flash. The top tier is Kimi K3. Ruble figures use the CBR rate for Sep 18, 2026.

Prices in rubles as of 18.09.2026 at the CBR rate: 1 $ = 84.51 ₽.

56modelsin the catalog
2,96 ₽from / 1M input$0.035
1.0Mmax contexttokens
Docsdocumentation

Qwen API price per 1M tokens

Currency
Sort
ModelContextInput / 1MOutput / 1MCapabilities
Qwen3.8 Flash
2026-08-26
1M12,68 ₽$0.1539,72 ₽$0.47
reasoningtoolsvisionstructured
Qwen3.8 Max
2026-08-03
1M169 ₽$2507 ₽$6
reasoningtoolsvisionstructured
DeepSeek V4 Flash 0731
2026-07-31
1M16,9 ₽$0.233,8 ₽$0.4
reasoningtoolsstructuredopen weights
Kimi K3
2026-07-16
1.0M254 ₽$31 268 ₽$15
reasoningtoolsvisionstructuredopen weights
GLM-5.2
2026-06-13
1M118 ₽$1.4372 ₽$4.4
reasoningtoolsstructuredopen weights
Qwen3.7 Plus
2026-06-02
1M42,25 ₽$0.5254 ₽$3
reasoningtoolsvision
Qwen3.7 Max
2026-05-21
1M211 ₽$2.5634 ₽$7.5
reasoningtools
Qwen3.6 Flash
2026-04-27
1M15,85 ₽$0.187595,07 ₽$1.125
reasoningtoolsvisionstructured
Qwen3.6 27B
2026-04-22
262K50,71 ₽$0.6304 ₽$3.6
reasoningtoolsvisionstructuredopen weights
Qwen3.6 Max Preview
2026-04-20
262K110 ₽$1.3659 ₽$7.8
reasoningtools
Qwen3.6 35B-A3B
2026-04-17
262K20,96 ₽$0.248125 ₽$1.485
reasoningtoolsvisionstructuredopen weights
Qwen3.6 Plus
2026-04-02
1M42,25 ₽$0.5254 ₽$3
reasoningtoolsvision
Qwen3.5 122B-A10B
2026-02-23
262K33,8 ₽$0.4270 ₽$3.2
reasoningtoolsvisionstructuredopen weights
Qwen3.5 27B
2026-02-23
262K25,35 ₽$0.3203 ₽$2.4
reasoningtoolsvisionstructuredopen weights
Qwen3.5 35B-A3B
2026-02-23
262K21,13 ₽$0.25169 ₽$2
reasoningtoolsvisionstructuredopen weights
Qwen3.5 Plus
2026-02-16
1M33,8 ₽$0.4203 ₽$2.4
reasoningtoolsvision
Qwen3.5 397B-A17B
2026-02-15
262K50,71 ₽$0.6304 ₽$3.6
reasoningtoolsvisionstructuredopen weights
Qwen3 Max
2025-09-23
262K101 ₽$1.2507 ₽$6
tools
Qwen3-VL Plus
2025-09-23
262K16,9 ₽$0.2135 ₽$1.6
reasoningtoolsvision
Qwen3-LiveTranslate Flash Realtime
2025-09-22
53K845 ₽$10845 ₽$10
vision
Qwen3-Omni Flash
2025-09-15
66K36,34 ₽$0.43140 ₽$1.66
reasoningtoolsvision
Qwen3-Omni Flash Realtime
2025-09-15
66K43,94 ₽$0.52168 ₽$1.99
toolsvision
Qwen3-ASR Flash
2025-09-08
53K2,96 ₽$0.0352,96 ₽$0.035
Qwen3-Next 80B-A3B (Thinking)
2025-09-01
131K42,25 ₽$0.5507 ₽$6
reasoningtoolsopen weights
Qwen3-Next 80B-A3B Instruct
2025-09-01
131K42,25 ₽$0.5169 ₽$2
toolsopen weights
Qwen Flash
2025-07-28
1M4,23 ₽$0.0533,8 ₽$0.4
reasoningtools
Qwen3 Coder Flash
2025-07-28
1M25,35 ₽$0.3127 ₽$1.5
tools
Qwen3 Coder Plus
2025-07-23
1.0M84,51 ₽$1423 ₽$5
toolsopen weights
Qwen-Omni Turbo Realtime
2025-05-08
33K22,82 ₽$0.2790,42 ₽$1.07
toolsvision
Qwen3 14B
2025-04-01
131K29,58 ₽$0.35118 ₽$1.4
reasoningtoolsopen weights
Qwen3 235B-A22B
2025-04-01
131K59,16 ₽$0.7237 ₽$2.8
reasoningtoolsopen weights
Qwen3 32B
2025-04-01
131K59,16 ₽$0.7237 ₽$2.8
reasoningtoolsopen weights
Qwen3 8B
2025-04-01
131K15,21 ₽$0.1859,16 ₽$0.7
reasoningtoolsopen weights
Qwen3-Coder 30B-A3B Instruct
2025-04-01
262K38,03 ₽$0.45190 ₽$2.25
toolsopen weights
Qwen3-Coder 480B-A35B Instruct
2025-04-01
262K127 ₽$1.5634 ₽$7.5
toolsopen weights
Qwen3-VL 235B-A22B
2025-04-01
131K59,16 ₽$0.7237 ₽$2.8
reasoningtoolsvisionopen weights
Qwen3-VL 30B-A3B
2025-04-01
131K16,9 ₽$0.267,61 ₽$0.8
reasoningtoolsvisionopen weights
QVQ Max
2025-03-25
131K101 ₽$1.2406 ₽$4.8
reasoningtoolsvision
QwQ Plus
2025-03-05
131K67,61 ₽$0.8203 ₽$2.4
reasoningtools
Qwen-Omni Turbo
2025-01-19
33K5,92 ₽$0.0722,82 ₽$0.27
toolsvision
Qwen-MT Plus
2025-01-01
16K208 ₽$2.46623 ₽$7.37
Qwen-MT Turbo
2025-01-01
16K13,52 ₽$0.1641,41 ₽$0.49
Qwen2.5-Omni 7B
2024-12-01
33K8,45 ₽$0.133,8 ₽$0.4
toolsvisionopen weights
Qwen Turbo
2024-11-01
1M4,23 ₽$0.0516,9 ₽$0.2
reasoningtools
Qwen-VL OCR
2024-10-28
34K60,85 ₽$0.7260,85 ₽$0.72
vision
Qwen2.5 14B Instruct
2024-09-01
131K29,58 ₽$0.35118 ₽$1.4
toolsopen weights
Qwen2.5 32B Instruct
2024-09-01
131K59,16 ₽$0.7237 ₽$2.8
toolsopen weights
Qwen2.5 72B Instruct
2024-09-01
131K118 ₽$1.4473 ₽$5.6
toolsopen weights
Qwen2.5 7B Instruct
2024-09-01
131K14,79 ₽$0.17559,16 ₽$0.7
toolsopen weights
Qwen2.5-VL 72B Instruct
2024-09-01
131K237 ₽$2.8710 ₽$8.4
toolsvisionopen weights
Qwen2.5-VL 7B Instruct
2024-09-01
131K29,58 ₽$0.3588,73 ₽$1.05
toolsvisionopen weights
Qwen-VL Max
2024-04-08
131K67,61 ₽$0.8270 ₽$3.2
toolsvision
Qwen Max
2024-04-03
33K135 ₽$1.6541 ₽$6.4
tools
Qwen Plus
2024-01-25
1M33,8 ₽$0.4101 ₽$1.2
reasoningtools
Qwen-VL Plus
2024-01-25
131K17,75 ₽$0.2153,24 ₽$0.63
toolsvision
Qwen Plus Character (Japanese)
2024-01-01
8K42,25 ₽$0.5118 ₽$1.4
tools

Models shown: 56 / 56

About the Alibaba API

Qwen is Alibaba Cloud's (Tongyi) model line-up. The catalog lists 51 models: the hosted Qwen Max, Plus, Turbo and Flash tiers are API-only, while Qwen2.5, Qwen3, Qwen3.5 and Qwen3.6 ship with open weights, alongside the multimodal Qwen-VL and Qwen-Omni models and the visual-reasoning QVQ Max.

Qwen publishes API prices per million tokens, billed separately for input and output, so the cost of a request depends on prompt and completion length. The table above lists 56 models with per-1M pricing, context window and maximum output.

Frequently asked questions

How much does the Qwen API cost?

There is no single rate — it depends on the model. The lowest input price across Qwen models is 2.96 ₽ per 1M tokens, converted from the provider's dollar price list, and the top-tier model is Kimi K3. The table above lists input and output rates for every Alibaba model. Data as of September 2026.

Which Qwen model is the cheapest?

By input price it is Qwen3-ASR Flash at 2.96 ₽ per 1M input tokens. The final bill also depends on the output rate, which varies far more between models: the longer the answers, the bigger the gap. Compare the Input / 1M and Output / 1M columns before you commit to a model.

How many Qwen models are in the catalog?

Alibaba models in the catalog: 56, as of September 2026. The list is built automatically from the provider's public data — new models appear at the next sync, while retired ones stay in place with a deprecated status. Every model has its own page with prices and limits.

How are Qwen prices converted to rubles?

Alibaba publishes its price list in dollars per 1M tokens, and the ruble figures are a conversion at the official Bank of Russia rate, currently 1 $ = 84.51 ₽. The rate and the date it was taken are shown in the page header, and the currency switch brings back the original dollar prices.

How do the Qwen model lines differ?

The lines differ in size, price and purpose: there are lightweight models starting at 2.96 ₽ per 1M input tokens, and top-tier ones such as Kimi K3. The table columns show exactly what changes — context window, input and output prices, capabilities. Details live on each model page.

How often are Qwen prices updated?

The catalog syncs every day at 06:00 Moscow time: provider price lists, limits, the model list and the Bank of Russia rate are all re-read. If Alibaba changes a rate, the new price shows up here after the next sync. Current figures are as of September 2026.

Do Qwen models support reasoning?

Yes, some Qwen models support a reasoning mode — the Capabilities column flags it in the table, and each model page lists it separately. Keep in mind that reasoning consumes output tokens, so a request costs more even when the per-1M rate stays the same.

Does Qwen have open-weight models?

Yes, some Qwen models ship with published weights — the model page states this together with the license. You can run them on your own infrastructure, where the cost depends on your hardware rather than on a price list. Catalog prices refer to the provider's own API.

How much do 1,000 Qwen tokens cost?

Providers quote prices per 1M tokens, so 1,000 tokens cost exactly a thousandth of that. The cheapest option in the catalog is Qwen3-ASR Flash at $0.035 per 1M input tokens.