Qwen API pricing per 1M tokens
Qwen API pricing starts at $0.035 per 1M input tokens with Qwen3-ASR Flash. The top tier is Kimi K3. Ruble figures use the CBR rate for Sep 18, 2026.
Prices in rubles as of 18.09.2026 at the CBR rate: 1 $ = 84.51 ₽.
Qwen API price per 1M tokens
| Model | Context | Input / 1M | Output / 1M | Capabilities | |
|---|---|---|---|---|---|
| Qwen3.8 Flash 2026-08-26 | 1M | 12,68 ₽$0.15 | 39,72 ₽$0.47 | reasoningtoolsvisionstructured | |
| Qwen3.8 Max 2026-08-03 | 1M | 169 ₽$2 | 507 ₽$6 | reasoningtoolsvisionstructured | |
| DeepSeek V4 Flash 0731 2026-07-31 | 1M | 16,9 ₽$0.2 | 33,8 ₽$0.4 | reasoningtoolsstructuredopen weights | |
| Kimi K3 2026-07-16 | 1.0M | 254 ₽$3 | 1 268 ₽$15 | reasoningtoolsvisionstructuredopen weights | |
| GLM-5.2 2026-06-13 | 1M | 118 ₽$1.4 | 372 ₽$4.4 | reasoningtoolsstructuredopen weights | |
| Qwen3.7 Plus 2026-06-02 | 1M | 42,25 ₽$0.5 | 254 ₽$3 | reasoningtoolsvision | |
| Qwen3.7 Max 2026-05-21 | 1M | 211 ₽$2.5 | 634 ₽$7.5 | reasoningtools | |
| Qwen3.6 Flash 2026-04-27 | 1M | 15,85 ₽$0.1875 | 95,07 ₽$1.125 | reasoningtoolsvisionstructured | |
| Qwen3.6 27B 2026-04-22 | 262K | 50,71 ₽$0.6 | 304 ₽$3.6 | reasoningtoolsvisionstructuredopen weights | |
| Qwen3.6 Max Preview 2026-04-20 | 262K | 110 ₽$1.3 | 659 ₽$7.8 | reasoningtools | |
| Qwen3.6 35B-A3B 2026-04-17 | 262K | 20,96 ₽$0.248 | 125 ₽$1.485 | reasoningtoolsvisionstructuredopen weights | |
| Qwen3.6 Plus 2026-04-02 | 1M | 42,25 ₽$0.5 | 254 ₽$3 | reasoningtoolsvision | |
| Qwen3.5 122B-A10B 2026-02-23 | 262K | 33,8 ₽$0.4 | 270 ₽$3.2 | reasoningtoolsvisionstructuredopen weights | |
| Qwen3.5 27B 2026-02-23 | 262K | 25,35 ₽$0.3 | 203 ₽$2.4 | reasoningtoolsvisionstructuredopen weights | |
| Qwen3.5 35B-A3B 2026-02-23 | 262K | 21,13 ₽$0.25 | 169 ₽$2 | reasoningtoolsvisionstructuredopen weights | |
| Qwen3.5 Plus 2026-02-16 | 1M | 33,8 ₽$0.4 | 203 ₽$2.4 | reasoningtoolsvision | |
| Qwen3.5 397B-A17B 2026-02-15 | 262K | 50,71 ₽$0.6 | 304 ₽$3.6 | reasoningtoolsvisionstructuredopen weights | |
| Qwen3 Max 2025-09-23 | 262K | 101 ₽$1.2 | 507 ₽$6 | tools | |
| Qwen3-VL Plus 2025-09-23 | 262K | 16,9 ₽$0.2 | 135 ₽$1.6 | reasoningtoolsvision | |
| Qwen3-LiveTranslate Flash Realtime 2025-09-22 | 53K | 845 ₽$10 | 845 ₽$10 | vision | |
| Qwen3-Omni Flash 2025-09-15 | 66K | 36,34 ₽$0.43 | 140 ₽$1.66 | reasoningtoolsvision | |
| Qwen3-Omni Flash Realtime 2025-09-15 | 66K | 43,94 ₽$0.52 | 168 ₽$1.99 | toolsvision | |
| Qwen3-ASR Flash 2025-09-08 | 53K | 2,96 ₽$0.035 | 2,96 ₽$0.035 | — | |
| Qwen3-Next 80B-A3B (Thinking) 2025-09-01 | 131K | 42,25 ₽$0.5 | 507 ₽$6 | reasoningtoolsopen weights | |
| Qwen3-Next 80B-A3B Instruct 2025-09-01 | 131K | 42,25 ₽$0.5 | 169 ₽$2 | toolsopen weights | |
| Qwen Flash 2025-07-28 | 1M | 4,23 ₽$0.05 | 33,8 ₽$0.4 | reasoningtools | |
| Qwen3 Coder Flash 2025-07-28 | 1M | 25,35 ₽$0.3 | 127 ₽$1.5 | tools | |
| Qwen3 Coder Plus 2025-07-23 | 1.0M | 84,51 ₽$1 | 423 ₽$5 | toolsopen weights | |
| Qwen-Omni Turbo Realtime 2025-05-08 | 33K | 22,82 ₽$0.27 | 90,42 ₽$1.07 | toolsvision | |
| Qwen3 14B 2025-04-01 | 131K | 29,58 ₽$0.35 | 118 ₽$1.4 | reasoningtoolsopen weights | |
| Qwen3 235B-A22B 2025-04-01 | 131K | 59,16 ₽$0.7 | 237 ₽$2.8 | reasoningtoolsopen weights | |
| Qwen3 32B 2025-04-01 | 131K | 59,16 ₽$0.7 | 237 ₽$2.8 | reasoningtoolsopen weights | |
| Qwen3 8B 2025-04-01 | 131K | 15,21 ₽$0.18 | 59,16 ₽$0.7 | reasoningtoolsopen weights | |
| Qwen3-Coder 30B-A3B Instruct 2025-04-01 | 262K | 38,03 ₽$0.45 | 190 ₽$2.25 | toolsopen weights | |
| Qwen3-Coder 480B-A35B Instruct 2025-04-01 | 262K | 127 ₽$1.5 | 634 ₽$7.5 | toolsopen weights | |
| Qwen3-VL 235B-A22B 2025-04-01 | 131K | 59,16 ₽$0.7 | 237 ₽$2.8 | reasoningtoolsvisionopen weights | |
| Qwen3-VL 30B-A3B 2025-04-01 | 131K | 16,9 ₽$0.2 | 67,61 ₽$0.8 | reasoningtoolsvisionopen weights | |
| QVQ Max 2025-03-25 | 131K | 101 ₽$1.2 | 406 ₽$4.8 | reasoningtoolsvision | |
| QwQ Plus 2025-03-05 | 131K | 67,61 ₽$0.8 | 203 ₽$2.4 | reasoningtools | |
| Qwen-Omni Turbo 2025-01-19 | 33K | 5,92 ₽$0.07 | 22,82 ₽$0.27 | toolsvision | |
| Qwen-MT Plus 2025-01-01 | 16K | 208 ₽$2.46 | 623 ₽$7.37 | — | |
| Qwen-MT Turbo 2025-01-01 | 16K | 13,52 ₽$0.16 | 41,41 ₽$0.49 | — | |
| Qwen2.5-Omni 7B 2024-12-01 | 33K | 8,45 ₽$0.1 | 33,8 ₽$0.4 | toolsvisionopen weights | |
| Qwen Turbo 2024-11-01 | 1M | 4,23 ₽$0.05 | 16,9 ₽$0.2 | reasoningtools | |
| Qwen-VL OCR 2024-10-28 | 34K | 60,85 ₽$0.72 | 60,85 ₽$0.72 | vision | |
| Qwen2.5 14B Instruct 2024-09-01 | 131K | 29,58 ₽$0.35 | 118 ₽$1.4 | toolsopen weights | |
| Qwen2.5 32B Instruct 2024-09-01 | 131K | 59,16 ₽$0.7 | 237 ₽$2.8 | toolsopen weights | |
| Qwen2.5 72B Instruct 2024-09-01 | 131K | 118 ₽$1.4 | 473 ₽$5.6 | toolsopen weights | |
| Qwen2.5 7B Instruct 2024-09-01 | 131K | 14,79 ₽$0.175 | 59,16 ₽$0.7 | toolsopen weights | |
| Qwen2.5-VL 72B Instruct 2024-09-01 | 131K | 237 ₽$2.8 | 710 ₽$8.4 | toolsvisionopen weights | |
| Qwen2.5-VL 7B Instruct 2024-09-01 | 131K | 29,58 ₽$0.35 | 88,73 ₽$1.05 | toolsvisionopen weights | |
| Qwen-VL Max 2024-04-08 | 131K | 67,61 ₽$0.8 | 270 ₽$3.2 | toolsvision | |
| Qwen Max 2024-04-03 | 33K | 135 ₽$1.6 | 541 ₽$6.4 | tools | |
| Qwen Plus 2024-01-25 | 1M | 33,8 ₽$0.4 | 101 ₽$1.2 | reasoningtools | |
| Qwen-VL Plus 2024-01-25 | 131K | 17,75 ₽$0.21 | 53,24 ₽$0.63 | toolsvision | |
| Qwen Plus Character (Japanese) 2024-01-01 | 8K | 42,25 ₽$0.5 | 118 ₽$1.4 | tools |
Models shown: 56 / 56
About the Alibaba API
Qwen is Alibaba Cloud's (Tongyi) model line-up. The catalog lists 51 models: the hosted Qwen Max, Plus, Turbo and Flash tiers are API-only, while Qwen2.5, Qwen3, Qwen3.5 and Qwen3.6 ship with open weights, alongside the multimodal Qwen-VL and Qwen-Omni models and the visual-reasoning QVQ Max.
Qwen publishes API prices per million tokens, billed separately for input and output, so the cost of a request depends on prompt and completion length. The table above lists 56 models with per-1M pricing, context window and maximum output.
Frequently asked questions
How much does the Qwen API cost?
There is no single rate — it depends on the model. The lowest input price across Qwen models is 2.96 ₽ per 1M tokens, converted from the provider's dollar price list, and the top-tier model is Kimi K3. The table above lists input and output rates for every Alibaba model. Data as of September 2026.
Which Qwen model is the cheapest?
By input price it is Qwen3-ASR Flash at 2.96 ₽ per 1M input tokens. The final bill also depends on the output rate, which varies far more between models: the longer the answers, the bigger the gap. Compare the Input / 1M and Output / 1M columns before you commit to a model.
How many Qwen models are in the catalog?
Alibaba models in the catalog: 56, as of September 2026. The list is built automatically from the provider's public data — new models appear at the next sync, while retired ones stay in place with a deprecated status. Every model has its own page with prices and limits.
How are Qwen prices converted to rubles?
Alibaba publishes its price list in dollars per 1M tokens, and the ruble figures are a conversion at the official Bank of Russia rate, currently 1 $ = 84.51 ₽. The rate and the date it was taken are shown in the page header, and the currency switch brings back the original dollar prices.
How do the Qwen model lines differ?
The lines differ in size, price and purpose: there are lightweight models starting at 2.96 ₽ per 1M input tokens, and top-tier ones such as Kimi K3. The table columns show exactly what changes — context window, input and output prices, capabilities. Details live on each model page.
How often are Qwen prices updated?
The catalog syncs every day at 06:00 Moscow time: provider price lists, limits, the model list and the Bank of Russia rate are all re-read. If Alibaba changes a rate, the new price shows up here after the next sync. Current figures are as of September 2026.
Do Qwen models support reasoning?
Yes, some Qwen models support a reasoning mode — the Capabilities column flags it in the table, and each model page lists it separately. Keep in mind that reasoning consumes output tokens, so a request costs more even when the per-1M rate stays the same.
Does Qwen have open-weight models?
Yes, some Qwen models ship with published weights — the model page states this together with the license. You can run them on your own infrastructure, where the cost depends on your hardware rather than on a price list. Catalog prices refer to the provider's own API.
How much do 1,000 Qwen tokens cost?
Providers quote prices per 1M tokens, so 1,000 tokens cost exactly a thousandth of that. The cheapest option in the catalog is Qwen3-ASR Flash at $0.035 per 1M input tokens.