Qwen3.6 Flash pricing and specs

Qwen vision-language model for visual reasoning, documents, and agent tasks

15,85 ₽input / 1M$0.1875
95,07 ₽output / 1M$1.125
1Mcontexttokens
66Kmax outputtokens

Qwen3.6 Flash price per 1M tokens

Type₽ / 1M$ / 1M
Input tokens15,85 ₽$0.1875
Output tokens95,07 ₽$1.125
Cache write19,81 ₽$0.2344

Prices in rubles as of 18.09.2026 at the CBR rate: 1 $ = 84.51 ₽.

Qwen3.6 Flash specs and limits

Context window
1M
Max output
66K
Input
text, image, video
Output
text
Family
qwen3.6
Released
2026-04-27
Updated
2026-04-27
Also known as
qwen flash, qwen3 flash, qwen 3 flash, qwen3.5 flash, qwen3.7 flash, alibaba qwen flash

What Qwen3.6 Flash can do

  • Reasoning
  • Tool calling
  • Structured output
  • Attachments / files
  • Vision (images)
  • Temperature control
  • Open weights

Frequently asked questions

How much does the Qwen3.6 Flash API cost?

Input costs $0.1875 per 1M tokens (15.85 ₽) and output costs $1.125 per 1M (95.07 ₽). Alibaba bills the prompt and the completion separately, so the price of a call depends on how long both are. Ruble figures use the Bank of Russia rate of 84.51 ₽ per dollar for September 2026.

How much do 1,000 Qwen3.6 Flash tokens cost?

List prices are quoted per 1M tokens, so 1,000 tokens cost a thousandth of that: divide $0.1875 for input and $1.125 for output by 1,000. The same rates in rubles are 15.85 ₽ and 95.07 ₽ per 1M tokens. Both prompt tokens and completion tokens are billed.

How many tokens does Qwen3.6 Flash hold?

The Qwen3.6 Flash context window is 1M tokens (1,000,000). Everything shares that budget: the system prompt, the dialogue history, attached files and the answer the model writes. The figure comes from the Alibaba model card as of September 2026.

What are the Qwen3.6 Flash limits?

The context window is 1M tokens (1,000,000) and a single response is capped at 66K tokens, so longer output has to be generated in parts. Rate limits are set by Alibaba per account and depend on your plan rather than on the model, so they are not listed here.

How is the Qwen3.6 Flash price in rubles calculated?

Providers quote their tariffs in US dollars, so the ruble amounts here are a conversion at the official Bank of Russia rate of 84.51 ₽ per dollar. For Qwen3.6 Flash that gives 15.85 ₽ per 1M input tokens and 95.07 ₽ per 1M output tokens. When the rate moves, the ruble price moves with it.

How is Qwen3.6 Flash different from other Qwen models?

Qwen3.6 Flash belongs to the Qwen3.6 family at Alibaba. Models in the Qwen line differ in context size, price per 1M tokens and supported capabilities, so compare them by the numbers on their cards rather than by the name. Every Qwen model is listed on the provider page.

Can Qwen3.6 Flash reason?

Yes, Qwen3.6 Flash declares a reasoning mode: it works through intermediate steps before it answers. Those steps add output tokens, so a call with a long reasoning chain costs more and takes longer than a plain completion. Capabilities declared for the model: reasoning, tool calling, image input, structured output.

What capabilities does Qwen3.6 Flash support?

The declared capabilities of Qwen3.6 Flash are: reasoning, tool calling, image input, structured output. It accepts text, images, video as input, and every attachment consumes tokens from the shared context window of 1M tokens. The list comes from the Alibaba model card and is refreshed with the catalog as of September 2026.

Other Alibaba models