Qwen-Omni Turbo Realtime pricing and specs

Qwen omni model for text, vision, audio, and multimodal agent tasks

22,82 ₽input / 1M$0.27
90,42 ₽output / 1M$1.07
33Kcontexttokens
2Kmax outputtokens

Qwen-Omni Turbo Realtime price per 1M tokens

Type₽ / 1M$ / 1M
Input tokens22,82 ₽$0.27
Output tokens90,42 ₽$1.07

Prices in rubles as of 18.09.2026 at the CBR rate: 1 $ = 84.51 ₽.

Qwen-Omni Turbo Realtime specs and limits

Context window
33K
Max output
2K
Input
text, image, audio
Output
text, audio
Family
qwen
Released
2025-05-08
Updated
2025-05-08
Knowledge cutoff
2024-04
Also known as
qwen omni turbo realtime, qwen3 omni turbo realtime, qwen 3 omni turbo realtime, qwen3.5 omni turbo realtime, qwen3.6 omni turbo realtime, qwen3.7 omni turbo realtime

What Qwen-Omni Turbo Realtime can do

  • Reasoning
  • Tool calling
  • Structured output
  • Attachments / files
  • Vision (images)
  • Temperature control
  • Open weights

Frequently asked questions

How much does the Qwen-Omni Turbo Realtime API cost?

Input costs $0.27 per 1M tokens (22.82 ₽) and output costs $1.07 per 1M (90.42 ₽). Alibaba bills the prompt and the completion separately, so the price of a call depends on how long both are. Ruble figures use the Bank of Russia rate of 84.51 ₽ per dollar for September 2026.

How much do 1,000 Qwen-Omni Turbo Realtime tokens cost?

List prices are quoted per 1M tokens, so 1,000 tokens cost a thousandth of that: divide $0.27 for input and $1.07 for output by 1,000. The same rates in rubles are 22.82 ₽ and 90.42 ₽ per 1M tokens. Both prompt tokens and completion tokens are billed.

How many tokens does Qwen-Omni Turbo Realtime hold?

The Qwen-Omni Turbo Realtime context window is 33K tokens (32,768). Everything shares that budget: the system prompt, the dialogue history, attached files and the answer the model writes. The figure comes from the Alibaba model card as of September 2026.

What are the Qwen-Omni Turbo Realtime limits?

The context window is 33K tokens (32,768) and a single response is capped at 2K tokens, so longer output has to be generated in parts. Rate limits are set by Alibaba per account and depend on your plan rather than on the model, so they are not listed here.

How is the Qwen-Omni Turbo Realtime price in rubles calculated?

Providers quote their tariffs in US dollars, so the ruble amounts here are a conversion at the official Bank of Russia rate of 84.51 ₽ per dollar. For Qwen-Omni Turbo Realtime that gives 22.82 ₽ per 1M input tokens and 90.42 ₽ per 1M output tokens. When the rate moves, the ruble price moves with it.

How is Qwen-Omni Turbo Realtime different from other Qwen models?

Qwen-Omni Turbo Realtime belongs to the Qwen family at Alibaba. Models in the Qwen line differ in context size, price per 1M tokens and supported capabilities, so compare them by the numbers on their cards rather than by the name. Every Qwen model is listed on the provider page.

What capabilities does Qwen-Omni Turbo Realtime support?

The declared capabilities of Qwen-Omni Turbo Realtime are: tool calling, image input. It accepts text, images, audio as input, and every attachment consumes tokens from the shared context window of 33K tokens. The list comes from the Alibaba model card and is refreshed with the catalog as of September 2026.

Other Alibaba models