Mistral Small (latest) pricing and specs

open weights

Efficient Mistral model for fast chat, extraction, and production assistants

12,68 ₽input / 1M$0.15
50,71 ₽output / 1M$0.6
256Kcontexttokens
256Kmax outputtokens

Mistral Small (latest) price per 1M tokens

Type₽ / 1M$ / 1M
Input tokens12,68 ₽$0.15
Output tokens50,71 ₽$0.6

Prices in rubles as of 18.09.2026 at the CBR rate: 1 $ = 84.51 ₽.

Mistral Small (latest) specs and limits

Context window
256K
Max output
256K
Input
text, image
Output
text
Family
mistral-small
Released
2026-03-16
Updated
2026-03-16
Knowledge cutoff
2025-06
Also known as
mistral ai small (latest), mistralai small (latest), le chat small (latest), mistral api small (latest), mistral llm small (latest), mistral models small (latest)

What Mistral Small (latest) can do

  • Reasoning
  • Tool calling
  • Structured output
  • Attachments / files
  • Vision (images)
  • Temperature control
  • Open weights

Model weights

Frequently asked questions

How much does the Mistral Small (latest) API cost?

Input costs $0.15 per 1M tokens (12.68 ₽) and output costs $0.6 per 1M (50.71 ₽). Mistral bills the prompt and the completion separately, so the price of a call depends on how long both are. Ruble figures use the Bank of Russia rate of 84.51 ₽ per dollar for September 2026.

How much do 1,000 Mistral Small (latest) tokens cost?

List prices are quoted per 1M tokens, so 1,000 tokens cost a thousandth of that: divide $0.15 for input and $0.6 for output by 1,000. The same rates in rubles are 12.68 ₽ and 50.71 ₽ per 1M tokens. Both prompt tokens and completion tokens are billed.

How many tokens does Mistral Small (latest) hold?

The Mistral Small (latest) context window is 256K tokens (256,000). Everything shares that budget: the system prompt, the dialogue history, attached files and the answer the model writes. The figure comes from the Mistral model card as of September 2026.

What are the Mistral Small (latest) limits?

The context window is 256K tokens (256,000) and a single response is capped at 256K tokens, so longer output has to be generated in parts. Rate limits are set by Mistral per account and depend on your plan rather than on the model, so they are not listed here.

How is the Mistral Small (latest) price in rubles calculated?

Providers quote their tariffs in US dollars, so the ruble amounts here are a conversion at the official Bank of Russia rate of 84.51 ₽ per dollar. For Mistral Small (latest) that gives 12.68 ₽ per 1M input tokens and 50.71 ₽ per 1M output tokens. When the rate moves, the ruble price moves with it.

How is Mistral Small (latest) different from other Mistral models?

Mistral Small (latest) belongs to the Mistral Small family at Mistral. Models in the Mistral line differ in context size, price per 1M tokens and supported capabilities, so compare them by the numbers on their cards rather than by the name. Every Mistral model is listed on the provider page.

Can Mistral Small (latest) reason?

Yes, Mistral Small (latest) declares a reasoning mode: it works through intermediate steps before it answers. Those steps add output tokens, so a call with a long reasoning chain costs more and takes longer than a plain completion. Capabilities declared for the model: reasoning, tool calling, image input.

What capabilities does Mistral Small (latest) support?

The declared capabilities of Mistral Small (latest) are: reasoning, tool calling, image input. It accepts text, images as input, and every attachment consumes tokens from the shared context window of 256K tokens. The list comes from the Mistral model card and is refreshed with the catalog as of September 2026.

Does Mistral Small (latest) have open weights?

Yes, Mistral Small (latest) ships with open weights, so it can be downloaded and served on your own hardware instead of being used only through the Mistral API. What you may do with it is set by the license shown on the card. Catalog prices cover hosted API access.

Other Mistral models