Gemini 3.5 Flash Lite pricing and specs

Fast Gemini model balancing multimodal reasoning, tool use, and cost

25,35 ₽input / 1M$0.3
211 ₽output / 1M$2.5
1.0Mcontexttokens
66Kmax outputtokens

Gemini 3.5 Flash Lite price per 1M tokens

Type₽ / 1M$ / 1M
Input tokens25,35 ₽$0.3
Output tokens211 ₽$2.5
Cache read2,54 ₽$0.03

Prices in rubles as of 18.09.2026 at the CBR rate: 1 $ = 84.51 ₽.

Gemini 3.5 Flash Lite specs and limits

Context window
1.0M
Max output
66K
Input
text, image, video, audio, pdf
Output
text
Family
gemini-flash-lite
Released
2026-07-21
Updated
2026-07-21
Knowledge cutoff
2026-03
Also known as
google gemini 3.5 flash lite, gemini api 3.5 flash lite, google ai 3.5 flash lite, google deepmind 3.5 flash lite, gemini pro 3.5 flash lite, gemini flash 3.5 flash lite

What Gemini 3.5 Flash Lite can do

  • Reasoning
  • Tool calling
  • Structured output
  • Attachments / files
  • Vision (images)
  • Temperature control
  • Open weights

Gemini 3.5 Flash Lite benchmarks

BenchmarkScoreMetric
SWE-Bench Pro54.2resolve rate
Terminal-Bench54accuracy
MLE-Bench39.2average position score
GDPval-AA1140Elo
OSWorld-Verified74success rate
CharXiv Reasoning74.5accuracy
CharXiv Reasoning76.5accuracy
GDM-MRCR72.2accuracy
GDM-MRCR21.3accuracy

Frequently asked questions

How much does the Gemini 3.5 Flash Lite API cost?

Input costs $0.3 per 1M tokens (25.35 ₽) and output costs $2.5 per 1M (211 ₽). Google bills the prompt and the completion separately, so the price of a call depends on how long both are. Ruble figures use the Bank of Russia rate of 84.51 ₽ per dollar for September 2026.

How much do 1,000 Gemini 3.5 Flash Lite tokens cost?

List prices are quoted per 1M tokens, so 1,000 tokens cost a thousandth of that: divide $0.3 for input and $2.5 for output by 1,000. The same rates in rubles are 25.35 ₽ and 211 ₽ per 1M tokens. Both prompt tokens and completion tokens are billed.

How many tokens does Gemini 3.5 Flash Lite hold?

The Gemini 3.5 Flash Lite context window is 1.0M tokens (1,048,576). Everything shares that budget: the system prompt, the dialogue history, attached files and the answer the model writes. The figure comes from the Google model card as of September 2026.

What are the Gemini 3.5 Flash Lite limits?

The context window is 1.0M tokens (1,048,576) and a single response is capped at 66K tokens, so longer output has to be generated in parts. Rate limits are set by Google per account and depend on your plan rather than on the model, so they are not listed here.

How is the Gemini 3.5 Flash Lite price in rubles calculated?

Providers quote their tariffs in US dollars, so the ruble amounts here are a conversion at the official Bank of Russia rate of 84.51 ₽ per dollar. For Gemini 3.5 Flash Lite that gives 25.35 ₽ per 1M input tokens and 211 ₽ per 1M output tokens. When the rate moves, the ruble price moves with it.

How is Gemini 3.5 Flash Lite different from other Gemini models?

Gemini 3.5 Flash Lite belongs to the Gemini Flash Lite family at Google. Models in the Gemini line differ in context size, price per 1M tokens and supported capabilities, so compare them by the numbers on their cards rather than by the name. Every Gemini model is listed on the provider page.

Can Gemini 3.5 Flash Lite reason?

Yes, Gemini 3.5 Flash Lite declares a reasoning mode: it works through intermediate steps before it answers. Those steps add output tokens, so a call with a long reasoning chain costs more and takes longer than a plain completion. Capabilities declared for the model: reasoning, tool calling, image input, structured output.

What capabilities does Gemini 3.5 Flash Lite support?

The declared capabilities of Gemini 3.5 Flash Lite are: reasoning, tool calling, image input, structured output. It accepts text, images, video, audio, PDF as input, and every attachment consumes tokens from the shared context window of 1.0M tokens. The list comes from the Google model card and is refreshed with the catalog as of September 2026.

Other Google models