GLM API pricing per 1M tokens

GLM API pricing starts at $0.07 per 1M input tokens with GLM-4.7-FlashX. The top tier is GLM-5V-Turbo. Ruble figures use the CBR rate for Sep 18, 2026.

Prices in rubles as of 18.09.2026 at the CBR rate: 1 $ = 84.51 ₽.

15modelsin the catalog
5,92 ₽from / 1M input$0.07
1Mmax contexttokens
Docsdocumentation

GLM API price per 1M tokens

Currency
Sort
ModelContextInput / 1MOutput / 1MCapabilities
GLM-5.3-Flash
2026-08-26
1M6,34 ₽$0.07521,13 ₽$0.25
reasoningtoolsvisionstructuredopen weights
GLM-5.3
2026-08-14
1M118 ₽$1.4372 ₽$4.4
reasoningtoolsstructuredopen weights
GLM-5.2
2026-06-13
1M118 ₽$1.4372 ₽$4.4
reasoningtoolsstructured
GLM-5V-Turbo
2026-04-01
200K423 ₽$51 859 ₽$22
reasoningtoolsvision
GLM-5.1
2026-03-27
200K118 ₽$1.4372 ₽$4.4
reasoningtoolsstructured
GLM-5
2026-02-11
205K84,51 ₽$1270 ₽$3.2
reasoningtoolsopen weights
GLM-4.7-Flash
2026-01-19
200K0 ₽$00 ₽$0
reasoningtoolsopen weights
GLM-4.7-FlashX
2026-01-19
200K5,92 ₽$0.0733,8 ₽$0.4
reasoningtoolsopen weights
GLM-4.7
2025-12-22
205K50,71 ₽$0.6186 ₽$2.2
reasoningtoolsopen weights
GLM-4.6V
2025-12-08
128K25,35 ₽$0.376,06 ₽$0.9
reasoningtoolsvisionopen weights
GLM-4.6
2025-09-30
205K50,71 ₽$0.6186 ₽$2.2
reasoningtoolsopen weights
GLM-4.5V
2025-08-11
64K50,71 ₽$0.6152 ₽$1.8
reasoningtoolsvisionopen weights
GLM-4.5
2025-07-28
131K50,71 ₽$0.6186 ₽$2.2
reasoningtoolsopen weights
GLM-4.5-Air
2025-07-28
131K16,9 ₽$0.292,96 ₽$1.1
reasoningtoolsopen weights
GLM-4.5-Flash
2025-07-28
131K0 ₽$00 ₽$0
reasoningtoolsopen weights

Models shown: 15 / 15

About the Zhipu AI API

GLM is the model line-up of the Chinese company Zhipu AI (Z.ai). The catalog lists 13 models: the main line from GLM-4.5 to GLM-5.2, the lighter GLM-4.5-Air, and the GLM Flash models listed at zero token price; everything up to and including GLM-5 ships with open weights, while GLM-5.1, GLM-5.2 and GLM-5V-Turbo are API-only.

GLM publishes API prices per million tokens, billed separately for input and output, so the cost of a request depends on prompt and completion length. The table above lists 15 models with per-1M pricing, context window and maximum output.

Frequently asked questions

How much does the GLM API cost?

There is no single rate — it depends on the model. The lowest input price across GLM models is 5.92 ₽ per 1M tokens, converted from the provider's dollar price list, and the top-tier model is GLM-5V-Turbo. The table above lists input and output rates for every Zhipu AI model. Data as of September 2026.

Which GLM model is the cheapest?

By input price it is GLM-4.7-FlashX at 5.92 ₽ per 1M input tokens. The final bill also depends on the output rate, which varies far more between models: the longer the answers, the bigger the gap. Compare the Input / 1M and Output / 1M columns before you commit to a model.

How many GLM models are in the catalog?

Zhipu AI models in the catalog: 15, as of September 2026. The list is built automatically from the provider's public data — new models appear at the next sync, while retired ones stay in place with a deprecated status. Every model has its own page with prices and limits.

How are GLM prices converted to rubles?

Zhipu AI publishes its price list in dollars per 1M tokens, and the ruble figures are a conversion at the official Bank of Russia rate, currently 1 $ = 84.51 ₽. The rate and the date it was taken are shown in the page header, and the currency switch brings back the original dollar prices.

How do the GLM model lines differ?

The lines differ in size, price and purpose: there are lightweight models starting at 5.92 ₽ per 1M input tokens, and top-tier ones such as GLM-5V-Turbo. The table columns show exactly what changes — context window, input and output prices, capabilities. Details live on each model page.

How often are GLM prices updated?

The catalog syncs every day at 06:00 Moscow time: provider price lists, limits, the model list and the Bank of Russia rate are all re-read. If Zhipu AI changes a rate, the new price shows up here after the next sync. Current figures are as of September 2026.

Do GLM models support reasoning?

Yes, some GLM models support a reasoning mode — the Capabilities column flags it in the table, and each model page lists it separately. Keep in mind that reasoning consumes output tokens, so a request costs more even when the per-1M rate stays the same.

Does GLM have open-weight models?

Yes, some GLM models ship with published weights — the model page states this together with the license. You can run them on your own infrastructure, where the cost depends on your hardware rather than on a price list. Catalog prices refer to the provider's own API.

How much do 1,000 GLM tokens cost?

Providers quote prices per 1M tokens, so 1,000 tokens cost exactly a thousandth of that. The cheapest option in the catalog is GLM-4.7-FlashX at $0.07 per 1M input tokens.