GLM 5.3 Flash

Z.AI · 1,310,720 context · Updated 3h ago

At a glance

GLM 5.3 Flash's lowest observed input price is $0.06 per 1M tokens at Nous Research. Its lowest observed output price is $0.20 per 1M tokens at Nous Research. The model supports a 1,310,720-token context window. 4 venues are currently tracked.

Based on the latest successful observations through Aug 27, 2026, 8:32 PM UTC. Sources: OpenRouter, Vercel AI Gateway, Cloudflare AI Gateway, Nous Research. How prices are compared.

Compare venues

4 observed offersWhat prices include

Best available

Nous Research

realtime · Nous Research public API

Input / 1M

$0.06

Output / 1M

$0.20

Vs next venue

20% lower

Nous ResearchBest price
Nous Research public API · realtime
Input / 1M$0.06
Output / 1M$0.20
Cache read / 1M$0.012
OpenRouter
OpenRouter public API · realtime
Input / 1M$0.075
Output / 1M$0.25
Cache read / 1M$0.015
Vercel AI Gateway
Vercel AI Gateway public API · realtime
Input / 1M$0.15
Output / 1M$0.50
Cache read / 1M$0.03
Cloudflare AI Gateway
Cloudflare-hosted Workers AI · realtime
Input / 1M$0.15
Output / 1M$0.50
Cache read / 1M-

Price history

Collecting hourly data

6/24 observations

Specifications

Model ID

z-ai/glm-5.3-flash

Developer

Z.AI

Context window

1,310,720 tokens

Frequently asked questions

How much does GLM 5.3 Flash cost?

The lowest observed GLM 5.3 Flash price is $0.06/1M input and $0.2/1M output at Nous Research. Prices are USD per 1 million tokens.

Which venue is cheapest for GLM 5.3 Flash?

Nous Research currently has the lowest observed prices for GLM 5.3 Flash: $0.2/1M output versus $0.5/1M at Cloudflare AI Gateway.