GLM 5.3 Flash
Z.AI · 1,310,720 context · Updated 3h ago
At a glance
GLM 5.3 Flash's lowest observed input price is $0.06 per 1M tokens at Nous Research. Its lowest observed output price is $0.20 per 1M tokens at Nous Research. The model supports a 1,310,720-token context window. 4 venues are currently tracked.
Based on the latest successful observations through Aug 27, 2026, 8:32 PM UTC. Sources: OpenRouter, Vercel AI Gateway, Cloudflare AI Gateway, Nous Research. How prices are compared.
Compare venues
4 observed offersWhat prices include
Best available
Nous Research
realtime · Nous Research public API
Input / 1M
$0.06
Output / 1M
$0.20
Vs next venue
20% lower
| Venue | Cache read / 1M | Action | ||
|---|---|---|---|---|
Nous ResearchBest price Nous Research public API · realtime | Input / 1M$0.06 | Output / 1M$0.20↓20% vs next | Cache read / 1M$0.012 | |
OpenRouter OpenRouter public API · realtime | Input / 1M$0.075 | Output / 1M$0.25 | Cache read / 1M$0.015 | |
Vercel AI Gateway Vercel AI Gateway public API · realtime | Input / 1M$0.15 | Output / 1M$0.50 | Cache read / 1M$0.03 | |
Cloudflare AI Gateway Cloudflare-hosted Workers AI · realtime | Input / 1M$0.15 | Output / 1M$0.50 | Cache read / 1M- |
Price history
Collecting hourly data
6/24 observations
Specifications
Model ID
z-ai/glm-5.3-flash
Developer
Z.AI
Context window
1,310,720 tokens
Frequently asked questions
How much does GLM 5.3 Flash cost?
The lowest observed GLM 5.3 Flash price is $0.06/1M input and $0.2/1M output at Nous Research. Prices are USD per 1 million tokens.
Which venue is cheapest for GLM 5.3 Flash?
Nous Research currently has the lowest observed prices for GLM 5.3 Flash: $0.2/1M output versus $0.5/1M at Cloudflare AI Gateway.