Z.ai · live price
GLM 5.3 API pricing
GLM-5.3 is a large-scale reasoning model from Z.ai, built for complex software engineering and long-horizon agent tasks. It supports text input and output with a 1M-token context window, and improves...
- Input
- $0.222
- per 1M tokens
- Output
- $3.39
- per 1M tokens
- Cached input
- $0.177
- per 1M tokens
- Context
- 1.0M
- 944K max output
List price from OpenRouter, which passes Z.ai's price through without markup. Refreshed hourly. Released August 18, 2026. Model id z-ai/glm-5.3.
What GLM 5.3 costs in practice
1,000 chat messages
$1.92
1,000 input + 500 output tokens each
1,000 RAG questions
$3.47
8,000 input + 500 output tokens each
100 coding-agent steps
$1.34
30,000 input + 2,000 output tokens each, no caching
GLM 5.3 price history
GLM 5.3 price by host
- Baidu$0.155 in · $0.488 out · 1.0M ctx · 98.4% uptime 24h
- Reka$0.37 in · $1.14 out · 262K ctx · 99.8% uptime 24h
- Morph$0.06 in · $2.69 out · 1.0M ctx · 98.3% uptime 24h
- SiliconFlow$0.7 in · $2.20 out · 1.0M ctx · 100.0% uptime 24h
- DeepInfra$0.563 in · $2.50 out · 1.0M ctx · 99.1% uptime 24h
- Novita$0.783 in · $2.46 out · 1.0M ctx · 100.0% uptime 24h
- Phala$0.84 in · $2.64 out · 1.0M ctx · 99.3% uptime 24h
- Sail Research$0.2 in · $3.40 out · 1.0M ctx · 95.4% uptime 24h
- Wafer$0.222 in · $3.39 out · 1.0M ctx · 99.6% uptime 24h
- DigitalOcean$0.91 in · $2.86 out · 1.0M ctx · 100.0% uptime 24h
- Inceptron$0.6 in · $3.39 out · 1.0M ctx · 99.3% uptime 24h
- GMICloud$0.98 in · $3.08 out · 1.0M ctx · 99.7% uptime 24h
- Relace$0.145 in · $4 out · 1.0M ctx · 99.7% uptime 24h
- AkashML$1.05 in · $3.56 out · 1.0M ctx · 99.9% uptime 24h
- InferenceNet$0.24 in · $4.40 out · 1.0M ctx · 99.3% uptime 24h
- Makora$0.85 in · $3.93 out · 980K ctx · 98.4% uptime 24h
- Alibaba$1.19 in · $3.74 out · 1M ctx · 99.9% uptime 24h
- Decart$1.19 in · $3.74 out · 1.0M ctx · 99.7% uptime 24h
- Friendli$1.26 in · $3.96 out · 1.0M ctx · 100.0% uptime 24h
- Mistral$1.40 in · $4.40 out · 1.0M ctx · 99.9% uptime 24h
- BaseTen$1.40 in · $4.40 out · 1.0M ctx · 99.8% uptime 24h
- Crusoe$1.40 in · $4.40 out · 1.0M ctx · 99.9% uptime 24h
- PrimeIntellect$1.40 in · $4.40 out · 1.0M ctx · 100.0% uptime 24h
- Venice$1.40 in · $4.40 out · 1M ctx · 99.4% uptime 24h
- Together$1.40 in · $4.40 out · 1.0M ctx · 97.2% uptime 24h
- Parasail$1.40 in · $4.40 out · 1.0M ctx · 99.6% uptime 24h
- Modal$1.40 in · $4.40 out · 1.0M ctx · 98.5% uptime 24h
- Fireworks$1.40 in · $4.40 out · 1.0M ctx · 99.1% uptime 24h
- Cloudflare$1.40 in · $4.40 out · 1.0M ctx · 99.5% uptime 24h
- AtlasCloud$1.40 in · $4.40 out · 1.0M ctx · 99.9% uptime 24h
- Z.AI$1.40 in · $4.40 out · 1.0M ctx · 99.8% uptime 24h
Where to buy Z.ai credits for less
Disclosure: this site is run by the team behind smaaart, one of the options compared. Tables are ranked by price only, from live quotes and published fees, including when smaaart isn't first. How we compare
GLM 5.3 cost calculator
Cheaper alternatives to GLM 5.3
| Model | Input / 1M | Output / 1M |
|---|---|---|
| OpenAI: GPT-6 Luna Pro | $0.1 | $0.5 |
| OpenAI: GPT-6 Luna | $0.1 | $0.5 |
| Qwen: Qwen3.8 Omni Flash | $0.15 | $0.47 |
| DeepSeek: DeepSeek V4.1 Flash | $0.03 | $0.5 |
| Qwen: Qwen3.8 Flash | $0.15 | $0.47 |
| DeepSeek: DeepSeek V4 Flash Vision Exp | $0.216 | $0.647 |
Other Z.ai models
| Model | Input / 1M | Output / 1M |
|---|---|---|
| GLM 5.3 Flash | $0.15 | $0.5 |
| GLM 5.3 FlashX | $0.37 | $1.25 |
| GLM 5.3 Prime | $2.80 | $8.80 |
| GLM 5.2 | $1.40 | $4.40 |
| GLM 5.1 | $0.965 | $3.03 |
| GLM 5V Turbo | $1.20 | $4 |