getcheapaicredits

Z.ai · live price

GLM 4.7 Flash API pricing

As a 30B-class SOTA model, GLM-4.7-Flash offers a new option that balances performance and efficiency. It is further optimized for agentic coding use cases, strengthening coding capabilities, long-horizon task planning,...

Small and cheap (under $1 per 1M)Tool callingReasoning
Input
$0.061
per 1M tokens
Output
$0.4
per 1M tokens
Cached input
–
per 1M tokens
Context
200K
118K max output

List price from OpenRouter, which passes Z.ai's price through without markup. Refreshed hourly. Released January 19, 2026. Model id z-ai/glm-4.7-flash.

What GLM 4.7 Flash costs in practice

At list price, before any batch or caching discount.

1,000 chat messages

$0.26

1,000 input + 500 output tokens each

1,000 RAG questions

$0.68

8,000 input + 500 output tokens each

100 coding-agent steps

$0.26

30,000 input + 2,000 output tokens each, no caching

GLM 4.7 Flash price history

No price change since we started tracking on 2026-10-01. We record prices daily; changes will show here.

GLM 4.7 Flash price by host

Venice serves it below list price right now. Route to it to pay less on every token.

Where to buy Z.ai credits for less

GLM 4.7 Flash costs the same per token everywhere it's sold through OpenRouter. What changes is the price of the credits. Live quotes for $100 of usage:
#Where you buyYou pay
1
OrbioCardCheapest right now
$85.10
incl. $2.60 card processing
2
smaaartCard
$95.00
+ card processing and tax at checkout (not in this number)
3
Direct from the labCard
$100.00
no top-up fee
4
OpenRouterCrypto
$105.00
fees included
5
OpenRouterCard
$105.50
fees included

Disclosure: this site is run by the team behind smaaart, one of the options compared. Tables are ranked by price only, from live quotes and published fees, including when smaaart isn't first. How we compare

GLM 4.7 Flash cost calculator

Enter your monthly tokens to price your own usage.
Usage at list price$7.03/ month(quotes below are for $7.00: purchases run from $5.00 to $10,000.00)

Cheaper alternatives to GLM 4.7 Flash

Other labs' models in the same price range or the one below, cheaper per token.
ModelInput / 1MOutput / 1M
Qwen: Qwen3.7 Flash$0.03$0.13
DeepSeek: DeepSeek V4 Flash 0423$0.042$0.084
Google: Gemma 4 26B A4B $0.09$0.3
Qwen: Qwen3.5-9B$0.1$0.15
Qwen: Qwen3.5-Flash$0.065$0.26
Mistral: Ministral 3 3B 2512$0.1$0.1

Other Z.ai models

ModelInput / 1MOutput / 1M
GLM 5$0.6$1.92
GLM 4.7$0.6$2.20
GLM 4.6V$0.3$0.9
GLM 5 Turbo$1.20$4
GLM 5V Turbo$1.20$4
GLM 5.1$0.965$3.03
All Z.ai API prices

GLM 4.7 Flash pricing FAQ