GLM (Z.AI) · Cheap + Fast
GLM 5.3 Flash
glm/glm-5.3-flash — GLM (Z.AI)'s cheap + fast model, accessible through Leanroute's OpenAI-compatible AI Gateway. Pay provider list price with zero per-request markup; flat $15/mo BYOK subscription covers the platform fee.
Pricing snapshot: August 2026. Re-validated monthly against the upstream GLM (Z.AI) pricing page.
Pricing
| Type | USD per 1M tokens |
|---|---|
| Input tokens | $0.150 |
| Cached input (80% off) | $0.030 |
| Output tokens | $0.500 |
Leanroute's markup: $0. You pay the provider list price above, plus a flat $15/mo BYOK subscription (or 5% at top-up for prepaid credits). Compare pricing at /calculator.
Capabilities
Example request
curl https://api.leanroute.dev/v1/chat/completions \
-H "Authorization: Bearer $LEANROUTE_API_KEY" \
-H "Content-Type: application/json" \
-d '{
"model": "glm/glm-5.3-flash",
"messages": [
{ "role": "user", "content": "Hello, GLM 5.3 Flash!" }
]
}'OpenAI-wire-compatible — point any OpenAI SDK at api.leanroute.dev/v1 and set model: "glm/glm-5.3-flash". Full endpoint reference at /docs/api.
Other GLM (Z.AI) models
- GLM 4 Plus$0.690/$0.690 · Flagship
- GLM 4 Flash$0.014/$0.014 · Cheap + Fast
- GLM 5.1$1.40/$4.40 · Flagship
- GLM 5.2$1.40/$4.40 · Flagship
- GLM 5.3$1.40/$4.40 · Flagship
- GLM 5$1.00/$3.20 · Flagship
- GLM 4.7$0.600/$2.20 · Cheap + Fast
- GLM 4.7 Flashx$0.070/$0.400 · Cheap + Fast
- GLM 4.6$0.600/$2.20 · Cheap + Fast
- GLM 4.5$0.600/$2.20 · Cheap + Fast
Same-tier alternatives across providers
- GPT 4o Mini$0.150/$0.600 · OpenAI
- GPT 5 Nano$0.050/$0.400 · OpenAI
- GPT 5 Mini$0.250/$2.00 · OpenAI
- GPT 5.6 Luna$0.200/$1.20 · OpenAI
- GPT 5.3 Codex$1.75/$14.00 · OpenAI
- Claude Haiku 4 5$1.00/$5.00 · Anthropic
- Gemini 2.5 Flash$0.300/$2.50 · Google
- Gemini 2.5 Flash Lite$0.150/$1.25 · Google
Leanroute's router picks the cheapest same-tier model automatically when cheaper-model routing is on. Config at /dashboard/settings.
See also: Full model catalog · API reference · Savings calculator · Leanroute pricing