Home / Z.ai / GLM / GLM 5.3 Flash
GLM 5.3 Flash API price
glm-5.3-flash
The cheapest GLM 5.3 Flash API is $0.07/M input at Nahcrof (inference host) - all 6 providers compared below, per 1M tokens.
FlagshipZ.ai / Zhipu1M context6 sellers
Input
$0.07
per 1M tokens · Nahcrof (inference host)
up to $0.15
Output
$0.22
per 1M tokens · Nahcrof (inference host)
up to $0.50
Where to buy it
Cheapest GLM 5.3 Flash API providers compared - 6 sellers, from $0.07/M.
Cost calculator
| Provider | Input $/M | Output $/M | Cached in | TPS | Your cost |
|---|---|---|---|---|---|
| Nahcrof (inference host)stale, Q8_0 | 0.070cheapest | 0.220 | 0.010 | 16 | |
| OrcaRouter (aggregator) | 0.075 | 0.250 | - | - | |
| InferX (inference host)70% off | 0.090 | 0.300 | 0.0090 | - | |
| RunInfra (inference host) | 0.100 | 0.400 | 0.010 | 254 | |
| Together AI (inference host) | 0.150 | 0.500 | 0.030 | - | |
| TokenRouter (aggregator) | 0.150 | 0.500 | 0.030 | - |
2.1× between the cheapest and the dearest of 6 sellers.
Against its tier
- This model$0.29
- Tier median · Gemini 3.1 Pro$11.90
- Dearest · GPT 5.5 Pro (04-23 snapshot)$210.00
#1 of 38 flagship · cheaper than 100%
Details
- Context window
- 1M1,000,000 tokens
- Developer
- Z.ai / Zhipu
- Intelligence
- 41.9
- Sellers
- 62.1× spread
- Cached input
- $0.01where the source publishes one
- Cheapest at
- Nahcrof (inference host)
- Combined
- $0.29in + out per 1M
- Last checked
- 2026-09-17
- Throughput
- 16 tok/sas published by Nahcrof (inference host)