GLM 5.3 Prime
ModelActiveby Z.ai (Zhipu AI) · family “glm-prime”
GLM-5.3-Prime is the high-speed variant of Z.ai's GLM-5.3, inheriting its full capabilities while delivering 1.5–2× the output throughput through inference acceleration. It supports text input and output with a 1M-token... (description from the OpenRouter listing)
in: textout: textReasoningTool useStructured output
Epoch Capabilities Index
Not scored by Epoch AI
Released
23 Sep 2026
Context window
1M tokens
Max output
131K tokens
Input $ / 1M tokens
$2.80
Output $ / 1M tokens
$8.80
Cached input $ / 1M
$0.56
Each benchmark shown separately with its own source
Benchmark results (0)
No benchmark results recorded for this model.
Source: Epoch AI Benchmarking Hub (CC BY 4.0). “External” rows are leaderboard results Epoch collects from third parties. Best reported setting per benchmark is shown.
API list price over time
Price history
$ per 1M tokens
Unchanged since first recorded on 27 Sep 2026 (OpenRouter listing); re-read on every data refresh, most recently 6h ago. Steps show when the price changed.
Events
- Model launchZ.ai releases GLM 5.3 PrimeOpenRouter models API