15% off all models 🎉 Every model at 85% of OpenRouter list price.Browse models →
Maker: Z.ai
GLM 5.3
Z.ai's large-scale GLM reasoning model for complex software engineering and long-horizon agent tasks, with improved coding and token efficiency over GLM 5.2.
- Interface model ID
- glm-5.3
- Context
- 1.049M context
- Max output
- Not published
- API provider
- 1 API provider
- Fast mode
- 0 fast routes
ApiFlux uptimeStatus unknown
Model transit availability
ApiFlux pricing by access route
USD per 1M tokens15% off
- Input
- $1.19
$1.40 - Output
- $3.74
$4.40
Z.AI
Not published
Cache reads are priced; a cache write price was not published.
Input $1.19Output $3.74Cache read $0.221Cache write —1-hour cache write —1.049M Context · Not published max outputProvider statusOfficial status unavailable
Pricing last updated: July 25, 2026
Long-context pricing applies only after the threshold shown. Cache and token prices are listed per 1M tokens.
Suitable for
- Text and chat requests that fit the published context window.
- Workloads that benefit from a unified OpenAI-compatible ApiFlux endpoint.
Unsuitable for
- Native image, audio, or video generation priced with non-token units.
- Workloads that require an unpublished cache or output-limit guarantee.