All models
GLM 5.3
Popular1,000,000 tokens of context
GLM-5.3 is a large-scale reasoning model from Z.ai, built for complex software engineering and long-horizon agent tasks. It supports text input and output with a 1M-token context window, and improves on GLM-5.2 in coding and in the balance between performance and token efficiency
Provider
WellFlow Premium
Provider model
glm-5.3
Pricing
per 1M tokens
Input
Wellflow price: $0.08provider price: $1.40
Output
provider price: $4.40Wellflow price: $0.20
Savings
95%
Prompt caching
The repeated part of a request (system prompt, documents, chat history) is stored in the cache. Reading it back costs less than regular input.
- Cache read
- Wellflow price: $0.04
- -50% vs input price
- Cache write
- Wellflow price: $0.08
Performance
Availability and speed of the model through Wellflow.
Uptime
OperationalPartial outageOutage