wellflow.devAI models, for less
wellflow
All models

GLM 5.3

Popular

1,000,000 tokens of context

GLM-5.3 is a large-scale reasoning model from Z.ai, built for complex software engineering and long-horizon agent tasks. It supports text input and output with a 1M-token context window, and improves on GLM-5.2 in coding and in the balance between performance and token efficiency

Provider

WellFlow Premium

Provider model

glm-5.3

ReasoningLong context

Pricing

per 1M tokens

Input

Wellflow price: $0.08provider price: $1.40

Output

provider price: $4.40Wellflow price: $0.20

Savings

95%

Prompt caching

The repeated part of a request (system prompt, documents, chat history) is stored in the cache. Reading it back costs less than regular input.

Cache read
Wellflow price: $0.04
-50% vs input price
Cache write
Wellflow price: $0.08

Performance

Availability and speed of the model through Wellflow.

Uptime

Speed

Connect Wellflow to your IDE