wellflow.devAI models, for less
wellflow
All models

GLM 5.2

Best value

1,000,000 tokens of context

GLM 5.2 is a large-scale reasoning model from Z.ai. It supports text input and output and is suited for long-horizon agent workflows, project-level software engineering, and complex multi-step automation. Reasoning efforts high and xhigh are supported; xhigh maps to max reasoning. It is particularly strong at coding and tool use across long-running tasks, able to maintain engineering context and follow standards consistently through a full development workflow, from requirements to multi-platform deployment, in a single task

Provider

WellFlow Premium

Provider model

glm-5.2

ReasoningVision

Pricing

per 1M tokens

Input

Wellflow price: $0.08provider price: $1.40

Output

provider price: $4.40Wellflow price: $0.20

Savings

95%

Prompt caching

The repeated part of a request (system prompt, documents, chat history) is stored in the cache. Reading it back costs less than regular input.

Cache read
Wellflow price: $0.04
-50% vs input price
Cache write
Wellflow price: $0.08

Performance

Availability and speed of the model through Wellflow.

Uptime

Speed

Connect Wellflow to your IDE