GLM 5 Turbo with Jev
GLM 5 Turbo is a model from Z.ai, released 2026-03-15. It costs $1.20 per million input tokens and $4.00 per million output tokens, reads up to 203K tokens of context, and writes up to 131,072 tokens in one reply. Through llm11, set model to z-ai/glm-5-turbo.
- Input price
- $1.20 / 1M
- Output price
- $4.00 / 1M
- Context window
- 203K tokens
- Max output
- 131K tokens
- Cached input
- $0.240 / 1M
- Released
- 2026-03-15
- Accepts
- text
Read from the provider's prompt cache
Read from the live catalogue and refreshed daily. At blended list price, 28% of the 332 priced models we can call cost more.
Using GLM 5 Turbo with Jev
Jev is TypeSafe AI's System One model. It reads each request first and decides which model in your pool answers, with a calibrated confidence on that call. GLM 5 Turbo does the answering when Jev sends it work, or when you name it yourself. At list price GLM 5 Turbo falls in the llm11-balanced band, but no pack routes to it because Z.ai is not one of the labs the packs draw from. You can still pin it: pass z-ai/glm-5-turbo as model and routing steps aside while verification still runs.
Jev is new, so we have not published results for GLM 5 Turbo with and without it, and this page does not invent any. Every response comes with a receipt that names the model that answered and what it cost against your most expensive model, so you can measure the pairing on your own traffic.
What a request costs
List price arithmetic, the same sum a receipt does. llm11 adds nothing per request; the only fee is on buying credits.
| Request | Tokens in / out | One | A thousand |
|---|---|---|---|
| Short chat turn | 500 / 200 | $0.00140 | $1.40 |
| Question over retrieved documents | 4,000 / 500 | $0.00680 | $6.80 |
| Long document summary | 30,000 / 1,000 | $0.040 | $40.00 |
Cheaper models to try
The catalogue has no benchmark score for GLM 5 Turbo, so this is a price ordering. It does not claim these answer as well.
| Model | Input / 1M | Output / 1M | Context | Released |
|---|---|---|---|---|
| OpenAI GPT-5.4 Mini | $0.750 | $4.50 | 400K | |
| SpaceXAI Grok 4.3 | $1.25 | $2.50 | 1M | |
| SpaceXAI Grok 4.20 | $1.25 | $2.50 | 2M | |
| SpaceXAI Grok 4.20 Multi-Agent | $1.25 | $2.50 | 2M |
Call it
Same OpenAI request shape. Naming the model pins it, so nothing is routed, and the answer is still checked.
python
from openai import OpenAI
client = OpenAI(base_url="https://www.llm11.com/v1", api_key="llm11_live_...")
res = client.chat.completions.create(
model="z-ai/glm-5-turbo",
messages=[{"role": "user", "content": "Hello"}],
)More from Z.ai
| Model | Input / 1M | Output / 1M | Context | Released |
|---|---|---|---|---|
| GLM 5.3 Prime | $2.80 | $8.80 | 1M | |
| GLM 5.3 FlashX | $0.370 | $1.25 | 1.05M | |
| GLM 5.3 Flash | $0.150 | $0.500 | 1.31M | |
| GLM 5.3 | $0.190 | $4.00 | 1.31M | |
| GLM 5.2 | $0.234 | $4.40 | 1.05M | |
| GLM 5.1 | $0.965 | $3.03 | 205K |
Questions
- What is GLM 5 Turbo with Jev?
- Jev is TypeSafe AI's System One model. It reads each request first and decides which model in your pool answers, with a calibrated confidence on that call. GLM 5 Turbo does the answering when Jev sends it work, or when you name it yourself. At list price GLM 5 Turbo falls in the llm11-balanced band, but no pack routes to it because Z.ai is not one of the labs the packs draw from. You can still pin it: pass z-ai/glm-5-turbo as model and routing steps aside while verification still runs. Routing decisions are made by Jev, TypeSafe AI's System One model.
- Is GLM 5 Turbo a System One model?
- No. GLM 5 Turbo answers requests. The System One model is Jev, which decides which model answers each one, so the two do different jobs and llm11 uses both.
- How much does GLM 5 Turbo cost?
- $1.20 per million input tokens and $4.00 per million output tokens. A 4,000 token prompt with a 500 token answer costs about $0.00680, so a thousand of them is about $6.80. llm11 passes provider list price through and charges 5% when you buy credits.
- What is the GLM 5 Turbo context window?
- 202,752 tokens, with up to 131,072 tokens in a single reply.
- Does GLM 5 Turbo support tool calling and structured output?
- The catalogue lists tool calling, structured output and reasoning controls.
- Can I use GLM 5 Turbo through llm11?
- Yes. Set model to z-ai/glm-5-turbo on the OpenAI-compatible endpoint and the request goes to it directly, with verification still running.
- Does Jev route to GLM 5 Turbo?
- Routing decisions are made by Jev, TypeSafe AI's System One model. At list price GLM 5 Turbo falls in the llm11-balanced band, but no pack routes to it because Z.ai is not one of the labs the packs draw from. You can still pin it: pass z-ai/glm-5-turbo as model and routing steps aside while verification still runs.
- What is a cheaper alternative to GLM 5 Turbo?
- Nearest in price and below it: OpenAI GPT-5.4 Mini at $0.750 in and $4.50 out, SpaceXAI Grok 4.3 at $1.25 in and $2.50 out and SpaceXAI Grok 4.20 at $1.25 in and $2.50 out. The catalogue has no benchmark score for GLM 5 Turbo, so this is a price ordering. It does not claim these answer as well.