GLM-5 Turbo

glm-5-turbo by Z.ai Released March 16, 2026
ReasoningFunction callingStructured outputPrompt cachingWeb searchComputer useImage inputFile inputAudio inputVideo inputEU hostingFree endpointUp to % off
Input from
$1.2 / 1M
Output from
$4 / 1M
Cached input from
$0.24 / 1M
Context window
203K
Endpoints
2
from 2 providers
Regions
Asia, Europe

About GLM-5 Turbo

GLM-5 Turbo is a new model from Z.ai designed for fast inference and strong performance in agent-driven environments such as OpenClaw scenarios.

GLM-5 Turbo is a chat model by Z.ai, available on Eden AI through 2 endpoints from 2 providers, from $1.2 per million input tokens and $4 per million output tokens, with a context window of up to 203K tokens.

Providers and prices

2 endpoints from 2 providers, cheapest first, in USD per million tokens. Every one is called with the same Eden AI key: pin one with its endpoint ID, or let Eden AI route for you. .

Endpoint
Region
Context
Input / 1M
Output / 1M
Cached / 1M
Supports
TensorX (Europe)Official−%tensorx/z-ai/glm-5-turbo
Europe
203K
$1.2$1.2
$4
$0.3
ReasoningToolsJSONCacheVisionWeb
Z.ai (Asia)Official−%zai/glm-5-turbo
Asia
203K
$1.2$1.2
$4
$0.24
ReasoningToolsJSONCacheVisionWeb

Performance

Measured on real Eden AI traffic over the last 30 days. Endpoints with too little traffic are not shown.

Fastest first token
7.81 s
TensorX (Europe)
Lowest latency
13.4 s
TensorX (Europe)
Highest throughput
28.99 tok/s
Z.ai (Asia)
Best availability
99.12%
Z.ai (Asia)
Endpoint
First token
Latency p50
Latency p95
Throughput
Availability
Cache hits
Tool errors
JSON errors
TensorX (Europe)
7.81 s
13.4 s
93.63 s
25.75 tok/s
95.53%
81.75%
5.88%
0%
Z.ai (Asia)
9.07 s
39.58 s
115.06 s
28.99 tok/s
99.12%
8.48%
0%
80.46%

Estimate your monthly cost

Enter your expected volume to compare every endpoint at once.

Call GLM-5 Turbo with Eden AI

Use the model ID to let Eden AI pick an endpoint, or an endpoint ID from the table above to pin one.

import requests

response = requests.post(
    "https://api.edenai.run/v3/chat/completions",
    headers={"Authorization": "Bearer YOUR_EDENAI_API_KEY"},
    json={
        "model": "",
        "messages": [{"role": "user", "content": "Hello!"}],
    },
)
print(response.json())

Questions about GLM-5 Turbo

On Eden AI, prices for this model start at $1.2 per million input tokens and $4 per million output tokens. Prices differ by endpoint; the providers table lists each one. .

Up to 203K tokens, depending on the endpoint.

2 endpoints from 2 providers, available in: Asia, Europe. All of them are called with the same Eden AI API key.

Yes. At least one endpoint is hosted in Europe; pin it with its endpoint ID from the providers table to keep requests in the EU.

Over the last 30 days, the fastest first token was 7.81 s (TensorX (Europe)) and the highest throughput was 28.99 tokens per second (Z.ai (Asia)).