GLM-5.3 Fast

glm-5.3-fast by Z.ai Released September 15, 2026
ReasoningFunction callingStructured outputPrompt cachingWeb searchComputer useImage inputFile inputAudio inputVideo inputEU hostingFree endpointUp to % off
Input from
$2.1 / 1M
Output from
$6.6 / 1M
Cached input from
$0.39 / 1M
Context window
1.05M
Endpoints
1
from 1 providers
Regions
United States

About GLM-5.3 Fast

GLM-5.3 Fast is a chat model by Z.ai, available on Eden AI through 1 endpoint from 1 provider, from $2.1 per million input tokens and $6.6 per million output tokens, with a context window of up to 1.05M tokens.

Providers and prices

1 endpoints from 1 providers, cheapest first, in USD per million tokens. Every one is called with the same Eden AI key: pin one with its endpoint ID, or let Eden AI route for you. .

Endpoint
Region
Context
Input / 1M
Output / 1M
Cached / 1M
Supports
Fireworks AI (United States)Official−%fireworks_ai/accounts/fireworks/routers/glm-5p3-fast
United States
1.05M
$2.1$2.1
$6.6
$0.39
ReasoningToolsJSONCacheVisionWeb

Performance

Measured on real Eden AI traffic over the last 30 days. Endpoints with too little traffic are not shown.

Fastest first token
0.67 s
Fireworks AI (United States)
Lowest latency
6.95 s
Fireworks AI (United States)
Highest throughput
64.03 tok/s
Fireworks AI (United States)
Best availability
100%
Fireworks AI (United States)
Endpoint
First token
Latency p50
Latency p95
Throughput
Availability
Cache hits
Tool errors
JSON errors
Fireworks AI (United States)
0.67 s
6.95 s
34.81 s
64.03 tok/s
100%
55.51%
0%
9.09%

Estimate your monthly cost

Enter your expected volume to compare every endpoint at once.

Call GLM-5.3 Fast with Eden AI

Use the model ID to let Eden AI pick an endpoint, or an endpoint ID from the table above to pin one.

import requests

response = requests.post(
    "https://api.edenai.run/v3/chat/completions",
    headers={"Authorization": "Bearer YOUR_EDENAI_API_KEY"},
    json={
        "model": "",
        "messages": [{"role": "user", "content": "Hello!"}],
    },
)
print(response.json())

Questions about GLM-5.3 Fast

On Eden AI, prices for this model start at $2.1 per million input tokens and $6.6 per million output tokens. Prices differ by endpoint; the providers table lists each one. .

Up to 1.05M tokens, depending on the endpoint.

1 endpoints from 1 providers, available in: United States. All of them are called with the same Eden AI API key.

Yes. At least one endpoint is hosted in Europe; pin it with its endpoint ID from the providers table to keep requests in the EU.

Over the last 30 days, the fastest first token was 0.67 s (Fireworks AI (United States)) and the highest throughput was 64.03 tokens per second (Fireworks AI (United States)).