Claude Haiku 4 5 (2025-10-01)

claude-haiku-4-5-20251001 by Anthropic Released October 16, 2025
ReasoningFunction callingStructured outputPrompt cachingWeb searchComputer useImage inputFile inputAudio inputVideo inputEU hostingFree endpointUp to % off
Input from
$1 / 1M
Output from
$5 / 1M
Cached input from
$0.1 / 1M
Context window
200K
Endpoints
4
from 2 providers
Regions
Global, Europe, United States

About Claude Haiku 4 5 (2025-10-01)

Claude Haiku 4.5 is Anthropic’s fastest and most efficient model, delivering near-frontier intelligence at a fraction of the cost and latency of larger Claude models.

Claude Haiku 4 5 (2025-10-01) is a chat model by Anthropic, available on Eden AI through 4 endpoints from 2 providers, from $1 per million input tokens and $5 per million output tokens, with a context window of up to 200K tokens.

Providers and prices

4 endpoints from 2 providers, cheapest first, in USD per million tokens. Every one is called with the same Eden AI key: pin one with its endpoint ID, or let Eden AI route for you. .

Endpoint
Region
Context
Input / 1M
Output / 1M
Cached / 1M
Supports
Amazon Web ServicesOfficial−%amazon/anthropic.claude-haiku-4-5-20251001-v1:0
Global
200K
$1$1
$5
$0.1
ReasoningToolsJSONCacheVisionWeb
AnthropicOfficial−%anthropic/claude-haiku-4-5-20251001
Global
200K
$1$1
$5
$0.1
ReasoningToolsJSONCacheVisionWeb
Amazon Web Services (Europe)Official−%amazon/anthropic.claude-haiku-4-5-20251001-v1:0@eu
Europe
200K
$1.1$1.1
$5.5
$0.11
ReasoningToolsJSONCacheVisionWeb
Amazon Web Services (United States)Official−%amazon/anthropic.claude-haiku-4-5-20251001-v1:0@us
United States
200K
$1.1$1.1
$5.5
$0.11
ReasoningToolsJSONCacheVisionWeb

Performance

Measured on real Eden AI traffic over the last 30 days. Endpoints with too little traffic are not shown.

Fastest first token
0.81 s
Anthropic
Lowest latency
1.57 s
Amazon Web Services (Europe)
Highest throughput
78.44 tok/s
Amazon Web Services (Europe)
Best availability
99.99%
Anthropic
Endpoint
First token
Latency p50
Latency p95
Throughput
Availability
Cache hits
Tool errors
JSON errors
Amazon Web Services
0.84 s
1.57 s
6.95 s
78.25 tok/s
98.79%
86.85%
0.52%
11.54%
Anthropic
0.81 s
1.65 s
3.61 s
51.11 tok/s
99.99%
1.01%
0%
0%
Amazon Web Services (Europe)
0.84 s
1.57 s
6.97 s
78.44 tok/s
98.78%
87.21%
0.49%
27.27%
Amazon Web Services (United States)
s
s
s
tok/s
%
%
%
%

Estimate your monthly cost

Enter your expected volume to compare every endpoint at once.

Call Claude Haiku 4 5 (2025-10-01) with Eden AI

Use the model ID to let Eden AI pick an endpoint, or an endpoint ID from the table above to pin one.

import requests

response = requests.post(
    "https://api.edenai.run/v3/chat/completions",
    headers={"Authorization": "Bearer YOUR_EDENAI_API_KEY"},
    json={
        "model": "",
        "messages": [{"role": "user", "content": "Hello!"}],
    },
)
print(response.json())

Questions about Claude Haiku 4 5 (2025-10-01)

On Eden AI, prices for this model start at $1 per million input tokens and $5 per million output tokens. Prices differ by endpoint; the providers table lists each one. .

Up to 200K tokens, depending on the endpoint.

4 endpoints from 2 providers, available in: Global, Europe, United States. All of them are called with the same Eden AI API key.

Yes. At least one endpoint is hosted in Europe; pin it with its endpoint ID from the providers table to keep requests in the EU.

Over the last 30 days, the fastest first token was 0.81 s (Anthropic) and the highest throughput was 78.44 tokens per second (Amazon Web Services (Europe)).