GPT-5.4 Mini

gpt-5.4-mini by OpenAI Released March 18, 2026
ReasoningFunction callingStructured outputPrompt cachingWeb searchComputer useImage inputFile inputAudio inputVideo inputEU hostingFree endpointUp to % off
Input from
$0.75 / 1M
Output from
$4.5 / 1M
Cached input from
$0.075 / 1M
Context window
400K
Endpoints
6
from 3 providers
Regions
Global, Europe, United States

About GPT-5.4 Mini

GPT-5.4 mini brings the core capabilities of GPT-5.4 to a faster, more efficient model optimized for high-throughput workloads.

GPT-5.4 Mini is a chat model by OpenAI, available on Eden AI through 6 endpoints from 3 providers, from $0.75 per million input tokens and $4.5 per million output tokens, with a context window of up to 400K tokens.

Providers and prices

6 endpoints from 3 providers, cheapest first, in USD per million tokens. Every one is called with the same Eden AI key: pin one with its endpoint ID, or let Eden AI route for you. .

Endpoint
Region
Context
Input / 1M
Output / 1M
Cached / 1M
Supports
Microsoft AzureOfficial−%azure/gpt-5.4-mini
Global
272K
$0.75$0.75
$4.5
$0.075
ReasoningToolsJSONCacheVisionWeb
Databricks (Europe)Official−%databricks/databricks-gpt-5-4-mini@eu
Europe
272K
$0.75$0.75
$4.5
$0.075
ReasoningToolsJSONCacheVisionWeb
DatabricksOfficial−%databricks/databricks-gpt-5-4-mini
Global
272K
$0.75$0.75
$4.5
$0.075
ReasoningToolsJSONCacheVisionWeb
OpenAIOfficial−%openai/gpt-5.4-mini
Global
400K
$0.75$0.75
$4.5
$0.075
ReasoningToolsJSONCacheVisionWeb
Microsoft Azure (Europe)Official−%azure/gpt-5.4-mini@eu
Europe
272K
$0.825$0.825
$4.95
$0.0825
ReasoningToolsJSONCacheVisionWeb
Microsoft Azure (United States)Official−%azure/gpt-5.4-mini@us
United States
272K
$0.825$0.825
$4.95
$0.0825
ReasoningToolsJSONCacheVisionWeb

Performance

Measured on real Eden AI traffic over the last 30 days. Endpoints with too little traffic are not shown.

Fastest first token
0.49 s
Databricks (Europe)
Lowest latency
1.18 s
Microsoft Azure (Europe)
Highest throughput
144.43 tok/s
OpenAI
Best availability
100%
Databricks (Europe)
Endpoint
First token
Latency p50
Latency p95
Throughput
Availability
Cache hits
Tool errors
JSON errors
Microsoft Azure
0.63 s
1.18 s
5.54 s
74.31 tok/s
99.94%
54.23%
1.21%
0%
Databricks (Europe)
0.49 s
1.18 s
2.9 s
98.94 tok/s
100%
75.92%
14.29%
0%
Databricks
0.49 s
1.19 s
2.98 s
98.41 tok/s
100%
75.93%
18.92%
0%
OpenAI
0.88 s
2.95 s
19.27 s
144.43 tok/s
100%
5.64%
3.3%
1.08%
Microsoft Azure (Europe)
0.63 s
1.18 s
5.53 s
74.29 tok/s
99.94%
54.32%
1.19%
0%
Microsoft Azure (United States)
s
s
s
tok/s
%
%
%
%

Estimate your monthly cost

Enter your expected volume to compare every endpoint at once.

Call GPT-5.4 Mini with Eden AI

Use the model ID to let Eden AI pick an endpoint, or an endpoint ID from the table above to pin one.

import requests

response = requests.post(
    "https://api.edenai.run/v3/chat/completions",
    headers={"Authorization": "Bearer YOUR_EDENAI_API_KEY"},
    json={
        "model": "",
        "messages": [{"role": "user", "content": "Hello!"}],
    },
)
print(response.json())

Questions about GPT-5.4 Mini

On Eden AI, prices for this model start at $0.75 per million input tokens and $4.5 per million output tokens. Prices differ by endpoint; the providers table lists each one. .

Up to 400K tokens, depending on the endpoint.

6 endpoints from 3 providers, available in: Global, Europe, United States. All of them are called with the same Eden AI API key.

Yes. At least one endpoint is hosted in Europe; pin it with its endpoint ID from the providers table to keep requests in the EU.

Over the last 30 days, the fastest first token was 0.49 s (Databricks (Europe)) and the highest throughput was 144.43 tokens per second (OpenAI).