GPT-4.1 Mini

gpt-4.1-mini by OpenAI Released April 15, 2025
ReasoningFunction callingStructured outputPrompt cachingWeb searchComputer useImage inputFile inputAudio inputVideo inputEU hostingFree endpointUp to % off
Input from
$0.4 / 1M
Output from
$1.6 / 1M
Cached input from
$0.1 / 1M
Context window
1.05M
Endpoints
4
from 2 providers
Regions
Global, Europe, United States

About GPT-4.1 Mini

GPT-4.1 Mini is a mid-sized model delivering performance competitive with GPT-4o at substantially lower latency and cost.

GPT-4.1 Mini is a chat model by OpenAI, available on Eden AI through 4 endpoints from 2 providers, from $0.4 per million input tokens and $1.6 per million output tokens, with a context window of up to 1.05M tokens.

Providers and prices

4 endpoints from 2 providers, cheapest first, in USD per million tokens. Every one is called with the same Eden AI key: pin one with its endpoint ID, or let Eden AI route for you. .

Endpoint
Region
Context
Input / 1M
Output / 1M
Cached / 1M
Supports
Microsoft AzureOfficial−%azure/gpt-4.1-mini
Global
1.05M
$0.4$0.4
$1.6
$0.1
ReasoningToolsJSONCacheVisionWeb
OpenAIOfficial−%openai/gpt-4.1-mini
Global
1.05M
$0.4$0.4
$1.6
$0.1
ReasoningToolsJSONCacheVisionWeb
Microsoft Azure (Europe)Official−%azure/gpt-4.1-mini@eu
Europe
1.05M
$0.44$0.44
$1.76
$0.11
ReasoningToolsJSONCacheVisionWeb
Microsoft Azure (United States)Official−%azure/gpt-4.1-mini@us
United States
1.05M
$0.44$0.44
$1.76
$0.11
ReasoningToolsJSONCacheVisionWeb

Performance

Measured on real Eden AI traffic over the last 30 days. Endpoints with too little traffic are not shown.

Fastest first token
0.54 s
OpenAI
Lowest latency
0.6 s
OpenAI
Highest throughput
31.27 tok/s
Microsoft Azure (United States)
Best availability
100%
Microsoft Azure (Europe)
Endpoint
First token
Latency p50
Latency p95
Throughput
Availability
Cache hits
Tool errors
JSON errors
Microsoft Azure
0.99 s
1.17 s
9.2 s
28.19 tok/s
100%
38.57%
0%
0.64%
OpenAI
0.54 s
0.6 s
1.2 s
16 tok/s
100%
36.38%
0.09%
0.39%
Microsoft Azure (Europe)
0.82 s
1.03 s
3.69 s
23.36 tok/s
100%
24.82%
0%
0.64%
Microsoft Azure (United States)
1.46 s
2.75 s
17.9 s
31.27 tok/s
100%
82.71%
0%
0%

Estimate your monthly cost

Enter your expected volume to compare every endpoint at once.

Call GPT-4.1 Mini with Eden AI

Use the model ID to let Eden AI pick an endpoint, or an endpoint ID from the table above to pin one.

import requests

response = requests.post(
    "https://api.edenai.run/v3/chat/completions",
    headers={"Authorization": "Bearer YOUR_EDENAI_API_KEY"},
    json={
        "model": "",
        "messages": [{"role": "user", "content": "Hello!"}],
    },
)
print(response.json())

Questions about GPT-4.1 Mini

On Eden AI, prices for this model start at $0.4 per million input tokens and $1.6 per million output tokens. Prices differ by endpoint; the providers table lists each one. .

Up to 1.05M tokens, depending on the endpoint.

4 endpoints from 2 providers, available in: Global, Europe, United States. All of them are called with the same Eden AI API key.

Yes. At least one endpoint is hosted in Europe; pin it with its endpoint ID from the providers table to keep requests in the EU.

Over the last 30 days, the fastest first token was 0.54 s (OpenAI) and the highest throughput was 31.27 tokens per second (Microsoft Azure (United States)).