GPT-5.6 Luna

gpt-5.6-luna by OpenAI Released July 9, 2026
ReasoningFunction callingStructured outputPrompt cachingWeb searchComputer useImage inputFile inputAudio inputVideo inputEU hostingFree endpointUp to % off
Input from
$0.2 / 1M
Output from
$1.2 / 1M
Cached input from
$0.02 / 1M
Context window
1.05M
Endpoints
8
from 4 providers
Regions
Global, Europe, United States

About GPT-5.6 Luna

GPT-5.6 Luna is a fast, cost-efficient model in OpenAI's GPT-5.6 series.

GPT-5.6 Luna is a chat model by OpenAI, available on Eden AI through 8 endpoints from 4 providers, from $0.2 per million input tokens and $1.2 per million output tokens, with a context window of up to 1.05M tokens.

Providers and prices

8 endpoints from 4 providers, cheapest first, in USD per million tokens. Every one is called with the same Eden AI key: pin one with its endpoint ID, or let Eden AI route for you. Above 272K tokens: $0.4 input / $1.8 output per 1M.

Endpoint
Region
Context
Input / 1M
Output / 1M
Cached / 1M
Supports
Microsoft AzureOfficial−%azure/gpt-5.6-luna
Global
922K
$0.2$0.2
$1.2
$0.02
ReasoningToolsJSONCacheVisionWeb
Databricks (Europe)Official−%databricks/databricks-gpt-5-6-luna@eu
Europe
922K
$0.2$0.2
$1.2
$0.02
ReasoningToolsJSONCacheVisionWeb
DatabricksOfficial−%databricks/databricks-gpt-5-6-luna
Global
922K
$0.2$0.2
$1.2
$0.02
ReasoningToolsJSONCacheVisionWeb
OpenAIOfficial−%openai/gpt-5.6-luna
Global
1.05M
$0.2$0.2
$1.2
$0.02
ReasoningToolsJSONCacheVisionWeb
Amazon Web ServicesOfficial−%amazon/openai.gpt-5.6-luna
Global
1M
$0.22$0.22
$1.32
$0.022
ReasoningToolsJSONCacheVisionWeb
Amazon Web Services (United States)Official−%amazon/openai.gpt-5.6-luna@us
United States
1M
$0.22$0.22
$1.32
$0.022
ReasoningToolsJSONCacheVisionWeb
Microsoft Azure (Europe)Official−%azure/gpt-5.6-luna@eu
Europe
922K
$0.22$0.22
$1.32
$0.022
ReasoningToolsJSONCacheVisionWeb
Microsoft Azure (United States)Official−%azure/gpt-5.6-luna@us
United States
922K
$0.22$0.22
$1.32
$0.022
ReasoningToolsJSONCacheVisionWeb

Performance

Measured on real Eden AI traffic over the last 30 days. Endpoints with too little traffic are not shown.

Fastest first token
1.61 s
Microsoft Azure
Lowest latency
1.74 s
Microsoft Azure (Europe)
Highest throughput
79.92 tok/s
Databricks (Europe)
Best availability
100%
Amazon Web Services
Endpoint
First token
Latency p50
Latency p95
Throughput
Availability
Cache hits
Tool errors
JSON errors
Microsoft Azure
1.61 s
1.87 s
14.67 s
34.25 tok/s
99.77%
84.4%
1.98%
1%
Databricks (Europe)
2.5 s
3.55 s
12.46 s
79.92 tok/s
99.99%
51.13%
20.73%
40.23%
Databricks
2.8 s
3.62 s
12.54 s
78.94 tok/s
99.99%
58.28%
19.55%
40.19%
OpenAI
2.05 s
2.4 s
7.8 s
66.09 tok/s
99.99%
72.39%
0.12%
0.75%
Amazon Web Services
4.89 s
6.09 s
12.66 s
19.53 tok/s
100%
93.4%
0.94%
82.5%
Amazon Web Services (United States)
s
s
s
tok/s
%
%
%
%
Microsoft Azure (Europe)
1.71 s
1.74 s
13.32 s
33.18 tok/s
99.78%
82.32%
1.91%
0.86%
Microsoft Azure (United States)
s
s
s
tok/s
%
%
%
%

Estimate your monthly cost

Enter your expected volume to compare every endpoint at once.

Call GPT-5.6 Luna with Eden AI

Use the model ID to let Eden AI pick an endpoint, or an endpoint ID from the table above to pin one.

import requests

response = requests.post(
    "https://api.edenai.run/v3/chat/completions",
    headers={"Authorization": "Bearer YOUR_EDENAI_API_KEY"},
    json={
        "model": "",
        "messages": [{"role": "user", "content": "Hello!"}],
    },
)
print(response.json())

Questions about GPT-5.6 Luna

On Eden AI, prices for this model start at $0.2 per million input tokens and $1.2 per million output tokens. Prices differ by endpoint; the providers table lists each one. Above 272K tokens: $0.4 input / $1.8 output per 1M.

Up to 1.05M tokens, depending on the endpoint.

8 endpoints from 4 providers, available in: Global, Europe, United States. All of them are called with the same Eden AI API key.

Yes. At least one endpoint is hosted in Europe; pin it with its endpoint ID from the providers table to keep requests in the EU.

Over the last 30 days, the fastest first token was 1.61 s (Microsoft Azure) and the highest throughput was 79.92 tokens per second (Databricks (Europe)).