GPT-6 Luna

gpt-6-luna by OpenAI Released September 23, 2026
ReasoningFunction callingStructured outputPrompt cachingWeb searchComputer useImage inputFile inputAudio inputVideo inputEU hostingFree endpointUp to % off
Input from
$0.1 / 1M
Output from
$0.5 / 1M
Cached input from
$0.01 / 1M
Context window
1.05M
Endpoints
8
from 4 providers
Regions
Global, Europe, United States

About GPT-6 Luna

GPT-6 Luna is the fast, cost-efficient model in OpenAI's GPT-6 series, positioned below GPT-6 Sol.

GPT-6 Luna is a chat model by OpenAI, available on Eden AI through 8 endpoints from 4 providers, from $0.1 per million input tokens and $0.5 per million output tokens, with a context window of up to 1.05M tokens.

Providers and prices

8 endpoints from 4 providers, cheapest first, in USD per million tokens. Every one is called with the same Eden AI key: pin one with its endpoint ID, or let Eden AI route for you. Above 272K tokens: $0.2 input / $0.75 output per 1M.

Endpoint
Region
Context
Input / 1M
Output / 1M
Cached / 1M
Supports
Databricks (Europe)Official−%databricks/databricks-gpt-6-luna@eu
Europe
$0.1$0.1
$0.5
$0.01
ReasoningToolsJSONCacheVisionWeb
DatabricksOfficial−%databricks/databricks-gpt-6-luna
Global
$0.1$0.1
$0.5
$0.01
ReasoningToolsJSONCacheVisionWeb
Microsoft AzureOfficial−%azure/gpt-6-luna
Global
922K
$0.1$0.1
$0.5
$0.01
ReasoningToolsJSONCacheVisionWeb
Amazon Web ServicesOfficial−%amazon/openai.gpt-6-luna
Global
1.05M
$0.1$0.1
$0.5
$0.01
ReasoningToolsJSONCacheVisionWeb
Amazon Web Services (United States)Official−%amazon/openai.gpt-6-luna@us
United States
1.05M
$0.1$0.1
$0.5
$0.01
ReasoningToolsJSONCacheVisionWeb
OpenAIOfficial−%openai/gpt-6-luna
Global
1.05M
$0.1$0.1
$0.5
$0.01
ReasoningToolsJSONCacheVisionWeb
Microsoft Azure (United States)Official−%azure/gpt-6-luna@us
United States
922K
$0.11$0.11
$0.55
$0.011
ReasoningToolsJSONCacheVisionWeb
Microsoft Azure (Europe)Official−%azure/gpt-6-luna@eu
Europe
922K
$0.12$0.12
$0.6
$0.012
ReasoningToolsJSONCacheVisionWeb

Performance

Measured on real Eden AI traffic over the last 30 days. Endpoints with too little traffic are not shown.

Fastest first token
1.62 s
Microsoft Azure (Europe)
Lowest latency
1.86 s
Microsoft Azure (Europe)
Highest throughput
106.75 tok/s
Microsoft Azure
Best availability
100%
Databricks (Europe)
Endpoint
First token
Latency p50
Latency p95
Throughput
Availability
Cache hits
Tool errors
JSON errors
Databricks (Europe)
2.64 s
4.42 s
18.17 s
88.55 tok/s
100%
80.02%
5.48%
7.74%
Databricks
2.66 s
4.26 s
16.84 s
87.19 tok/s
99.99%
78.52%
5.44%
6.15%
Microsoft Azure
1.66 s
1.86 s
7.61 s
106.75 tok/s
99.99%
75.62%
16.73%
0.74%
Amazon Web Services
3.43 s
2.84 s
22.45 s
51.58 tok/s
99.23%
53.09%
1.61%
0.79%
Amazon Web Services (United States)
s
s
s
tok/s
%
%
%
%
OpenAI
1.8 s
5.15 s
21.15 s
81.17 tok/s
99.88%
82.28%
0.39%
2.17%
Microsoft Azure (United States)
s
s
s
tok/s
%
%
%
%
Microsoft Azure (Europe)
1.62 s
1.86 s
7.49 s
106.59 tok/s
100%
75.25%
17.14%
0.74%

Estimate your monthly cost

Enter your expected volume to compare every endpoint at once.

Call GPT-6 Luna with Eden AI

Use the model ID to let Eden AI pick an endpoint, or an endpoint ID from the table above to pin one.

import requests

response = requests.post(
    "https://api.edenai.run/v3/chat/completions",
    headers={"Authorization": "Bearer YOUR_EDENAI_API_KEY"},
    json={
        "model": "",
        "messages": [{"role": "user", "content": "Hello!"}],
    },
)
print(response.json())

Questions about GPT-6 Luna

On Eden AI, prices for this model start at $0.1 per million input tokens and $0.5 per million output tokens. Prices differ by endpoint; the providers table lists each one. Above 272K tokens: $0.2 input / $0.75 output per 1M.

Up to 1.05M tokens, depending on the endpoint.

8 endpoints from 4 providers, available in: Global, Europe, United States. All of them are called with the same Eden AI API key.

Yes. At least one endpoint is hosted in Europe; pin it with its endpoint ID from the providers table to keep requests in the EU.

Over the last 30 days, the fastest first token was 1.62 s (Microsoft Azure (Europe)) and the highest throughput was 106.75 tokens per second (Microsoft Azure).