GPT-6 Astra

gpt-6-astra by OpenAI Released September 5, 2026
ReasoningFunction callingStructured outputPrompt cachingWeb searchComputer useImage inputFile inputAudio inputVideo inputEU hostingFree endpointUp to % off
Input from
$10 / 1M
Output from
$50 / 1M
Cached input from
$1 / 1M
Context window
1.05M
Endpoints
7
from 4 providers
Regions
Global, Europe, United States

About GPT-6 Astra

GPT-6 Astra is OpenAI's flagship model for demanding end-to-end work.

GPT-6 Astra is a chat model by OpenAI, available on Eden AI through 7 endpoints from 4 providers, from $10 per million input tokens and $50 per million output tokens, with a context window of up to 1.05M tokens.

Providers and prices

7 endpoints from 4 providers, cheapest first, in USD per million tokens. Every one is called with the same Eden AI key: pin one with its endpoint ID, or let Eden AI route for you. Above 272K tokens: $20 input / $75 output per 1M.

Endpoint
Region
Context
Input / 1M
Output / 1M
Cached / 1M
Supports
Databricks (Europe)Official−%databricks/databricks-gpt-6-astra@eu
Europe
$10$10
$50
$1
ReasoningToolsJSONCacheVisionWeb
DatabricksOfficial−%databricks/databricks-gpt-6-astra
Global
$10$10
$50
$1
ReasoningToolsJSONCacheVisionWeb
Microsoft AzureOfficial−%azure/gpt-6-astra
Global
922K
$10$10
$50
$1
ReasoningToolsJSONCacheVisionWeb
OpenAIOfficial−%openai/gpt-6-astra
Global
1.05M
$10$10
$50
$1
ReasoningToolsJSONCacheVisionWeb
Amazon Web ServicesOfficial−%amazon/openai.gpt-6-astra
Global
1.05M
$11$11
$55
$1.1
ReasoningToolsJSONCacheVisionWeb
Amazon Web Services (United States)Official−%amazon/openai.gpt-6-astra@us
United States
1.05M
$11$11
$55
$1.1
ReasoningToolsJSONCacheVisionWeb
Microsoft Azure (United States)Official−%azure/gpt-6-astra@us
United States
922K
$11$11
$55
$1.1
ReasoningToolsJSONCacheVisionWeb

Performance

Measured on real Eden AI traffic over the last 30 days. Endpoints with too little traffic are not shown.

Fastest first token
0.6 s
Microsoft Azure
Lowest latency
7.1 s
OpenAI
Highest throughput
44.73 tok/s
Databricks (Europe)
Best availability
99.95%
Databricks (Europe)
Endpoint
First token
Latency p50
Latency p95
Throughput
Availability
Cache hits
Tool errors
JSON errors
Databricks (Europe)
2.99 s
8.47 s
43.17 s
44.73 tok/s
99.95%
53.14%
15.34%
16.67%
Databricks
2.59 s
8.57 s
42.98 s
44.1 tok/s
99.95%
56.94%
15.34%
17.95%
Microsoft Azure
0.6 s
10.35 s
70.88 s
30.34 tok/s
88.63%
90.25%
1.19%
16.67%
OpenAI
2.45 s
7.1 s
44.66 s
32.53 tok/s
99.75%
88.36%
0.22%
0.18%
Amazon Web Services
s
s
s
tok/s
%
%
%
%
Amazon Web Services (United States)
s
s
s
tok/s
%
%
%
%
Microsoft Azure (United States)
s
s
s
tok/s
%
%
%
%

Estimate your monthly cost

Enter your expected volume to compare every endpoint at once.

Call GPT-6 Astra with Eden AI

Use the model ID to let Eden AI pick an endpoint, or an endpoint ID from the table above to pin one.

import requests

response = requests.post(
    "https://api.edenai.run/v3/chat/completions",
    headers={"Authorization": "Bearer YOUR_EDENAI_API_KEY"},
    json={
        "model": "",
        "messages": [{"role": "user", "content": "Hello!"}],
    },
)
print(response.json())

Questions about GPT-6 Astra

On Eden AI, prices for this model start at $10 per million input tokens and $50 per million output tokens. Prices differ by endpoint; the providers table lists each one. Above 272K tokens: $20 input / $75 output per 1M.

Up to 1.05M tokens, depending on the endpoint.

7 endpoints from 4 providers, available in: Global, Europe, United States. All of them are called with the same Eden AI API key.

Yes. At least one endpoint is hosted in Europe; pin it with its endpoint ID from the providers table to keep requests in the EU.

Over the last 30 days, the fastest first token was 0.6 s (Microsoft Azure) and the highest throughput was 44.73 tokens per second (Databricks (Europe)).