Gemini 3 Flash Preview

gemini-3-flash-preview by Google Cloud Released December 18, 2025
ReasoningFunction callingStructured outputPrompt cachingWeb searchComputer useImage inputFile inputAudio inputVideo inputEU hostingFree endpointUp to % off
Input from
$0.5 / 1M
Output from
$3 / 1M
Cached input from
$0.05 / 1M
Context window
1.05M
Endpoints
2
from 2 providers
Regions
Global

About Gemini 3 Flash Preview

Gemini 3 Flash Preview is a high speed, high value thinking model designed for agentic workflows, multi turn chat, and coding assistance.

Gemini 3 Flash Preview is a chat model by Google Cloud, available on Eden AI through 2 endpoints from 2 providers, from $0.5 per million input tokens and $3 per million output tokens, with a context window of up to 1.05M tokens.

Providers and prices

2 endpoints from 2 providers, cheapest first, in USD per million tokens. Every one is called with the same Eden AI key: pin one with its endpoint ID, or let Eden AI route for you. .

Endpoint
Region
Context
Input / 1M
Output / 1M
Cached / 1M
Supports
Google CloudOfficial−%google/gemini-3-flash-preview
Global
1.05M
$0.5$0.5
$3
$0.05
ReasoningToolsJSONCacheVisionWeb
Google CloudOfficial−%vertex/gemini-3-flash-preview
Global
1.05M
$0.5$0.5
$3
$0.05
ReasoningToolsJSONCacheVisionWeb

Performance

Measured on real Eden AI traffic over the last 30 days. Endpoints with too little traffic are not shown.

Fastest first token
3.71 s
Google Cloud
Lowest latency
8.43 s
Google Cloud
Highest throughput
171.06 tok/s
Google Cloud
Best availability
100%
Google Cloud
Endpoint
First token
Latency p50
Latency p95
Throughput
Availability
Cache hits
Tool errors
JSON errors
Google Cloud
3.71 s
8.43 s
21.84 s
171.06 tok/s
99.7%
10.64%
0.63%
0.23%
Google Cloud
7.56 s
16.44 s
26.62 s
115.15 tok/s
100%
2.16%
0%
0%

Estimate your monthly cost

Enter your expected volume to compare every endpoint at once.

Call Gemini 3 Flash Preview with Eden AI

Use the model ID to let Eden AI pick an endpoint, or an endpoint ID from the table above to pin one.

import requests

response = requests.post(
    "https://api.edenai.run/v3/chat/completions",
    headers={"Authorization": "Bearer YOUR_EDENAI_API_KEY"},
    json={
        "model": "",
        "messages": [{"role": "user", "content": "Hello!"}],
    },
)
print(response.json())

Questions about Gemini 3 Flash Preview

On Eden AI, prices for this model start at $0.5 per million input tokens and $3 per million output tokens. Prices differ by endpoint; the providers table lists each one. .

Up to 1.05M tokens, depending on the endpoint.

2 endpoints from 2 providers, available in: Global. All of them are called with the same Eden AI API key.

Yes. At least one endpoint is hosted in Europe; pin it with its endpoint ID from the providers table to keep requests in the EU.

Over the last 30 days, the fastest first token was 3.71 s (Google Cloud) and the highest throughput was 171.06 tokens per second (Google Cloud).