Qwen3 Max (2026-01-23)

qwen3-max-2026-01-23 by Qwen Released September 24, 2025
ReasoningFunction callingStructured outputPrompt cachingWeb searchComputer useImage inputFile inputAudio inputVideo inputEU hostingFree endpointUp to 35% off
Input from
$0.78 / 1M
Output from
$3.9 / 1M
Cached input from
$ / 1M
Context window
262K
Endpoints
2
from 1 providers
Regions
Global, Europe

About Qwen3 Max (2026-01-23)

Compared with the snapshot as of September 23, 2025, the Qwen-3 series Max model in this release achieves an effective integration of thinking and non-thinking modes, resulting in a comprehensive and substantial improvement in the model’s overall performance. In thinking mode, the model simultaneously supports web search, web information extraction, and a code interpreter tool, enabling it to tackle more complex and challenging problems with greater accuracy by leveraging external tools while engaging in slow, deliberative reasoning. This version is based on a snapshot taken on January 23, 2026.

Qwen3 Max (2026-01-23) is a chat model by Qwen, available on Eden AI through 2 endpoints from 1 provider, from $0.78 per million input tokens and $3.9 per million output tokens, with a context window of up to 262K tokens.

Providers and prices

2 endpoints from 1 providers, cheapest first, in USD per million tokens. Every one is called with the same Eden AI key: pin one with its endpoint ID, or let Eden AI route for you. .

Endpoint
Region
Context
Input / 1M
Output / 1M
Cached / 1M
Supports
Qwen (Europe)Official−35%qwen/qwen3-max-2026-01-23@eu
Europe
262K
$1.2$0.78
$3.9
$
ReasoningToolsJSONCacheVisionWeb
QwenOfficial−35%qwen/qwen3-max-2026-01-23
Global
262K
$1.2$0.78
$3.9
$
ReasoningToolsJSONCacheVisionWeb

Performance

Measured on real Eden AI traffic over the last 30 days. Endpoints with too little traffic are not shown.

Fastest first token
s
Lowest latency
0.76 s
Qwen (Europe)
Highest throughput
17.91 tok/s
Qwen (Europe)
Best availability
100%
Qwen (Europe)
Endpoint
First token
Latency p50
Latency p95
Throughput
Availability
Cache hits
Tool errors
JSON errors
Qwen (Europe)
s
0.76 s
12.31 s
17.91 tok/s
100%
2.3%
0%
0%
Qwen
s
0.76 s
12.31 s
17.89 tok/s
100%
4.48%
0%
0%

Estimate your monthly cost

Enter your expected volume to compare every endpoint at once.

Call Qwen3 Max (2026-01-23) with Eden AI

Use the model ID to let Eden AI pick an endpoint, or an endpoint ID from the table above to pin one.

import requests

response = requests.post(
    "https://api.edenai.run/v3/chat/completions",
    headers={"Authorization": "Bearer YOUR_EDENAI_API_KEY"},
    json={
        "model": "",
        "messages": [{"role": "user", "content": "Hello!"}],
    },
)
print(response.json())

Questions about Qwen3 Max (2026-01-23)

On Eden AI, prices for this model start at $0.78 per million input tokens and $3.9 per million output tokens. Prices differ by endpoint; the providers table lists each one. .

Up to 262K tokens, depending on the endpoint.

2 endpoints from 1 providers, available in: Global, Europe. All of them are called with the same Eden AI API key.

Yes. At least one endpoint is hosted in Europe; pin it with its endpoint ID from the providers table to keep requests in the EU.

Over the last 30 days, the fastest first token was s () and the highest throughput was 17.91 tokens per second (Qwen (Europe)).