Qwen3.7 Flash (2026-07-15)

qwen3.7-flash-2026-07-15 by Qwen Released July 28, 2026
ReasoningFunction callingStructured outputPrompt cachingWeb searchComputer useImage inputFile inputAudio inputVideo inputEU hostingFree endpointUp to 35% off
Input from
$0.0195 / 1M
Output from
$0.0845 / 1M
Cached input from
$0.0039 / 1M
Context window
1M
Endpoints
1
from 1 providers
Regions
Global

About Qwen3.7 Flash (2026-07-15)

The Qwen3.7 native vision-language Flash model series delivers a comprehensive upgrade over 3.6-Flash in multimodal understanding and agent execution. This model particularly excels in enhanced multimodal foundations with stronger universal object recognition, further improved real-world perception and spatial intelligence, significantly upgraded multimodal agent capabilities for Search Agent and CI Agent scenarios with more stable end-to-end task execution, as well as optimized multimodal coding for a smoother vibe coding experience. This version is a snapshot from July 15, 2026.

Qwen3.7 Flash (2026-07-15) is a chat model by Qwen, available on Eden AI through 1 endpoint from 1 provider, from $0.019 per million input tokens and $0.085 per million output tokens, with a context window of up to 1M tokens.

Providers and prices

1 endpoints from 1 providers, cheapest first, in USD per million tokens. Every one is called with the same Eden AI key: pin one with its endpoint ID, or let Eden AI route for you. .

Endpoint
Region
Context
Input / 1M
Output / 1M
Cached / 1M
Supports
QwenOfficial−35%qwen/qwen3.7-flash-2026-07-15
Global
1M
$0.03$0.0195
$0.0845
$0.0039
ReasoningToolsJSONCacheVisionWeb

Performance

Measured on real Eden AI traffic over the last 30 days. Endpoints with too little traffic are not shown.

Fastest first token
1.74 s
Qwen
Lowest latency
3.47 s
Qwen
Highest throughput
38.09 tok/s
Qwen
Best availability
99.93%
Qwen
Endpoint
First token
Latency p50
Latency p95
Throughput
Availability
Cache hits
Tool errors
JSON errors
Qwen
1.74 s
3.47 s
26.11 s
38.09 tok/s
99.93%
64.78%
0%
4.17%

Estimate your monthly cost

Enter your expected volume to compare every endpoint at once.

Call Qwen3.7 Flash (2026-07-15) with Eden AI

Use the model ID to let Eden AI pick an endpoint, or an endpoint ID from the table above to pin one.

import requests

response = requests.post(
    "https://api.edenai.run/v3/chat/completions",
    headers={"Authorization": "Bearer YOUR_EDENAI_API_KEY"},
    json={
        "model": "",
        "messages": [{"role": "user", "content": "Hello!"}],
    },
)
print(response.json())

Questions about Qwen3.7 Flash (2026-07-15)

On Eden AI, prices for this model start at $0.0195 per million input tokens and $0.0845 per million output tokens. Prices differ by endpoint; the providers table lists each one. .

Up to 1M tokens, depending on the endpoint.

1 endpoints from 1 providers, available in: Global. All of them are called with the same Eden AI API key.

Yes. At least one endpoint is hosted in Europe; pin it with its endpoint ID from the providers table to keep requests in the EU.

Over the last 30 days, the fastest first token was 1.74 s (Qwen) and the highest throughput was 38.09 tokens per second (Qwen).