Skip to content
Runs in:SGMade in:China
Alibaba Cloud Qwen (DashScope Intl)

Qwen3.7 Max

Tier A — Frontier · 1M tokens

Tokonomix Editorial Team·Reviewed by Mes Kalkan··
Section 01

Speed analysis

Latency measured across all benchmark runs. P50 (median) and P95 (95th percentile) give a realistic picture of response speed under normal and peak load.

P50 latency (median)P95 latency3 runs
1179205529313807468307-3007-31ms
Section 02

Pricing history

Direct provider rates per million tokens, plus a typical-conversation cost estimate.

💰
API rates — Qwen3.7 Max
$1.25 per 1M input tokens
$3.75 per 1M output tokens
≈ $0.0015 per typical conversation (800 tokens)
Input vs output price (per 1M tokens)
per 1M input tokens$1.25
per 1M output tokens$3.75
No pricing history yet — will populate after the first metadata sync detects a price change.
Section 03

Tokens per second

Throughput in tokens per second, derived from measured P50 latency. Higher is better; fluctuations track provider-side load.

Throughput (tokens / s)75 / avg 118
16863

Estimated from P50 latency × 200 output tokens — the absolute number depends on this assumption; the trend is what matters.

Section 04

Availability

Availability

How often this model answers when we call it — measured across real API requests and live tests over the last 30 days. This is separate from quality: these numbers only tell you whether the model responds, not how good the answer is.

Last 7 days

100.0%

n=1

Last 30 days

100.0%

n=1

Median response time

4,171ms

n=1

Based on 10 measurements over the last 30 days.

Technical details

Only live API calls and live-test requests count — internal probes and benchmark runs are excluded.

Calls with a custom API key (BYOK) are excluded: those failures are key-specific, not a sign of model downtime.

Failed calls are NOT included in quality scores — quality is measured on successful responses only. Availability and quality are independent signals.

Median response time (p50) across successful calls with a recorded duration. Outliers (very slow or very fast calls) pull the median less than the average.

Total calls (30d)

1

OK responses (30d)

1

Total calls (7d)

1

OK responses (7d)

1

Section 05

Tokonomix benchmark verdicts

No benchmark verdicts yet for this model.

Last automated test
Jul 31, 2026 · 08:03 UTC · Speed benchmark
P50 latency
2657 ms
P95 latency
2729 ms
Errors
0 / 3 runs
Last reviewed by Tokonomix Team·July 31, 2026