Skip to content
Tier B — Production
Runs in:SGMade in:China
Alibaba Cloud Qwen (DashScope Intl)

Qwen3.7 Plus

Tier B — Production · 1M tokens

Tokonomix Editorial Team·Reviewed by Mes Kalkan··
Section 01

Speed analysis

Latency measured across all benchmark runs. P50 (median) and P95 (95th percentile) give a realistic picture of response speed under normal and peak load.

P50 latency (median)P95 latency3 runs
1265139015151639176407-3007-31ms
Section 02

Pricing history

Direct provider rates per million tokens, plus a typical-conversation cost estimate.

💰
API rates — Qwen3.7 Plus
$0.3200 per 1M input tokens
$1.28 per 1M output tokens
≈ $0.0004 per typical conversation (800 tokens)
Input vs output price (per 1M tokens)
per 1M input tokens$0.3200
per 1M output tokens$1.28
No pricing history yet — will populate after the first metadata sync detects a price change.
Section 03

Tokens per second

Throughput in tokens per second, derived from measured P50 latency. Higher is better; fluctuations track provider-side load.

Throughput (tokens / s)142 / avg 134
157104

Estimated from P50 latency × 200 output tokens — the absolute number depends on this assumption; the trend is what matters.

Section 04

Availability

Availability

How often this model answers when we call it — measured across real API requests and live tests over the last 30 days. This is separate from quality: these numbers only tell you whether the model responds, not how good the answer is.

Last 7 days

100.0%

n=1

Last 30 days

100.0%

n=1

Median response time

4,988ms

n=1

Based on 7 measurements over the last 30 days.

Technical details

Only live API calls and live-test requests count — internal probes and benchmark runs are excluded.

Calls with a custom API key (BYOK) are excluded: those failures are key-specific, not a sign of model downtime.

Failed calls are NOT included in quality scores — quality is measured on successful responses only. Availability and quality are independent signals.

Median response time (p50) across successful calls with a recorded duration. Outliers (very slow or very fast calls) pull the median less than the average.

Total calls (30d)

1

OK responses (30d)

1

Total calls (7d)

1

OK responses (7d)

1

Section 05

Tokonomix benchmark verdicts

No benchmark verdicts yet for this model.

Last automated test
Jul 31, 2026 · 08:03 UTC · Speed benchmark
P50 latency
1406 ms
P95 latency
1416 ms
Errors
0 / 3 runs
Last reviewed by Tokonomix Team·July 31, 2026