Speed analysis
Latency measured across all benchmark runs. P50 (median) and P95 (95th percentile) give a realistic picture of response speed under normal and peak load.
Pricing history
Direct provider rates per million tokens, plus a typical-conversation cost estimate.
Tokens per second
Throughput in tokens per second, derived from measured P50 latency. Higher is better; fluctuations track provider-side load.
Estimated from P50 latency × 200 output tokens — the absolute number depends on this assumption; the trend is what matters.
Availability
Availability
How often this model answers when we call it — measured across real API requests and live tests over the last 30 days. This is separate from quality: these numbers only tell you whether the model responds, not how good the answer is.
Last 7 days
100.0%
n=1
Last 30 days
100.0%
n=1
Median response time
4,171ms
n=1
Based on 10 measurements over the last 30 days.
Technical details
Only live API calls and live-test requests count — internal probes and benchmark runs are excluded.
Calls with a custom API key (BYOK) are excluded: those failures are key-specific, not a sign of model downtime.
Failed calls are NOT included in quality scores — quality is measured on successful responses only. Availability and quality are independent signals.
Median response time (p50) across successful calls with a recorded duration. Outliers (very slow or very fast calls) pull the median less than the average.
Total calls (30d)
1
OK responses (30d)
1
Total calls (7d)
1
OK responses (7d)
1
Tokonomix benchmark verdicts
No benchmark verdicts yet for this model.
Qwen3.7 Max
by Alibaba Cloud Qwen (DashScope Intl)
- Context window
- 1M tokens
- Input price
- $1.25 / 1M
- Output price
- $3.75 / 1M
- Tier
- Tier A — Frontier
- Modality
- Text
- API type
- REST · streaming
- Benchmark runs
- 3
More from Alibaba Cloud Qwen (DashScope Intl)