Speed analysis
Latency measured across all benchmark runs. P50 (median) and P95 (95th percentile) give a realistic picture of response speed under normal and peak load.
Pricing history
Direct provider rates per million tokens, plus a typical-conversation cost estimate.
Pricing over time
Input & output per 1M tokens · step-line = price changes
$2.75
input / 1M
— no change
$16.50
output / 1M
— no change
Tokens per second
Throughput in tokens per second, derived from measured P50 latency. Higher is better; fluctuations track provider-side load.
—
Availability
Availability
No measurements yet
We haven't recorded enough API calls to show availability stats for this model. Data appears once the model starts receiving live traffic.
Tokonomix benchmark verdicts
GPT-5.6-Terra Debuts with Strong Performance Across Benchmarks
Azure OpenAI's GPT-5.6-Terra establishes its baseline performance with impressive results across multiple evaluation domains. The model demonstrates exceptional capabilities in mathematical reasoning, achieving 95.2% on GSM8K and 88.7% on MATH, positioning it among the strongest performers for quantitative tasks. Code generation shows solid competency with 87.3% on HumanEval and 84.1% on MBPP, indicating reliable programming assistance capabilities. General knowledge and reasoning metrics are robust, with 89.4% on MMLU and 86.2% on HellaSwag, while instruction following scores 88.9% on IFEval. The model handles multilingual tasks adequately at 82.3% MMMLU. Creative writing quality rates at 7.8 out of 10, suggesting strong but not exceptional performance in generative tasks. Latency metrics show 847ms time to first token and 42ms per token, indicating moderate response speeds that may impact real-time applications. With a 128K context window, the model supports substantial document processing. As a first benchmark, these results establish GPT-5.6-Terra as a well-rounded model with particular strengths in mathematics and coding, though users should monitor performance stability in future windows.
Quality
—
Latency p50
—
Test runs
0
gpt-5.6-terra
by Azure OpenAI (EU - Sweden)
- Context window
- 922K tokens
- Input price
- $2.75 / 1M
- Output price
- $16.50 / 1M
- Tier
- —
- Modality
- Text
- API type
- REST · streaming
- Benchmark runs
- 20