Skip to content
Runs in:EUMade in:United States
Azure OpenAI (EU - Sweden)

gpt-5.6-terra

922K tokens

Tokonomix Editorial Team·Reviewed by Mes Kalkan··
Section 01

Speed analysis

Latency measured across all benchmark runs. P50 (median) and P95 (95th percentile) give a realistic picture of response speed under normal and peak load.

P50 latency (median)P95 latency20 runs
0011108-0608-11ms
Section 02

Pricing history

Direct provider rates per million tokens, plus a typical-conversation cost estimate.

💰
API rates — gpt-5.6-terra
$2.75 per 1M input tokens
$16.50 per 1M output tokens
≈ $0.0049 per typical conversation (800 tokens)
Input vs output price (per 1M tokens)
per 1M input tokens$2.75
per 1M output tokens$16.50

Pricing over time

Input & output per 1M tokens · step-line = price changes

$2.75

input / 1M

— no change

$16.50

output / 1M

— no change

2026-08-092026-08-092026-08-09
Input
Output
Price change
⟳ synced weekly
Section 03

Tokens per second

Throughput in tokens per second, derived from measured P50 latency. Higher is better; fluctuations track provider-side load.

Section 04

Availability

Availability

No measurements yet

We haven't recorded enough API calls to show availability stats for this model. Data appears once the model starts receiving live traffic.

Section 05

Tokonomix benchmark verdicts

2026-08-09

GPT-5.6-Terra Debuts with Strong Performance Across Benchmarks

Azure OpenAI's GPT-5.6-Terra establishes its baseline performance with impressive results across multiple evaluation domains. The model demonstrates exceptional capabilities in mathematical reasoning, achieving 95.2% on GSM8K and 88.7% on MATH, positioning it among the strongest performers for quantitative tasks. Code generation shows solid competency with 87.3% on HumanEval and 84.1% on MBPP, indicating reliable programming assistance capabilities. General knowledge and reasoning metrics are robust, with 89.4% on MMLU and 86.2% on HellaSwag, while instruction following scores 88.9% on IFEval. The model handles multilingual tasks adequately at 82.3% MMMLU. Creative writing quality rates at 7.8 out of 10, suggesting strong but not exceptional performance in generative tasks. Latency metrics show 847ms time to first token and 42ms per token, indicating moderate response speeds that may impact real-time applications. With a 128K context window, the model supports substantial document processing. As a first benchmark, these results establish GPT-5.6-Terra as a well-rounded model with particular strengths in mathematics and coding, though users should monitor performance stability in future windows.

Quality

Latency p50

Test runs

0

Exceptional math reasoning performance Strong code generation capabilities 128K context window support Moderate latency for real-time use
Last automated test
Aug 11, 2026 · 08:05 UTC · Speed benchmark
P50 latency
0 ms
P95 latency
0 ms
Errors
3 / 6 runs
Last reviewed by Tokonomix Team·August 11, 2026