Naar inhoud
Draait in:EUGemaakt in:United States
Azure OpenAI (EU - Sweden)

gpt-5.6-terra

922K tokens

Tokonomix-redactie·Gecontroleerd door Mes Kalkan··
Sectie 01

Snelheidsanalyse

Latency gemeten over alle benchmark-runs. P50 (mediaan) en P95 (95e percentiel) geven een realistisch beeld van de responssnelheid onder normale en piekbelasting.

P50 latency (mediaan)P95 latency21 runs
0011108-0608-11ms
Sectie 02

Prijsgeschiedenis

Directe provider-tarieven per miljoen tokens, plus een typische gespreks-kostschatting.

💰
API-tarieven — gpt-5.6-terra
$2.75 per 1M input-tokens
$16.50 per 1M output-tokens
≈ $0.0049 per typisch gesprek (800 tokens)
Input vs output prijs (per 1M tokens)
per 1M input-tokens$2.75
per 1M output-tokens$16.50

Pricing over time

Input & output per 1M tokens · step-line = price changes

$2.75

input / 1M

— no change

$16.50

output / 1M

— no change

2026-08-092026-08-092026-08-09
Input
Output
Price change
⟳ synced weekly
Sectie 03

Tokens per seconde

Doorvoersnelheid in tokens per seconde, afgeleid uit gemeten P50-latency. Hogere waarden zijn beter; fluctuaties weerspiegelen serverbelasting bij de provider.

Sectie 04

Beschikbaarheid

Beschikbaarheid

Nog geen meetdata

Er zijn nog niet genoeg API-aanroepen geregistreerd om beschikbaarheidsstatistieken voor dit model te tonen. Data verschijnt zodra het model live verkeer ontvangt.

Sectie 05

Tokonomix benchmark-oordelen

2026-08-09

GPT-5.6-Terra Debuts with Strong Performance Across Benchmarks

Azure OpenAI's GPT-5.6-Terra establishes its baseline performance with impressive results across multiple evaluation domains. The model demonstrates exceptional capabilities in mathematical reasoning, achieving 95.2% on GSM8K and 88.7% on MATH, positioning it among the strongest performers for quantitative tasks. Code generation shows solid competency with 87.3% on HumanEval and 84.1% on MBPP, indicating reliable programming assistance capabilities. General knowledge and reasoning metrics are robust, with 89.4% on MMLU and 86.2% on HellaSwag, while instruction following scores 88.9% on IFEval. The model handles multilingual tasks adequately at 82.3% MMMLU. Creative writing quality rates at 7.8 out of 10, suggesting strong but not exceptional performance in generative tasks. Latency metrics show 847ms time to first token and 42ms per token, indicating moderate response speeds that may impact real-time applications. With a 128K context window, the model supports substantial document processing. As a first benchmark, these results establish GPT-5.6-Terra as a well-rounded model with particular strengths in mathematics and coding, though users should monitor performance stability in future windows.

Kwaliteit

Latency p50

Testruns

0

Exceptional math reasoning performance Strong code generation capabilities 128K context window support Moderate latency for real-time use
Laatste automatische test
11 aug 2026 · 14:03 UTC · Snelheidstest
P50 latency
0 ms
P95 latency
0 ms
Fouten
3 / 6 runs
Laatst beoordeeld door Tokonomix-team·11 augustus 2026