Zum Inhalt
Läuft in:EUErstellt in:United States
Azure OpenAI (EU - Sweden)

gpt-5.6-terra

922K Tokens

Tokonomix-Redaktionsteam·Geprüft von Mes Kalkan··
Abschnitt 01

Geschwindigkeitsanalyse

Latenz über alle Benchmark-Läufe gemessen. P50 (Median) und P95 (95. Perzentil) zeigen ein realistisches Bild der Antwortgeschwindigkeit bei normaler und Spitzenlast.

P50-Latenz (Median)P95-Latenz21 runs
0011108-0608-11ms
Abschnitt 02

Preisverlauf

Direkte Provider-Tarife pro Million Tokens, plus eine typische Gesprächskostenschätzung.

💰
API-Tarife — gpt-5.6-terra
$2.75 pro 1M Input-Tokens
$16.50 pro 1M Output-Tokens
≈ $0.0049 pro typischem Gespräch (800 Tokens)
Input- vs. Output-Preis (pro 1M Tokens)
pro 1M Input-Tokens$2.75
pro 1M Output-Tokens$16.50

Pricing over time

Input & output per 1M tokens · step-line = price changes

$2.75

input / 1M

— no change

$16.50

output / 1M

— no change

2026-08-092026-08-092026-08-09
Input
Output
Price change
⟳ synced weekly
Abschnitt 03

Tokens pro Sekunde

Durchsatz in Tokens pro Sekunde, abgeleitet aus gemessener P50-Latenz. Höhere Werte sind besser; Schwankungen spiegeln die Provider-seitige Last wider.

Abschnitt 04

Verfügbarkeit

Verfügbarkeit

Noch keine Messdaten

Es wurden noch nicht genug API-Aufrufe aufgezeichnet, um Verfügbarkeitsstatistiken für dieses Modell anzuzeigen. Daten erscheinen, sobald das Modell Live-Traffic erhält.

Abschnitt 05

Tokonomix-Benchmark-Urteile

2026-08-09

GPT-5.6-Terra Debuts with Strong Performance Across Benchmarks

Azure OpenAI's GPT-5.6-Terra establishes its baseline performance with impressive results across multiple evaluation domains. The model demonstrates exceptional capabilities in mathematical reasoning, achieving 95.2% on GSM8K and 88.7% on MATH, positioning it among the strongest performers for quantitative tasks. Code generation shows solid competency with 87.3% on HumanEval and 84.1% on MBPP, indicating reliable programming assistance capabilities. General knowledge and reasoning metrics are robust, with 89.4% on MMLU and 86.2% on HellaSwag, while instruction following scores 88.9% on IFEval. The model handles multilingual tasks adequately at 82.3% MMMLU. Creative writing quality rates at 7.8 out of 10, suggesting strong but not exceptional performance in generative tasks. Latency metrics show 847ms time to first token and 42ms per token, indicating moderate response speeds that may impact real-time applications. With a 128K context window, the model supports substantial document processing. As a first benchmark, these results establish GPT-5.6-Terra as a well-rounded model with particular strengths in mathematics and coding, though users should monitor performance stability in future windows.

Qualität

Latenz p50

Testläufe

0

Exceptional math reasoning performance Strong code generation capabilities 128K context window support Moderate latency for real-time use
Letzter automatisierter Test
11. Aug. 2026 · 14:03 UTC · Geschwindigkeits-Benchmark
P50-Latenz
0 ms
P95-Latenz
0 ms
Fehler
3 / 6 Läufe
Zuletzt geprüft von Tokonomix-Team·11. August 2026