Skip to content
Tier A — Frontier
Runs in:USMade in:United States
Google Gemini

Gemini 3.8 Flash

Tier A — Frontier · 1.048576M tokens

Tokonomix Editorial Team·Reviewed by Mes Kalkan··
Section 01

Pricing

What you pay per million tokens when you use this model on Tokonomix, plus an estimate for a typical conversation.

💰
API rates — Gemini 3.8 Flash
$1.13 per 1M input tokens
$5.61 per 1M output tokens
≈ $0.0018 per typical conversation (800 tokens)
Input vs output price (per 1M tokens)
per 1M input tokens$1.13
per 1M output tokens$5.61
Section 02

Capabilities

toolssource: seed-newest-generation-2026-09-30visionjson modepdf inputreasoningaudio inputjson schemaparallel toolsprompt cachingoutputTokenLimit: 65536max output tokens: 65535
Section 03

Availability

Availability

How often this model answers when we call it — measured across real API requests and live tests over the last 30 days. This is separate from quality: these numbers only tell you whether the model responds, not how good the answer is.

Last 7 days

100.0%

n=2

Last 30 days

100.0%

n=2

Median response time

1,103ms

n=2

Based on 5 measurements over the last 30 days.

Technical details

Only live API calls and live-test requests count — internal probes and benchmark runs are excluded.

Calls with a custom API key (BYOK) are excluded: those failures are key-specific, not a sign of model downtime.

Failed calls are NOT included in quality scores — quality is measured on successful responses only. Availability and quality are independent signals.

Median response time (p50) across successful calls with a recorded duration. Outliers (very slow or very fast calls) pull the median less than the average.

Total calls (30d)

2

OK responses (30d)

2

Total calls (7d)

2

OK responses (7d)

2

Section 04

Tokonomix benchmark verdicts

No benchmark verdicts yet for this model.

Last automated test
Sep 30, 2026 · 08:02 UTC · Speed benchmark
P50 latency
848 ms
P95 latency
3312 ms
Errors
0 / 1 runs
Last reviewed by Tokonomix Team·September 30, 2026