SylphxModels

Google

Gemini Flash family · v3.8

Gemini 3.8 Flash

New
google/gemini-3.8-flash

google/gemini-3.8-flash on Sylphx Models · 1.0M of context. One Responses contract, with list prices published per million tokens.

Released
Sep 2, 2026 · 19 days old
Provider
Google
Context
1,000,000 tokens
Open in the playgroundSee a request sampleRead the docs
Input $0.9 / MOutput $4.50 / MCached input $0.09 / MFull rate sheet
Context window1,000,000tokens per request
Max outputNot publishedtokens per response
Input$0.9 / Mper million tokens
Output$4.50 / Mper million tokens
Requests · 7d569metered on Responses
Avg TPM · 7d113.71.1M tokens

Pricing

USD per million tokens
Input
$0.9
Output
$4.50
Cached input
$0.09

$1.80 for 750,000 input and 250,000 output tokens.

Assumption: one batch of 750,000 input plus 250,000 output tokens at the published list prices — recompute for your own mix. Blended $1.80 per million at a 3:1 mix.

Chat replies · 3:1 input:output
$1.80 / M
Long documents · 9:1 input:output
$1.26 / M
Heavy output · 1:1 input:output
$2.70 / M
Cache write
Not published
Currency and unit
USD / million tokens

Cached input reads cost 90% less than fresh input tokens.

Full rate sheet and billing rules.

Specification

Context window
1,000,000 tokens
Max output
Not published
Input modalities
Not published
Output modalities
Not published
Family
Gemini Flash
Version
v3.8
Provider
Google
Released
Sep 2, 2026 · 19 days old
Model id
google/gemini-3.8-flash

Data posture

Standard

Prompts and outputs sent to this model are not used to train models.

Training on what you send
No
Zero Data Retention
Available

Traffic

Daily traffic the platform metered on Responses for this exact model id. Totals are the window they name; a day with no traffic is zero, not missing.

Requests · last 7 days
569
Tokens · last 7 days
1.1M
Average TPM · last 7 days
113.7
Series total · 14d
27,899 req · 9.4B tok

14-day total 27,899 · peak 20,531 on Sep 9

Sep 8Sep 21
Daily numbers for the last 14 days
Daily requests and tokens metered for google/gemini-3.8-flash
DayRequestsTokens
2026-09-0885052,769,336
2026-09-0920,5317,614,215,284
2026-09-105,6391,730,890,139
2026-09-1130455,542
2026-09-1200
2026-09-132495
2026-09-146471,039
2026-09-1526176,166
2026-09-1635245,727
2026-09-1717117,216
2026-09-1873467,793
2026-09-1920103,502
2026-09-20634,802
2026-09-2113
curl https://api.models.sylphx.ai/v1/responses \
  -H "Authorization: Bearer $SYLPHX_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{
    "model": "google/gemini-3.8-flash",
    "input": "Summarise this incident report in three bullets.",
    "stream": false
  }'

How to call it

Send the official Responses document to the base URL above with your organization key and this exact model id. Streaming arrives as ordered events with one terminal event.

  • Retries: send an Idempotency-Key and a retry replays the original response instead of billing twice.
  • Errors: typed envelopes tell you whether to retry, wait, or fix the request.
  • Keys: mint and revoke organization keys from the console.

Independent evaluation

We do not publish benchmark scores for this model — we have not verified any. These external leaderboards run their own tests, so check them against your own workload.

Not our numbers
569 requests metered on google/gemini-3.8-flash over the last seven daysEvery Google model →