SylphxModels

Z.ai

GLM Flashx family · v5.3

GLM 5.3 Flashx

NewNewest in catalog
z-ai/glm-5.3-flashx

z-ai/glm-5.3-flashx on Sylphx Models · 1.0M of context. One Responses contract, with list prices published per million tokens.

Released
Sep 18, 2026 · 4 days old
Provider
z-ai
Context
1,000,000 tokens
Open in the playgroundSee a request sampleRead the docs
Input $0.444 / MOutput $1.50 / MCached input $0.09 / MFull rate sheet
Context window1,000,000tokens per request
Max outputNot publishedtokens per response
Input$0.444 / Mper million tokens
Output$1.50 / Mper million tokens
Requests · 7d3metered on Responses
Avg TPM · 7d<1106 tokens

Pricing

USD per million tokens
Input
$0.444
Output
$1.50
Cached input
$0.09

$0.71 for 750,000 input and 250,000 output tokens.

Assumption: one batch of 750,000 input plus 250,000 output tokens at the published list prices — recompute for your own mix. Blended $0.708 per million at a 3:1 mix.

Chat replies · 3:1 input:output
$0.708 / M
Long documents · 9:1 input:output
$0.5496 / M
Heavy output · 1:1 input:output
$0.972 / M
Cache write
Not published
Currency and unit
USD / million tokens

Cached input reads cost 80% less than fresh input tokens.

Full rate sheet and billing rules.

Specification

Context window
1,000,000 tokens
Max output
Not published
Input modalities
Not published
Output modalities
Not published
Family
GLM Flashx
Version
v5.3
Provider
z-ai
Released
Sep 18, 2026 · 4 days old
Model id
z-ai/glm-5.3-flashx

Data posture

Standard

Prompts and outputs sent to this model are not used to train models.

Training on what you send
No
Zero Data Retention
Available

Traffic

Daily traffic the platform metered on Responses for this exact model id. Totals are the window they name; a day with no traffic is zero, not missing.

Requests · last 7 days
3
Tokens · last 7 days
106
Average TPM · last 7 days
<1
Series total · 14d
3 req · 106 tok

14-day total 3 · peak 3 on Sep 21

Sep 8Sep 21
Daily numbers for the last 14 days
Daily requests and tokens metered for z-ai/glm-5.3-flashx
DayRequestsTokens
2026-09-0800
2026-09-0900
2026-09-1000
2026-09-1100
2026-09-1200
2026-09-1300
2026-09-1400
2026-09-1500
2026-09-1600
2026-09-1700
2026-09-1800
2026-09-1900
2026-09-2000
2026-09-213106
curl https://api.models.sylphx.ai/v1/responses \
  -H "Authorization: Bearer $SYLPHX_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{
    "model": "z-ai/glm-5.3-flashx",
    "input": "Summarise this incident report in three bullets.",
    "stream": false
  }'

How to call it

Send the official Responses document to the base URL above with your organization key and this exact model id. Streaming arrives as ordered events with one terminal event.

  • Retries: send an Idempotency-Key and a retry replays the original response instead of billing twice.
  • Errors: typed envelopes tell you whether to retry, wait, or fix the request.
  • Keys: mint and revoke organization keys from the console.

Independent evaluation

We do not publish benchmark scores for this model — we have not verified any. These external leaderboards run their own tests, so check them against your own workload.

Not our numbers
3 requests metered on z-ai/glm-5.3-flashx over the last seven daysEvery Z.ai model →