SylphxModels

Pricing

You pay for tokens. The prices are published.

Every model in the catalog has a list price per million tokens for input, cached input, and output. The same numbers appear on the model page, in the rate sheet below, and on the usage your console reports — so a bill is never a surprise.

Models listed278 providers
Cheapest input$0.0378qwen/qwen3.7-flash
Largest context1.05Mtokens per request
MeteringPer tokenon every Responses call

Estimate a monthly bill

Pick a model and a token mix. The estimate uses the same list prices as the rate sheet — no rounding tricks, no hidden platform fee in the number.

2,000,000 tokens

500,000 tokens

Estimated spend

$1.90

Input side

$0.95

Output side

$0.95

Estimate only: it uses today’s list prices for the selected model and assumes every token is billed at the input or output rate. Cached input is cheaper than the input rate, and failed requests that return an error status are not billed as completions.

How usage is metered

  • One meter on Responses. Each response reports the input and output tokens it used, and that count is what the price applies to.
  • Cached input is discounted. When a request reuses a prompt prefix, the cached tokens are billed at the cached-input rate shown for that model.
  • Retries do not double-bill. Send an Idempotency-Key and a retried create replays the original response instead of charging again.
  • Read your own numbers. The console shows requests, tokens, and settled spend for the workspace, broken down by model.
  • These are our own prices. We sell each model at the rate on this page, and the meter bills the same numbers — one rate sheet, no per-model markup table to decode and no fee added at checkout.

Charged against your organization’s keys at https://api.models.sylphx.ai/v1. Keys are prefixed sk-sx- and can be revoked at any time.

The rate sheet

USD per million tokens, per model. Cached input is what you pay when a prefix is reused; the percentage under it is the saving against the input rate. 488.8K requests were metered in the last seven days across this catalog.

27 of 27 models
Sylphx Models list prices per million tokens for every listed model
ModelInput / MCached input / MOutput / MContextEst. mix / M
DeepSeek: DeepSeek V4 Flash 0731deepseek/deepseek-v4-flash$0.5544$0.088284%$1.6631.0M~$0.8316
DeepSeek: DeepSeek V4 Prodeepseek/deepseek-v4-pro$1.663$0.166390%$4.99128K~$2.495
DeepSeek: DeepSeek V4.1 Flashdeepseek/deepseek-v4.1-flashnew$0.4725$0.050489%$1.891.0M~$0.8269
deepseek/deepseek-flashdeepseek/deepseek-flashnew$0.18$0.01890%$0.721.0M~$0.315
deepseek/deepseek-v4-flash-vision-expdeepseek/deepseek-v4-flash-vision-exp$0.5544$0.0554490%$1.6631.0M~$0.8316
Google: Gemma 4 26B A4B google/gemma-4-26b-a4b-it$0.1764$0.06364%$0.504128K~$0.2583
Google: Gemma 4 31Bgoogle/gemma-4-31b-it$0.189$0.075660%$0.504128K~$0.2677
google/gemini-3.7-flashgoogle/gemini-3.7-flash$0.75$0.07590%$3.751.0M~$1.50
google/gemini-3.8-flashgoogle/gemini-3.8-flashnew$0.9$0.0990%$4.501.0M~$1.80
Meta: Muse Spark 1.2meta/muse-spark-1.2$1.25$0.12590%$4.251.05M~$2.00
Meta: Muse Spark 1.3meta/muse-spark-1.3new$1.25$0.12590%$4.251.05M~$2.00
Meta: Muse Spark 1.3 Contributormeta/muse-spark-1.3-contributor$0.126$0.012690%$0.2521.05M~$0.1575
meta/muse-glimmer-30bmeta/muse-glimmer-30b$0.441$0.050489%$1.89128K~$0.8033
MiniMax: MiniMax M3minimax/minimax-m3$0.378$0.075680%$1.512200K~$0.6615
OpenAI: GPT-5.5openai/gpt-5.5$5.00$0.590%$30.00400K~$11.25
OpenAI: GPT-5.6 Lunaopenai/gpt-5.6-luna$0.504$0.050490%$3.024400K~$1.134
OpenAI: GPT-5.6 Luna Proopenai/gpt-5.6-luna-pro$0.504$0.050490%$3.024400K~$1.134
OpenAI: GPT-5.6 Solopenai/gpt-5.6-sol$4.00$0.490%$20.00400K~$8.00
OpenAI: GPT-5.6 Terraopenai/gpt-5.6-terra$2.00$0.290%$12.00400K~$4.50
Qwen: Qwen3.6 35B A3Bqwen/qwen3.6-35b-a3b$0.248$0.06375%$1.4851.0M~$0.5573
Qwen: Qwen3.7 Flashqwen/qwen3.7-flash$0.0378$0.0075680%$0.16381.0M~$0.0693
Qwen: Qwen3.7 Plusqwen/qwen3.7-plus$0.5$0.0806484%$3.001.0M~$1.125
qwen/qwen3.8-27bqwen/qwen3.8-27b$0.567$0.06389%$4.0321.0M~$1.433
qwen/qwen3.8-flashqwen/qwen3.8-flashnew$0.189$0.0201689%$0.59221.0M~$0.2898
x-ai/grok-4.6x-ai/grok-4.6$2.00$0.290%$6.00131K~$3.00
z-ai/glm-5.3z-ai/glm-5.3$1.40$0.1490%$4.401.0M~$2.15
z-ai/glm-5.3-flashz-ai/glm-5.3-flashnew$0.2268$0.06372%$0.7561.0M~$0.3591

“Est. mix / M” is a 3:1 input:output estimate for a million tokens, shown to compare models. Your bill uses the exact input, cached-input, and output counts on each response.

Billing questions

The short answers. For the request-and-response details, start at the quickstart.

What exactly am I charged for?
Tokens processed on Responses calls: input tokens at the input rate, cached prefix tokens at the cached-input rate, and generated tokens at the output rate for the model you named.
Is there a subscription or a minimum?
Pricing on this API is usage-based: you are charged for the tokens you use, at the prices above. There is no per-model access fee listed here, and adding a key costs nothing.
Do failed requests cost anything?
A request that fails validation or returns an error status does not produce a completed response to bill. Retries are safe when you send an Idempotency-Key: a replayed create returns the original response instead of charging a second time.
How do prices change?
The catalog is the rate sheet: a price change publishes on the model page, the pricing page, and the API’s own model payload at the same time. Historic usage keeps the price that applied when it was metered.
Where can I see what I have spent?
The console shows settled spend, request counts, and token totals for your workspace, with a per-model breakdown so you can see which model is driving the bill.
Start with one keySend a first request

Lowest listed input price: Qwen: Qwen3.7 Flash at $0.0378 per million tokens.

Pricing · Sylphx Models