[ zai ]

GLM-5.3 Flash

Billed at the provider's list price, with no markup. Pay per token in USDF, with no subscription and no minimum. Streaming and the OpenAI-compatible request format are supported.

zai logoZ.aiglm-5.3-flashAvailabletextContext window: 1,048,576 tokens

Input

$0.19

$ / 1M tokens

Output

$0.63

$ / 1M tokens

Cached input

$0.04

$ / 1M tokens

Context window

1,048,576

tokens

Private: $0.19 input / $0.63 output per 1M tokens on attested hardware. Private inference

Model ID
glm-5.3-flash
Modality
text
Unit
USD per 1M tokens
Status
Available
Endpoints
/v1/chat/completions
Context window
1,048,576 tokens
Max output
16,384 tokens
Pricing basis
List price
Routes enabled
1 of 5
Price sheet
2026-10-10.1

Routes

ProviderEnabledUpstream inputUpstream output
deepinfraYes$0.15$0.50
novitaNo$0.15$0.50
fireworksNo$0.15$0.50
redpill (private)No$0.15$0.50
huggingfaceNo$0.15$0.50

Make your first request

Use any OpenAI-compatible SDK. Change the base URL, keep everything else. Responses include a receipt with the exact cost.

No account: send the same request to /x402/v1/chat/completions without credentials and the gateway answers 402 with the quote. Pay per request.

curl https://api.usdf.fi/v1/chat/completions \
  -H "Authorization: Bearer $USDF_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{"model": "glm-5.3-flash", "messages": [{"role": "user", "content": "Hello"}]}'