[ zai ]
GLM-4.6
Billed at the provider's list price, with no markup. Pay per token in USDF, with no subscription and no minimum. Streaming and the OpenAI-compatible request format are supported.
Input
$0.50
$ / 1M tokens
Output
$2.00
$ / 1M tokens
Cached input
$0.10
$ / 1M tokens
Context window
202,752
tokens
- Model ID
- glm-4.6
- Modality
- text
- Unit
- USD per 1M tokens
- Status
- Available
- Endpoints
- /v1/chat/completions
- Context window
- 202,752 tokens
- Max output
- 131,072 tokens
- Pricing basis
- List price
- Routes enabled
- 1 of 1
- Price sheet
- 2026-10-10.1
Routes
| Provider | Enabled | Upstream input | Upstream output |
|---|---|---|---|
| deepinfra | Yes | $0.50 | $2.00 |
Make your first request
Use any OpenAI-compatible SDK. Change the base URL, keep everything else. Responses include a receipt with the exact cost.
No account: send the same request to /x402/v1/chat/completions without credentials and the gateway answers 402 with the quote. Pay per request.
curl https://api.usdf.fi/v1/chat/completions \
-H "Authorization: Bearer $USDF_API_KEY" \
-H "Content-Type: application/json" \
-d '{"model": "glm-4.6", "messages": [{"role": "user", "content": "Hello"}]}'