[ meta ]
Llama 3.1 8B Instruct
Billed at the provider's list price, with no markup. Pay per token in USDF, with no subscription and no minimum. Streaming and the OpenAI-compatible request format are supported.
Input
$0.02
$ / 1M tokens
Output
$0.04
$ / 1M tokens
Cached input
not offered
Context window
131,072
tokens
- Model ID
- llama-3.1-8b-instruct
- Modality
- text
- Unit
- USD per 1M tokens
- Status
- Available
- Endpoints
- /v1/chat/completions, /v1/completions
- Context window
- 131,072 tokens
- Max output
- 16,384 tokens
- Pricing basis
- List price
- Routes enabled
- 1 of 1
- Price sheet
- 2026-10-10.1
Routes
| Provider | Enabled | Upstream input | Upstream output |
|---|---|---|---|
| deepinfra | Yes | $0.02 | $0.04 |
Make your first request
Use any OpenAI-compatible SDK. Change the base URL, keep everything else. Responses include a receipt with the exact cost.
No account: send the same request to /x402/v1/chat/completions without credentials and the gateway answers 402 with the quote. Pay per request.
curl https://api.usdf.fi/v1/chat/completions \
-H "Authorization: Bearer $USDF_API_KEY" \
-H "Content-Type: application/json" \
-d '{"model": "llama-3.1-8b-instruct", "messages": [{"role": "user", "content": "Hello"}]}'