[ nousresearch ]
Hermes 3 Llama 3.1 70B
Billed at the provider's list price, with no markup. Pay per token in USDF, with no subscription and no minimum. Streaming and the OpenAI-compatible request format are supported.
Input
$0.70
$ / 1M tokens
Output
$0.70
$ / 1M tokens
Cached input
not offered
Context window
131,072
tokens
- Model ID
- hermes-3-llama-3.1-70b
- Modality
- text
- Unit
- USD per 1M tokens
- Status
- Available
- Endpoints
- /v1/chat/completions
- Context window
- 131,072 tokens
- Max output
- 131,072 tokens
- Pricing basis
- List price
- Routes enabled
- 1 of 1
- Price sheet
- 2026-10-10.1
Routes
| Provider | Enabled | Upstream input | Upstream output |
|---|---|---|---|
| deepinfra | Yes | $0.70 | $0.70 |
Make your first request
Use any OpenAI-compatible SDK. Change the base URL, keep everything else. Responses include a receipt with the exact cost.
No account: send the same request to /x402/v1/chat/completions without credentials and the gateway answers 402 with the quote. Pay per request.
curl https://api.usdf.fi/v1/chat/completions \
-H "Authorization: Bearer $USDF_API_KEY" \
-H "Content-Type: application/json" \
-d '{"model": "hermes-3-llama-3.1-70b", "messages": [{"role": "user", "content": "Hello"}]}'