[ meta ]
Llama 4 Scout 17B-16E Instruct
Billed at the provider's list price, with no markup. Pay per token in USDF, with no subscription and no minimum. Streaming and the OpenAI-compatible request format are supported.
Input
$0.10
$ / 1M tokens
Output
$0.30
$ / 1M tokens
Cached input
not offered
Context window
327,680
tokens
- Model ID
- llama-4-scout
- Modality
- text
- Unit
- USD per 1M tokens
- Status
- Available
- Endpoints
- /v1/chat/completions
- Context window
- 327,680 tokens
- Max output
- 16,384 tokens
- Pricing basis
- List price
- Routes enabled
- 1 of 2
- Price sheet
- 2026-10-10.1
Routes
| Provider | Enabled | Upstream input | Upstream output |
|---|---|---|---|
| deepinfra | Yes | $0.10 | $0.30 |
| huggingface | No | $0.09 | $0.29 |
Make your first request
Use any OpenAI-compatible SDK. Change the base URL, keep everything else. Responses include a receipt with the exact cost.
No account: send the same request to /x402/v1/chat/completions without credentials and the gateway answers 402 with the quote. Pay per request.
curl https://api.usdf.fi/v1/chat/completions \
-H "Authorization: Bearer $USDF_API_KEY" \
-H "Content-Type: application/json" \
-d '{"model": "llama-4-scout", "messages": [{"role": "user", "content": "Hello"}]}'