[ qwen ]
Qwen3 4B Instruct 2507
Billed at a published markup over the upstream list price. Pay per token in USDF, with no subscription and no minimum. Streaming and the OpenAI-compatible request format are supported.
No route to this model is enabled right now. Requests for it are refused with the error code model_unavailable before any balance is held, so nothing is charged. The listed prices apply once a route is enabled.
Input
$0.02
$ / 1M tokens
Output
$0.04
$ / 1M tokens
Cached input
—
not offered
Context window
262,144
tokens
- Model ID
- qwen3-4b-instruct-2507
- Modality
- text
- Unit
- USD per 1M tokens
- Status
- Unavailable
- Last verified
- —
- Endpoints
- /v1/chat/completions
- Context window
- 262,144 tokens
- Max output
- 8,192 tokens
- Pricing basis
- Markup
- Cache write (5 min)
- —
- Cache write (1 h)
- —
- Routes enabled
- 0 of 1
- Price sheet
- 2026-10-09.1
Routes
| Provider | Enabled | Upstream input | Upstream output |
|---|---|---|---|
| huggingface | No | $0.01 | $0.03 |
Make your first request
Use any OpenAI-compatible SDK. Change the base URL, keep everything else. Responses include a receipt with the exact cost.
No account: send the same request to /x402/v1/chat/completions without credentials and the gateway answers 402 with the quote. Pay per request.
This model has no enabled route right now, so the request below is refused with model_unavailable until a route is enabled. Nothing is charged for a refused request.
curl https://api.usdf.fi/v1/chat/completions \
-H "Authorization: Bearer $FOUNDRY_API_KEY" \
-H "Content-Type: application/json" \
-d '{"model": "qwen3-4b-instruct-2507", "messages": [{"role": "user", "content": "Hello"}]}'