[ qwen ]

Qwen2.5 VL 72B Instruct

Billed at a published markup over the upstream list price. Pay per token in USDF, with no subscription and no minimum. Streaming and the OpenAI-compatible request format are supported.

qwen2.5-vl-72b-instruct·Unavailable·text·Context window: 32,768 tokens

No route to this model is enabled right now. Requests for it are refused with the error code model_unavailable before any balance is held, so nothing is charged. The listed prices apply once a route is enabled.

Input

$1.27

$ / 1M tokens

Output

$1.27

$ / 1M tokens

Cached input

—

not offered

Context window

32,768

tokens

Model ID
qwen2.5-vl-72b-instruct
Modality
text
Unit
USD per 1M tokens
Status
Unavailable
Last verified
—
Endpoints
/v1/chat/completions
Context window
32,768 tokens
Max output
8,192 tokens
Pricing basis
Markup
Cache write (5 min)
—
Cache write (1 h)
—
Routes enabled
0 of 1
Price sheet
2026-10-09.1

Routes

ProviderEnabledUpstream inputUpstream output
huggingfaceNo$1.01$1.01

Make your first request

Use any OpenAI-compatible SDK. Change the base URL, keep everything else. Responses include a receipt with the exact cost.

No account: send the same request to /x402/v1/chat/completions without credentials and the gateway answers 402 with the quote. Pay per request.

This model has no enabled route right now, so the request below is refused with model_unavailable until a route is enabled. Nothing is charged for a refused request.

curl https://api.usdf.fi/v1/chat/completions \
  -H "Authorization: Bearer $FOUNDRY_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{"model": "qwen2.5-vl-72b-instruct", "messages": [{"role": "user", "content": "Hello"}]}'