Qwen3 Max API
Run qwen/qwen3-max through FastInfra's API.
Pay per token, no subscription, routed to the cheapest available provider.
Qwen3 Max pricing
Billed per token used. Prices sync automatically from wholesale providers.
| Direction | Price per 1M tokens |
|---|---|
| Input | $0.82 |
| Output | $4.10 |
2 routes available
Requests route to route-07 by default (lowest cost). FastInfra fails over automatically if a route is unavailable.
| Route label | Model ID |
|---|---|
route-07 |
qwen/qwen3-max |
route-21 |
qwen/qwen3-max |
Call Qwen3 Max in 30 seconds
Works with any OpenAI SDK — change the base URL and API key only.
Python
from openai import OpenAI
client = OpenAI(api_key="YOUR_API_KEY", base_url="https://api.fastinfra.ai/v1")
response = client.chat.completions.create(
model="qwen/qwen3-max",
messages=[{"role": "user", "content": "Hello!"}]
)
print(response.choices[0].message.content)
curl
curl https://api.fastinfra.ai/v1/chat/completions \
-H "Authorization: Bearer YOUR_API_KEY" \
-H "Content-Type: application/json" \
-d '{
"model": "qwen/qwen3-max",
"messages": [{"role": "user", "content": "Hello!"}]
}'
Qwen3 Max — common questions
How much does the qwen/qwen3-max API cost?
On FastInfra, qwen/qwen3-max costs $0.819/1M input, $4.095/1M output tokens. Billing is per token used, with no subscription.
Is qwen/qwen3-max compatible with the OpenAI SDK?
Yes. FastInfra exposes qwen/qwen3-max through an OpenAI-compatible endpoint at https://api.fastinfra.ai/v1 — point any OpenAI SDK at that base URL with a FastInfra API key and keep your existing code.
How does routing work for qwen/qwen3-max?
qwen/qwen3-max is available on 2 route(s): route-07, route-21. Requests use route-07 by default (lowest cost). FastInfra fails over automatically if a route is unavailable.
Related models
Compare side by side: Qwen3 Max vs Qwen2.5:0.5b · Qwen3 Max vs Qwen2.5:1.5b · Qwen3 Max vs Qwen2.5:14b