🚀 One API for every frontier model on FastInfra

Browse models

Transparent per-token pricing → compare providers on FastInfra

View pricing

OpenAI-compatible chat completions → start in minutes

Read the docs

⚡ Server Auction → Enterprise bare metal from $78.15/mo (173 in stock)

Browse deals

Llama 3.3 70B Instruct Turbo API

Run meta-llama/Llama-3.3-70B-Instruct-Turbo through FastInfra's API. Pay per token, no subscription, routed to the cheapest available provider.

Pricing

Llama 3.3 70B Instruct Turbo pricing

Billed per token used. Prices sync automatically from wholesale providers.

Direction Price per 1M tokens
Input$0.11
Output$0.34
Availability

2 routes available

Requests route to route-21 by default (lowest cost). FastInfra fails over automatically if a route is unavailable.

Route label Model ID
route-21 meta-llama/Llama-3.3-70B-Instruct-Turbo
route-03 meta-llama/Llama-3.3-70B-Instruct-Turbo
Quickstart

Call Llama 3.3 70B Instruct Turbo in 30 seconds

Works with any OpenAI SDK — change the base URL and API key only.

Python

from openai import OpenAI

client = OpenAI(api_key="YOUR_API_KEY", base_url="https://api.fastinfra.ai/v1")

response = client.chat.completions.create(
    model="meta-llama/Llama-3.3-70B-Instruct-Turbo",
    messages=[{"role": "user", "content": "Hello!"}]
)
print(response.choices[0].message.content)

curl

curl https://api.fastinfra.ai/v1/chat/completions \
  -H "Authorization: Bearer YOUR_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{
    "model": "meta-llama/Llama-3.3-70B-Instruct-Turbo",
    "messages": [{"role": "user", "content": "Hello!"}]
  }'
FAQ

Llama 3.3 70B Instruct Turbo — common questions

How much does the meta-llama/Llama-3.3-70B-Instruct-Turbo API cost?

On FastInfra, meta-llama/Llama-3.3-70B-Instruct-Turbo costs $0.105/1M input, $0.336/1M output tokens. Billing is per token used, with no subscription.

Is meta-llama/Llama-3.3-70B-Instruct-Turbo compatible with the OpenAI SDK?

Yes. FastInfra exposes meta-llama/Llama-3.3-70B-Instruct-Turbo through an OpenAI-compatible endpoint at https://api.fastinfra.ai/v1 — point any OpenAI SDK at that base URL with a FastInfra API key and keep your existing code.

How does routing work for meta-llama/Llama-3.3-70B-Instruct-Turbo?

meta-llama/Llama-3.3-70B-Instruct-Turbo is available on 2 route(s): route-21, route-03. Requests use route-21 by default (lowest cost). FastInfra fails over automatically if a route is unavailable.