Llama 3.3 70B Instruct FP8 Lora API
Run meta-llama/Llama-3.3-70B-Instruct-FP8-Lora through FastInfra's API.
Pay per token, no subscription, routed to the cheapest available provider.
Llama 3.3 70B Instruct FP8 Lora pricing
Billed per token used. Prices sync automatically from wholesale providers.
| Direction | Price per 1M tokens |
|---|---|
| Input | — |
| Output | — |
Served by 1 provider
Requests route to Together AI by default (cheapest).
Pin a specific provider with the :provider suffix.
| Provider | Upstream model ID |
|---|---|
| together | meta-llama/Llama-3.3-70B-Instruct-FP8-Lora |
Call Llama 3.3 70B Instruct FP8 Lora in 30 seconds
Works with any OpenAI SDK — change the base URL and API key only.
Python
from openai import OpenAI
client = OpenAI(api_key="YOUR_API_KEY", base_url="https://api.fastinfra.ai/v1")
response = client.chat.completions.create(
model="meta-llama/Llama-3.3-70B-Instruct-FP8-Lora",
messages=[{"role": "user", "content": "Hello!"}]
)
print(response.choices[0].message.content)
curl
curl https://api.fastinfra.ai/v1/chat/completions \
-H "Authorization: Bearer YOUR_API_KEY" \
-H "Content-Type: application/json" \
-d '{
"model": "meta-llama/Llama-3.3-70B-Instruct-FP8-Lora",
"messages": [{"role": "user", "content": "Hello!"}]
}'
Llama 3.3 70B Instruct FP8 Lora — common questions
How much does the meta-llama/Llama-3.3-70B-Instruct-FP8-Lora API cost?
On FastInfra, meta-llama/Llama-3.3-70B-Instruct-FP8-Lora costs Pay-per-token pricing via provider routing. Billing is per token used, with no subscription.
Is meta-llama/Llama-3.3-70B-Instruct-FP8-Lora compatible with the OpenAI SDK?
Yes. FastInfra exposes meta-llama/Llama-3.3-70B-Instruct-FP8-Lora through an OpenAI-compatible endpoint at https://api.fastinfra.ai/v1 — point any OpenAI SDK at that base URL with a FastInfra API key and keep your existing code.
Which providers serve meta-llama/Llama-3.3-70B-Instruct-FP8-Lora?
meta-llama/Llama-3.3-70B-Instruct-FP8-Lora is available from 1 provider(s): together. Requests route to Together AI by default (cheapest); append ":provider" to the model ID to pin one.
Related models
Compare side by side: Llama 3.3 70B Instruct FP8 Lora vs Llama3.1:8b · Llama 3.3 70B Instruct FP8 Lora vs Llama3.2:1b · Llama 3.3 70B Instruct FP8 Lora vs Llama3.2:3b