🎉 LTX 2.5 IS LIVE! 🎉 | The wait is over. We’ve just pushed LTX 2.5 to production, and it’s our best update yet.

⚡️ GPT-OSS-120b is now live | Blazing fast - better than TogetherAI

👉 Get $5 welcome credit | One API for every frontier model | Ends soon.

NVIDIA Nemotron 3.5 Lightning API

Run nvidia/NVIDIA-Nemotron-3.5-Lightning through FastInfra's API. Pay per token, no subscription, routed to the cheapest available provider.

Pricing

NVIDIA Nemotron 3.5 Lightning pricing

Billed per token used. Prices sync automatically from wholesale providers.

Direction Price per 1M tokens
Input$0.08
Output$0.21
Availability

Served by 1 provider

Requests route to DeepInfra by default (cheapest). Pin a specific provider with the :provider suffix.

Provider Upstream model ID
deepinfra nvidia/NVIDIA-Nemotron-3.5-Lightning
Quickstart

Call NVIDIA Nemotron 3.5 Lightning in 30 seconds

Works with any OpenAI SDK — change the base URL and API key only.

Python

from openai import OpenAI

client = OpenAI(api_key="YOUR_API_KEY", base_url="https://api.fastinfra.ai/v1")

response = client.chat.completions.create(
    model="nvidia/NVIDIA-Nemotron-3.5-Lightning",
    messages=[{"role": "user", "content": "Hello!"}]
)
print(response.choices[0].message.content)

curl

curl https://api.fastinfra.ai/v1/chat/completions \
  -H "Authorization: Bearer YOUR_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{
    "model": "nvidia/NVIDIA-Nemotron-3.5-Lightning",
    "messages": [{"role": "user", "content": "Hello!"}]
  }'
FAQ

NVIDIA Nemotron 3.5 Lightning — common questions

How much does the nvidia/NVIDIA-Nemotron-3.5-Lightning API cost?

On FastInfra, nvidia/NVIDIA-Nemotron-3.5-Lightning costs $0.084/1M input, $0.21/1M output tokens. Billing is per token used, with no subscription.

Is nvidia/NVIDIA-Nemotron-3.5-Lightning compatible with the OpenAI SDK?

Yes. FastInfra exposes nvidia/NVIDIA-Nemotron-3.5-Lightning through an OpenAI-compatible endpoint at https://api.fastinfra.ai/v1 — point any OpenAI SDK at that base URL with a FastInfra API key and keep your existing code.

Which providers serve nvidia/NVIDIA-Nemotron-3.5-Lightning?

nvidia/NVIDIA-Nemotron-3.5-Lightning is available from 1 provider(s): deepinfra. Requests route to DeepInfra by default (cheapest); append ":provider" to the model ID to pin one.