Llama Nemotron Embed Vl 1b v2 API
Run nvidia/llama-nemotron-embed-vl-1b-v2 through FastInfra's API.
Pay per token, no subscription, routed to the cheapest available provider.
Llama Nemotron Embed Vl 1b v2 pricing
Billed per token used. Prices sync automatically from wholesale providers.
| Direction | Price per 1M tokens |
|---|---|
| Input | $0.01 |
| Output | — |
Served by 1 provider
Requests route to DeepInfra by default (cheapest).
Pin a specific provider with the :provider suffix.
| Provider | Upstream model ID |
|---|---|
| deepinfra | nvidia/llama-nemotron-embed-vl-1b-v2 |
Call Llama Nemotron Embed Vl 1b v2 in 30 seconds
Works with any OpenAI SDK — change the base URL and API key only.
Python
from openai import OpenAI
client = OpenAI(api_key="YOUR_API_KEY", base_url="https://api.fastinfra.ai/v1")
response = client.chat.completions.create(
model="nvidia/llama-nemotron-embed-vl-1b-v2",
messages=[{"role": "user", "content": "Hello!"}]
)
print(response.choices[0].message.content)
curl
curl https://api.fastinfra.ai/v1/chat/completions \
-H "Authorization: Bearer YOUR_API_KEY" \
-H "Content-Type: application/json" \
-d '{
"model": "nvidia/llama-nemotron-embed-vl-1b-v2",
"messages": [{"role": "user", "content": "Hello!"}]
}'
Llama Nemotron Embed Vl 1b v2 — common questions
How much does the nvidia/llama-nemotron-embed-vl-1b-v2 API cost?
On FastInfra, nvidia/llama-nemotron-embed-vl-1b-v2 costs $0.0105/1M input, $0/1M output tokens. Billing is per token used, with no subscription.
Is nvidia/llama-nemotron-embed-vl-1b-v2 compatible with the OpenAI SDK?
Yes. FastInfra exposes nvidia/llama-nemotron-embed-vl-1b-v2 through an OpenAI-compatible endpoint at https://api.fastinfra.ai/v1 — point any OpenAI SDK at that base URL with a FastInfra API key and keep your existing code.
Which providers serve nvidia/llama-nemotron-embed-vl-1b-v2?
nvidia/llama-nemotron-embed-vl-1b-v2 is available from 1 provider(s): deepinfra. Requests route to DeepInfra by default (cheapest); append ":provider" to the model ID to pin one.
Related models
Compare side by side: Llama Nemotron Embed Vl 1b v2 vs Nemotron 3 Nano 30b a3b · Llama Nemotron Embed Vl 1b v2 vs Nemotron 3.5 Lightning · Llama Nemotron Embed Vl 1b v2 vs NVIDIA Nemotron 3.5 Lightning