🎉 LTX 2.5 IS LIVE! 🎉 | The wait is over. We’ve just pushed LTX 2.5 to production, and it’s our best update yet.

⚡️ GPT-OSS-120b is now live | Blazing fast - better than TogetherAI

👉 Get $5 welcome credit | One API for every frontier model | Ends soon.

Inkling Small API

Run thinkingmachines/inkling-small through FastInfra's API. Pay per token, no subscription, routed to the cheapest available provider.

Pricing

Inkling Small pricing

Billed per token used. Prices sync automatically from wholesale providers.

Direction Price per 1M tokens
Input$0.47
Output$1.26
Availability

Served by 3 providers

Requests route to DeepInfra by default (cheapest). Pin a specific provider with the :provider suffix.

Provider Upstream model ID
deepinfra thinkingmachines/Inkling-Small
openrouter thinkingmachines/inkling-small
together thinkingmachines/Inkling-Small
Quickstart

Call Inkling Small in 30 seconds

Works with any OpenAI SDK — change the base URL and API key only.

Python

from openai import OpenAI

client = OpenAI(api_key="YOUR_API_KEY", base_url="https://api.fastinfra.ai/v1")

response = client.chat.completions.create(
    model="thinkingmachines/inkling-small",
    messages=[{"role": "user", "content": "Hello!"}]
)
print(response.choices[0].message.content)

curl

curl https://api.fastinfra.ai/v1/chat/completions \
  -H "Authorization: Bearer YOUR_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{
    "model": "thinkingmachines/inkling-small",
    "messages": [{"role": "user", "content": "Hello!"}]
  }'
FAQ

Inkling Small — common questions

How much does the thinkingmachines/inkling-small API cost?

On FastInfra, thinkingmachines/inkling-small costs $0.4725/1M input, $1.26/1M output tokens. Billing is per token used, with no subscription.

Is thinkingmachines/inkling-small compatible with the OpenAI SDK?

Yes. FastInfra exposes thinkingmachines/inkling-small through an OpenAI-compatible endpoint at https://api.fastinfra.ai/v1 — point any OpenAI SDK at that base URL with a FastInfra API key and keep your existing code.

Which providers serve thinkingmachines/inkling-small?

thinkingmachines/inkling-small is available from 3 provider(s): deepinfra, openrouter, together. Requests route to DeepInfra by default (cheapest); append ":provider" to the model ID to pin one.