NVIDIA Nemotron 3 Nano 30B A3B BF16 API

Run nvidia/NVIDIA-Nemotron-3-Nano-30B-A3B-BF16 through FastInfra's OpenAI-compatible API. Pay per token, no subscription, routed to the cheapest available provider.

Read the docs
Pricing

NVIDIA Nemotron 3 Nano 30B A3B BF16 pricing

Billed per token used. Prices sync automatically from wholesale providers.

Direction Price per 1M tokens
Input
Output
Availability

Served by 1 provider

Requests route to Together AI by default (cheapest). Pin a specific provider with the :provider suffix.

Provider Upstream model ID
together nvidia/NVIDIA-Nemotron-3-Nano-30B-A3B-BF16
Quickstart

Call NVIDIA Nemotron 3 Nano 30B A3B BF16 in 30 seconds

Works with any OpenAI SDK — change the base URL and API key only.

Python

from openai import OpenAI

client = OpenAI(api_key="YOUR_API_KEY", base_url="https://api.fastinfra.ai/v1")

response = client.chat.completions.create(
    model="nvidia/NVIDIA-Nemotron-3-Nano-30B-A3B-BF16",
    messages=[{"role": "user", "content": "Hello!"}]
)
print(response.choices[0].message.content)

curl

curl https://api.fastinfra.ai/v1/chat/completions \
  -H "Authorization: Bearer YOUR_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{
    "model": "nvidia/NVIDIA-Nemotron-3-Nano-30B-A3B-BF16",
    "messages": [{"role": "user", "content": "Hello!"}]
  }'
FAQ

NVIDIA Nemotron 3 Nano 30B A3B BF16 — common questions

How much does the nvidia/NVIDIA-Nemotron-3-Nano-30B-A3B-BF16 API cost?

On FastInfra, nvidia/NVIDIA-Nemotron-3-Nano-30B-A3B-BF16 costs Pay-per-token pricing via provider routing. Billing is per token used, with no subscription.

Is nvidia/NVIDIA-Nemotron-3-Nano-30B-A3B-BF16 compatible with the OpenAI SDK?

Yes. FastInfra exposes nvidia/NVIDIA-Nemotron-3-Nano-30B-A3B-BF16 through an OpenAI-compatible endpoint at https://api.fastinfra.ai/v1 — point any OpenAI SDK at that base URL with a FastInfra API key and keep your existing code.

Which providers serve nvidia/NVIDIA-Nemotron-3-Nano-30B-A3B-BF16?

nvidia/NVIDIA-Nemotron-3-Nano-30B-A3B-BF16 is available from 1 provider(s): together. Requests route to Together AI by default (cheapest); append ":provider" to the model ID to pin one.