NVIDIA logo

NVIDIA Nemotron 3 Ultra 550B API

nvidia/nemotron-3-ultra-550b-a55b

Call NVIDIA Nemotron 3 Ultra 550B through UberLLM’s OpenAI-compatible API — from $0.5000/M input and $2.50/M output, routed automatically to the cheapest healthy provider. NVIDIA · 1,000,000 token context. No inference markup.

Try it live — no signup
Free demo · a few prompts/day

NVIDIA Nemotron 3 Ultra 550B pricing & providers

ProviderInput / 1MOutput / 1M
Surplus Intelligencebest$0.5000$2.50
OpenRouter$0.6000$3.60

How to use NVIDIA Nemotron 3 Ultra 550B with the OpenAI SDK

Already using OpenAI or OpenRouter? Change one line — the base URL — and call NVIDIA Nemotron 3 Ultra 550B with your existing code. Anthropic SDKs work against https://api.uberllm.dev/v1/messages.

python
from openai import OpenAI

client = OpenAI(
    base_url="https://api.uberllm.dev/v1",
    api_key="ull_your_key_here",
)

resp = client.chat.completions.create(
    model="nvidia/nemotron-3-ultra-550b-a55b",
    messages=[{"role": "user", "content": "Hello!"}],
)
print(resp.choices[0].message.content)

NVIDIA Nemotron 3 Ultra 550B API — FAQ

How much does the NVIDIA Nemotron 3 Ultra 550B API cost?

NVIDIA Nemotron 3 Ultra 550B costs from $0.5000 per million input tokens and $2.50 per million output tokens through UberLLM, which routes to the cheapest healthy provider. There is no inference markup.

How do I call NVIDIA Nemotron 3 Ultra 550B via API?

Use any OpenAI-compatible SDK: set the base URL to https://api.uberllm.dev/v1, your UberLLM API key, and model "nvidia/nemotron-3-ultra-550b-a55b". Anthropic SDKs work via https://api.uberllm.dev/v1/messages.

Is NVIDIA Nemotron 3 Ultra 550B OpenAI-compatible on UberLLM?

Yes. Every model on UberLLM — including NVIDIA Nemotron 3 Ultra 550B — is served through one OpenAI-compatible endpoint, so you change only the base URL and key.

More NVIDIA models