Skip to content
01

Providers + pricing

CapabilitiesProvider details
● Credits · billed by AnyRouter
NVIDIA
nvidia
$0
$0
● Your own key · billed by your provider · AnyRouter fee $0
AIHubMix
aihubmix-byok
$0
$0
Unavailable
NVIDIA
nvidia-byok
$0
$0
Unavailable
OpenRouter
openrouter-byok
$0
$0
Unavailable
Ollama Cloud
ollama-byok
$0
$0
Unavailable
Baseten
baseten-byok
$0
$0
Unavailable
CommandCode
commandcode-byok
$0
$0
Unavailable
Ramp Router
router-byok
$0
$0
Unavailable
Vercel AI Gateway
vercel-byok
$0
$0
Unavailable
● Free pool · donated keys · not available for this model — donate a key
02

Try it

POST /api/v1/chat/completions
import { createAnyRouter } from "@anyr/ai-sdk-provider"import { generateText } from "ai" const anyrouter = createAnyRouter() const { text } = await generateText({  model: anyrouter("nvidia/nemotron-3-ultra-550b-a55b"),  prompt: "Say hi in 3 words.",})console.log(text)
Set ANYROUTER_API_KEY · edits on the demo update this code
Live demo

Type a prompt and run it on this model.

03

Uptime + latency

—recent checks · all routes

No health checks recorded for this model yet.

24h
No traffic in the last 24h.
More performance detail
Usage analytics

Loading usage…

Uptime & Health
No uptime data yet

These providers haven't been health-probed for this model yet. The router still routes around upstreams that fail live requests — uptime fills in once probe history accrues.

Nemotron 3 Ultra 550BFree

Also accepted:nvidia/nemotron-3-ultranemotron-3-ultranvidia/nemotron-3-ultra-550b-a55b:free
  • AIHubMix
  • NVIDIA
  • NVIDIA
  • OpenRouter
  • Ollama Cloud
  • Baseten
  • CommandCode
  • Ramp Router
  • Vercel AI Gateway
Texttext → text1Mtoken context
Context
1M
Input
$0.50
Output
$2.50
TTFT
—
Uptime
—
Routes
9
04

Your access

Credits
Your own key · AnyRouter fee $0

Run Nemotron 3 Ultra 550B on your own key — your requests are billed by the provider. Pool callers pay AnyRouter credits.

No BYOK keys configured for this model yet.

Share a key with the pool to earn credits for every request it serves, covering your plan cost.

Free poolnot in pool yetDonate key →
Create API key for this model
Text generation
Context length1,000,000 tokens
Max output65,536 tokens
ArchitectureTransformer
Categorytext
ReleasedJun 4, 2026
Modalities
Capabilities
05

About

NVIDIA Nemotron-3-Ultra-550B-A55B is a 550B parameter (55B active) frontier model built on a LatentMoE hybrid architecture combining Mamba-2, MoE, and Attention with Multi-Token Prediction. Features a 1M token context window, configurable reasoning mode (enable_thinking), and strong multilingual support across English, French, Spanish, Italian, German, Japanese, Korean, Hindi, Brazilian Portuguese, and Chinese. Best suited for complex agentic workflows, long-context analysis, tool use, and high-stakes RAG.

Released 2026-06-04

API & code
Share cards
Nemotron 3 Ultra 550B share card
Nemotron 3 Ultra 550B
AIHubMix upstream share card
AIHubMix upstream
NVIDIA upstream share card
NVIDIA upstream
NVIDIA upstream share card
NVIDIA upstream
Hue upstream share card
Hue upstream
OpenRouter upstream share card
OpenRouter upstream
Ollama Cloud upstream share card
Ollama Cloud upstream
Baseten upstream share card
Baseten upstream
CommandCode upstream share card
CommandCode upstream
Ramp Router upstream share card
Ramp Router upstream
Vercel AI Gateway upstream share card
Vercel AI Gateway upstream
Back to models