# Nemotron 3 Ultra 550B

> NVIDIA Nemotron-3-Ultra-550B-A55B is a 550B parameter (55B active) frontier model built on a LatentMoE hybrid architecture combining Mamba-2, MoE, and Attention with Multi-Token Prediction. Features a 1M token context window, configurable reasoning mode (enable_thinking), and strong multilingual support across English, French, Spanish, Italian, German, Japanese, Korean, Hindi, Brazilian Portuguese, and Chinese. Best suited for complex agentic workflows, long-context analysis, tool use, and high-stakes RAG.

**ID**: `nvidia/nemotron-3-ultra-550b-a55b`  
**Creator**: nvidia  
**Category**: text  
**Context**: 1M tokens  
**Input**: $0.50/M  
**Output**: $2.50/M  
**Released**: 2026-06-04  
**Web page**: https://anyrouter.dev/model/nvidia/nemotron-3-ultra-550b-a55b

**Input modalities**: text  
**Output modalities**: text  

**Tokenizer**: Nemotron

**Capabilities**: chat, streaming, reasoning, function-calling, thinking

**Aliases**: `nvidia/nemotron-3-ultra`, `nemotron-3-ultra`, `nvidia/nemotron-3-ultra-550b-a55b:free`

## Usage

**Endpoint**: `POST https://anyrouter.dev/api/v1/chat/completions`

```bash
curl https://anyrouter.dev/api/v1/chat/completions \
  -H "Authorization: Bearer $ANYROUTER_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{
  "model": "nvidia/nemotron-3-ultra-550b-a55b",
  "messages": [
    {
      "role": "user",
      "content": "Say hi in 3 words."
    }
  ]
}'
```

## Providers

| Provider | Input | Output |
| --- | --- | --- |
| aihubmix-byok | $0 | $0 |
| nvidia | $0 | $0 |
| nvidia-byok | $0 | $0 |
| hue | $0 | $0 |
| openrouter-byok | $0 | $0 |
| openrouter-byok | $0 | $0 |
| ollama-byok | $0 | $0 |
| baseten-byok | $0 | $0 |
| commandcode-byok | $0 | $0 |
| router-byok | $0 | $0 |
| vercel-byok | $0 | $0 |

## Source

- Upstream docs: https://build.nvidia.com/nvidia/nemotron-3-ultra-550b-a55b
