# Nemotron 3 120B A12B

> NVIDIA's Mixture-of-Experts model with 120B total parameters and 12B active, optimized for efficient inference with strong reasoning capabilities.

**ID**: `nvidia/nemotron-3-120b-a12b`  
**Creator**: nvidia  
**Category**: text  
**Context**: 131K tokens  
**Pricing**: $0.50 / $1.50 per 1M · billed by your provider (BYOK)  
**Released**: 2025-03-15  
**Web page**: https://anyrouter.dev/model/nvidia/nemotron-3-120b-a12b

**Input modalities**: text  
**Output modalities**: text  

**Tokenizer**: Nemotron

**Capabilities**: chat, reasoning, function-calling, streaming, coding

**Supported parameters**: max_tokens, temperature, top_p, stop, tools, tool_choice, response_format

## Usage

**Endpoint**: `POST https://anyrouter.dev/api/v1/chat/completions`

```bash
curl https://anyrouter.dev/api/v1/chat/completions \
  -H "Authorization: Bearer $ANYROUTER_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{
  "model": "nvidia/nemotron-3-120b-a12b",
  "messages": [
    {
      "role": "user",
      "content": "Say hi in 3 words."
    }
  ]
}'
```

## Providers

| Provider | Input | Output |
| --- | --- | --- |
| openrouter-byok | $0.50/M (list) | $1.50/M (list) |
| openrouter-byok | $0.50/M (list) | $1.50/M (list) |

BYOK routes are billed by your own provider key — AnyRouter charges $0. Rates marked (list) are the provider's published list price.

## Benchmarks

- **GPQA Diamond**: 0.8
- **HLE**: 0.2

## Source

- Upstream docs: https://developers.cloudflare.com/workers-ai/models/nemotron-3-120b-a12b/
