Providers + pricing
| Capabilities | Provider details | ||||
|---|---|---|---|---|---|
| ● Credits · billed by AnyRouter | |||||
NVIDIA nvidia | $0 | $0 | |||
| ● Your own key · billed by your provider · AnyRouter fee $0 | |||||
NVIDIA nvidia-byok | $0 | $0 | Unavailable | ||
| ● Free pool · donated keys · not available for this model — donate a key | |||||
Try it
import OpenAI from "openai" const client = new OpenAI({ apiKey: process.env.ANYROUTER_API_KEY, baseURL: "https://anyrouter.dev/api/v1",}) const resp = await client.embeddings.create({ model: "nvidia/llama-nemotron-embed-vl-1b-v2", input: "The quick brown fox jumps over the lazy dog",})console.log(resp.data[0].embedding.slice(0, 8))
ANYROUTER_API_KEY · edits on the demo update this codeUptime + latency
No health checks recorded for this model yet.
More performance detail
Llama Nemotron Embed VL 1B v2Removed
This model has been removed. Requests that still use this id are routed to nvidia/nemotron-3-embed-1b.
- NVIDIA
- NVIDIA
Your access
Run Llama Nemotron Embed VL 1B v2 on your own key — your requests are billed by the provider. Pool callers pay AnyRouter credits.
No BYOK keys configured for this model yet.
Share a key with the pool to earn credits for every request it serves, covering your plan cost.
About
NVIDIA Llama-Nemotron-Embed-VL-1B-v2 is a vision-language embedding model for multimodal question-answering and retrieval over text, images, or combined image-text documents. A transformer encoder fine-tuned from Llama 3.2 1B with SigLip2 400M, it uses a tiling-based VLM architecture (Eagle 2 + nemoretriever-parse) for high-resolution image and complex visual-document understanding. Served via NVIDIA NIM.
Released 2026-06-01 · params: input · model · encoding_format · input_type


