Providers + pricing
| Capabilities | Provider details | ||||
|---|---|---|---|---|---|
| ● Credits · billed by AnyRouter | |||||
NVIDIA nvidia | $0 | $0 | |||
| ● Your own key · billed by your provider · AnyRouter fee $0 | |||||
NVIDIA nvidia-byok | $0 | $0 | Unavailable | ||
| ● Free pool · donated keys · not available for this model — donate a key | |||||
Try it
import { createAnyRouter } from "@anyr/ai-sdk-provider"import { generateText } from "ai" const anyrouter = createAnyRouter() const { text } = await generateText({ model: anyrouter("google/diffusiongemma-26b-a4b-it"), prompt: "Say hi in 3 words.",})console.log(text)
ANYROUTER_API_KEY · edits on the demo update this codeUptime + latency
No health checks recorded for this model yet.
More performance detail
DiffusionGemma 26B A4BDisabled since Aug 19, 2026
Chat completions return empty content for this discrete-diffusion model under normal chat token budgets. It is unlisted rather than advertised as a live chat model.
- NVIDIA
- NVIDIA
Your access
Run DiffusionGemma 26B A4B on your own key — your requests are billed by the provider. Pool callers pay AnyRouter credits.
No BYOK keys configured for this model yet.
Share a key with the pool to earn credits for every request it serves, covering your plan cost.
About
DiffusionGemma 26B A4B IT is an open-weights multimodal model from Google DeepMind that generates text via discrete diffusion. Built on the Gemma 4 26B A4B MoE architecture (25.2B total / 3.8B active), it emits tokens in parallel 256-token blocks for high-throughput generation, with a 256K context window, configurable thinking mode, native function calling, and 35+ language support.
Released 2026-06-10 · params: max_tokens · temperature · top_p · stop


