Skip to content
01

Providers + pricing

CapabilitiesProvider details
● Credits · billed by AnyRouter
NVIDIA
nvidia
$0
$0
● Your own key · billed by your provider · AnyRouter fee $0
NVIDIA
nvidia-byok
$0
$0
Unavailable
OpenRouter
openrouter-byok
$0
$0
Unavailable
● Free pool · donated keys · not available for this model — donate a key
02

Try it

POST /api/v1/chat/completions
import { createAnyRouter } from "@anyr/ai-sdk-provider"import { generateText } from "ai" const anyrouter = createAnyRouter() const { text } = await generateText({  model: anyrouter("deepseek-ai/deepseek-v4-flash-0731"),  prompt: "Say hi in 3 words.",})console.log(text)
Set ANYROUTER_API_KEY · edits on the demo update this code
03

Uptime + latency

—recent checks · all routes

No health checks recorded for this model yet.

24h
No traffic in the last 24h.
More performance detail
Usage analytics

Loading usage…

Uptime & Health
No uptime data yet

These providers haven't been health-probed for this model yet. The router still routes around upstreams that fail live requests — uptime fills in once probe history accrues.

DeepSeek V4 Flash 0731 (NIM)Removed

This model has been removed. Requests that still use this id are routed to deepseek/deepseek-v4-flash.

Also accepted:deepseek/deepseek-v4-flash-0731
  • NVIDIA
  • NVIDIA
  • OpenRouter
Texttext → text128Ktoken context
Context
128K
Input
$0.14
Output
$0.28
TTFT
—
Uptime
—
Routes
3
04

Your access

Credits
Your own key · AnyRouter fee $0

Run DeepSeek V4 Flash 0731 (NIM) on your own key — your requests are billed by the provider. Pool callers pay AnyRouter credits.

No BYOK keys configured for this model yet.

Share a key with the pool to earn credits for every request it serves, covering your plan cost.

Free poolnot in pool yetDonate key →
Create API key for this model
Text generation
Context length128,000 tokens
Max output102,400 tokens
ArchitectureTransformer
Categorytext
ReleasedJul 31, 2026
Modalities
Capabilities
05

About

DeepSeek-V4-Flash build 0731 on NVIDIA NIM — a dated NIM packaging of the DeepSeek V4 Flash MoE for low-latency chat and coding.

Released 2026-07-31 · params: max_tokens · temperature · top_p · stop · tools · tool_choice

API & code
Share cards
DeepSeek V4 Flash 0731 (NIM) share card
DeepSeek V4 Flash 0731 (NIM)
NVIDIA upstream share card
NVIDIA upstream
NVIDIA upstream share card
NVIDIA upstream
OpenRouter upstream share card
OpenRouter upstream
Back to models