Skip to content
01

Providers + pricing

CapabilitiesProvider details
● Your own key · billed by your provider · AnyRouter fee $0
AIHubMix
aihubmix-byok
$0.06list$0.40list
—
Unavailable
DeepInfra
deepinfra-byok
$0.06list$0.40list
$0
Unavailable
Z-AI
z-ai-byok
$0.06list$0.40list
—
Unavailable
Z-AI
z-ai-coding-byok
$0.06list$0.40list
—
Unavailable
Ollama Cloud
ollama-byok
$0.06list$0.40list
—
Unavailable
OpenRouter
openrouter-byok
$0.06list$0.40list
—
Unavailable
ZenMux
zenmux-byok
$0.06list$0.40list
—
Unavailable
● Free pool · donated keys · not available for this model — donate a key
02

Try it

POST /api/v1/chat/completions
import { createAnyRouter } from "@anyr/ai-sdk-provider"import { generateText } from "ai" const anyrouter = createAnyRouter() const { text } = await generateText({  model: anyrouter("z-ai/glm-4.7-flash"),  prompt: "Say hi in 3 words.",})console.log(text)
Set ANYROUTER_API_KEY · edits on the demo update this code
Live demo

Type a prompt and run it on this model.

03

Uptime + latency

—recent checks · all routes

No health checks recorded for this model yet.

24h
No traffic in the last 24h.
More performance detail
Usage analytics

Loading usage…

Uptime & Health
No uptime data yet

These providers haven't been health-probed for this model yet. The router still routes around upstreams that fail live requests — uptime fills in once probe history accrues.

GLM-4.7 Flash

Also accepted:zai-org/glm-4.7-flash
  • AIHubMix
  • DeepInfra
  • Z-AI
  • Z-AI
  • Ollama Cloud
  • OpenRouter
  • ZenMux
Texttext → text128Ktoken context
Context
128K
Input
$0.06
Output
$0.40
TTFT
200ms
Uptime
—
Routes
7
04

Your access

Credits
Your own key · AnyRouter fee $0

Run GLM-4.7 Flash on your own key — your requests are billed by the provider. Pool callers pay AnyRouter credits.

No BYOK keys configured for this model yet.

Share a key with the pool to earn credits for every request it serves, covering your plan cost.

Free poolnot in pool yetDonate key →
Create API key for this model
Text generation
Context length128,000 tokens
Max output128,000 tokens
ArchitectureTransformer
Categorytext
ReleasedApr 10, 2025
Modalities
Capabilities
05

About

GLM-4.7-Flash is a fast and efficient multilingual text generation model with a 131,072 token context window. Optimized for dialogue, instruction-following, and multi-turn tool calling across 100+ languages.

Released 2025-04-10

API & code
Share cards
GLM-4.7 Flash share card
GLM-4.7 Flash
AIHubMix upstream share card
AIHubMix upstream
Hue upstream share card
Hue upstream
DeepInfra upstream share card
DeepInfra upstream
Z-AI upstream share card
Z-AI upstream
Z-AI upstream share card
Z-AI upstream
Ollama Cloud upstream share card
Ollama Cloud upstream
OpenRouter upstream share card
OpenRouter upstream
ZenMux upstream share card
ZenMux upstream
Back to models