# GLM-5.3-Flash

> GLM-5.3-Flash is a native multimodal model from Z.ai (320B-A18B, MIT License, 1M-token context). It is suited for efficient coding and long-horizon agent tasks. Its hybrid sparse and linear attention architecture maintains accurate long-context behavior while reducing compute overhead.

**ID**: `z-ai/glm-5.3-flash`  
**Creator**: Zhipu AI  
**Category**: text  
**Context**: 1M tokens  
**Pricing**: $0.15 / $0.50 per 1M · billed by your provider (BYOK)  
**Released**: 2026-08-26  
**Web page**: https://anyrouter.dev/model/z-ai/glm-5.3-flash

**Input modalities**: text, image, video  
**Output modalities**: text  

**Tokenizer**: GLM

**Capabilities**: chat, streaming, reasoning, coding, agentic, function-calling, vision

**Aliases**: `zai-org/glm-5.3-flash`

## Usage

**Endpoint**: `POST https://anyrouter.dev/api/v1/chat/completions`

```bash
curl https://anyrouter.dev/api/v1/chat/completions \
  -H "Authorization: Bearer $ANYROUTER_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{
  "model": "z-ai/glm-5.3-flash",
  "messages": [
    {
      "role": "user",
      "content": "Say hi in 3 words."
    }
  ]
}'
```

## Providers

| Provider | Input | Output |
| --- | --- | --- |
| hue | $0.15/M | $0.50/M |
| aihubmix-byok | $0.15/M (list) | $0.50/M (list) |
| aihubmix-byok | $0.15/M (list) | $0.50/M (list) |
| opencode-zen-byok | $0.15/M (list) | $0.50/M (list) |
| commandcode-byok | $0.15/M (list) | $0.50/M (list) |
| hermes-agent | $0.15/M | $0.50/M |
| hermes-agent-byok | $0.15/M (list) | $0.50/M (list) |
| openrouter-byok | $0.15/M (list) | $0.50/M (list) |
| venice-byok | $0.15/M (list) | $0.50/M (list) |
| cline-byok | $0.15/M (list) | $0.50/M (list) |
| router-byok | $0.15/M (list) | $0.50/M (list) |

BYOK routes are billed by your own provider key — AnyRouter charges $0. Rates marked (list) are the provider's published list price.

## Source

- Upstream docs: https://z.ai/blog/glm-5.3-flash
