Z.ai: GLM 5.3 Flash
GLM-5.3-Flash is a native multimodal model from Z.ai. It is suited for efficient coding and long-horizon agent tasks. Its hybrid sparse and linear attention architecture maintains accurate long-context behavior while...
Anyone in the Project can @-mention Z.ai: GLM 5.3 Flash with the team's shared context - pooled credits, one chat, one memory.
Starter is free forever - 1 Project, 100 credits/month, 1 MCP. No card.
Specifications
- Provider
- z-ai
- Category
- llm
- Context length
- 1,048,576 tokens
- Max output
- 131,072 tokens
- Modalities
- text, image, video
- License
- proprietary
- Released
- 2026-08-26
Pricing
- Input
- $0.07/Mtok
- Output
- $0.25/Mtok
- Model ID
z-ai/glm-5.3-flash
Per-token prices show what the model costs upstream. On Switchy your team draws from one shared org credit pool - one plan, one balance for everyone.
Team cost calculator
5 seats · 80 msgs/day
Switchy meters this against your org's shared credit pool - one plan, one balance for everyone.
Providers
Performance
Benchmarks
Works well with
Top MCPs
Compatibility data comes from first-party telemetry; once we have enough co-usage signal, top MCPs for this model will appear here.
How Switchy teams use it
Starter prompts
Compare with
More language models
- Z.ai: GLM 5.3 Flash (batch)z-ai
- Z.ai: GLM 5 Turboz-ai
- Z.ai: GLM 5V Turboz-ai
- Z.ai: GLM Flash Latestz-ai
- Z.ai: GLM Latestz-ai
- AionLabs: Aion-2.0aion-labs
- AionLabs: Aion-3.0aion-labs
- AionLabs: Aion-3.0-Miniaion-labs
- AionLabs: Aion-RP 1.0 (8B)aion-labs
- Amazon: Nova 2 Liteamazon
- Amazon: Nova Lite 1.0amazon
- Amazon: Nova Micro 1.0amazon