reasoningZ.ai

Glm 5.3 Flash

Auto-discovered from the NVIDIA catalog on 2026-09-11.

z-ai/glm-5.3-flash
Running· verified 17 min ago · 8703ms

What is Glm 5.3 Flash?

Glm 5.3 Flash is Z.ai's reasoning model. Auto-discovered from the NVIDIA catalog on 2026-09-11. It runs on Speka's OpenAI-compatible /v1/chat/completions endpoint at $0.20 / 1M input and $0.20 / 1M output tokens, using the model ID z-ai/glm-5.3-flash.

Test it live

Live playgroundGlm 5.3 Flashin-browser demo

Test Glm 5.3 Flash live

Send a message and watch the model respond. Multi-turn — it remembers the conversation.