GLM 5.3 FlashX
by Z Ai
GLM-5.3-FlashX is the high-speed variant of Z.ai's GLM-5.3-Flash, a native multimodal model delivering inference speeds of up to 200 tokens/s. Built on the same hybrid sparse and l
Models
GLM 5.3 FlashX
Added 20 Sept 2026
Overview
GLM-5.3-FlashX is the high-speed variant of Z.ai's GLM-5.3-Flash, a native multimodal model delivering inference speeds of up to 200 tokens/s. Built on the same hybrid sparse and linear attention architecture...
How to use / API access
Call via OpenRouter with model id `z-ai/glm-5.3-flashx`.
- openrouter ·
z-ai/glm-5.3-flashxDocs
OpenRouter id: z-ai/glm-5.3-flashx
Notes
GLM 5.3 FlashX is listed in the EDNA Models directory from the OpenRouter catalogue.
Get the free Developer’s Field Guide
A 27-page field guide to the AI coding workflow with Claude. Claude Code, MCP servers, the prompt patterns that work, and what to delegate. Free.
Enter your work email. We send it straight over, plus a few short notes worth knowing. Unsubscribe any time.