Z.ai
z-ai/glm-5.3-flashx
Published by the vendor; the platform bills at the target's declared price.
Z.ai z-ai/glm-5.3-flashx
GLM-5.3-FlashX is the high-speed variant of Z.ai's GLM-5.3-Flash, a native multimodal model delivering inference speeds of up to 200 tokens/s. Built on the same hybrid sparse and linear attention architecture...
Cached input reads $0.07 per 1M
No routing target resolves to this model yet. Declare one and the router picks it up without retraining.
# targets.yaml - declare a target that runs GLM 5.3 FlashX
targets:
- id: glm-5-3-flashx
kind: llm
name: GLM 5.3 FlashX
metadata:
model: z-ai/glm-5.3-flashx # links the target to this catalogue entry
capabilities:
domains: [general]
supports_tools: true
context_window: 1048576
cost:
usd_per_1k_tokens: 0.000810Input list price per million tokens; GLM 5.3 FlashX is the 5th cheapest of 13 listed.
Other models by the same vendor in the catalogue.