Providers
Provider-specific identifiers, limits, and listed prices per million tokens. Every row links back to the provider's own documentation.
The model record exists, but no source-linked provider offer is available yet.
Submit a sourceGLM-5.3-Flash is a native multimodal model from Z.ai. It is suited for efficient coding and long-horizon agent tasks. Its hybrid sparse and linear attention architecture maintains accurate long-context behavior while...
Provider-specific identifiers, limits, and listed prices per million tokens. Every row links back to the provider's own documentation.
The model record exists, but no source-linked provider offer is available yet.
Submit a sourceRecorded from the source catalog and provider listings.
z-ai/glm-5.3-flashEvery result stays attached to its source, version, metric and harness. Scores from different versions are never merged, and the profile below plots each benchmark against its own population rather than on a shared scale.
Each axis is this model's percentile among the 8 benchmarks it has published results for, measured against every other model with a score on that same benchmark. Percentiles are used because benchmarks do not share a scale — a 60 on one is not a 60 on another. Hover any point for the raw score.
Each strip shows every published score for that benchmark, with this model marked. The lighter shape behind the ticks is the density of results, and the dashed line is the median.
Listed input price plotted against Coding Index, the benchmark with the widest published coverage that this model appears in.
The stepped line is the efficient frontier: at each price, the best score available for that money or less. A model sitting on it is not being beaten by anything cheaper. Price is log-scaled because listed rates span four orders of magnitude. Only models with both a listed price and a score on this benchmark can appear.
Field-level changes detected between successful source imports.
Every figure on this page traces back to one of these records.
Fetch the complete source-linked model record. No key, no account, no rate-limited tier.
GET https://model.kyssta.lol/api/v1/models/z-ai/glm-5.3-flashcurl "https://model.kyssta.lol/api/v1/models/z-ai/glm-5.3-flash"Answered directly from the stored record — nothing here is generated beyond the catalog's own fields.
GLM-5.3-Flash is a native multimodal model from Z.ai. It is suited for efficient coding and long-horizon agent tasks. Its hybrid sparse and linear attention architecture maintains accurate long-context behavior while. It is published by Z Ai and catalogued here from OpenRouter.
Z.ai: GLM 5.3 Flash accepts up to 1.04858M tokens of context and returns up to 131.072K output tokens.
Provider catalogs list support for image input.
GLM 5.2 is a large-scale reasoning model from Z.ai. It supports text input and output with a 1M-token context window, and is suited for long-horizon agent workflows, project-level software engineering,...
Z.ai: GLM 5.3GLM-5.3 is a large-scale reasoning model from Z.ai, built for complex software engineering and long-horizon agent tasks. It supports text input and output with a 1M-token context window, and improves...
Z.ai: GLM 5GLM-5 is Z.ai’s flagship open-source foundation model engineered for complex systems design and long-horizon agent workflows. Built for expert developers, it delivers production-grade performance on large-scale programming tasks, rivaling leading...
Z.ai: GLM 4.6Compared with GLM-4.5, this generation brings several key improvements: Longer context window: The context window has been expanded from 128K to 200K tokens, enabling the model to handle more complex...