Inference availability
Providers Provider-specific identifiers, limits, and listed prices per million tokens. Every row links back to the provider's own documentation.
Report a price
Providers offering GLM-4.5-Flash
Provider Provider model ID Context Max output Input Output Cache read Capabilities Docs
U UnoRouter
glm-4.5-flash:free
131.072K
98.304K
—
—
—
Reasoning Tools
Docs ↗
Z Zhipu AI
glm-4.5-flash
131.072K
98.304K
—
—
—
Reasoning Tools
Docs ↗
Z Z.AI
glm-4.5-flash
131.072K
98.304K
—
—
—
Reasoning Tools
Docs ↗
Capability badges appear only where the provider catalog explicitly lists support. A blank cell means the source is silent, not that the feature is absent.
Specification
Capabilities Recorded from the source catalog and provider listings.
✓
Reasoning
Yes
✓
Tool calling
Yes
?
Structured output
Unknown
×
Attachments
No
×
Vision input
No
×
Open weights
No
Model family glm-flash
Knowledge cutoff 2025-04
License Not documented
Release date 2025-07-28
Model ID zhipuai/glm-4.5-flash
Published evaluations
Benchmarks Every result stays attached to its source, version, metric and harness. Scores from different versions are never merged, and the profile below plots each benchmark against its own population rather than on a shared scale.
Benchmark registry
No published benchmark results ModelBench does not infer quality from price, context size, or model name. When a source publishes a comparable result, it appears here with its version and link.
Read the methodology
Catalog activity
Change log Field-level changes detected between successful source imports.
Full change log
No changes recorded This record has not changed within the retained import history.
Provenance
Sources & verification Every figure on this page traces back to one of these records.
Methodology
Last verified Sep 11, 2026
Status Source-linked
Confidence Medium
Catalog source Models.dev
Public API
Use this record Fetch the complete source-linked model record. No key, no account, no rate-limited tier.
API documentation
Endpoint Copy
GET https://model.kyssta.lol/api/v1/models/zhipuai/glm-4.5-flash
curl Copy
curl "https://model.kyssta.lol/api/v1/models/zhipuai/glm-4.5-flash"
Common questions
Frequently asked questions Answered directly from the stored record — nothing here is generated beyond the catalog's own fields.
What is GLM-4.5-Flash?
Efficient GLM model for fast reasoning, coding, and agent workflows. It is published by Zhipu AI and catalogued here from Models.dev.
What is the context length of GLM-4.5-Flash?
GLM-4.5-Flash accepts up to 131.072K tokens of context and returns up to 98.304K output tokens.
Does GLM-4.5-Flash support tool calling and structured output?
Provider catalogs list support for tool calling, and reasoning.
Which providers serve GLM-4.5-Flash?
3 providers list this model: UnoRouter, Zhipu AI, Z.AI.
Are the weights for GLM-4.5-Flash open?
No. This model is served through hosted APIs only.
When was GLM-4.5-Flash released?
The catalog records a release date of 2025-07-28, last verified Sep 11, 2026.