Zhipu AI logo

Zhipu AI: GLM-4.7-Flash

Budget GLM lane for fast coding help, routing, and everyday automation

Source-linked glm-flash Open weights Released 2026-01-19
API record Report
InputT
OutputT
Input price$0.06/M
Output price$0.4/M
Context200K
Max output131.072K
Providers16
Inference availability

Providers

Provider-specific identifiers, limits, and listed prices per million tokens. Every row links back to the provider's own documentation.

Report a price
Providers offering GLM-4.7-Flash
ProviderProvider model IDContextMax outputInputOutputCache readCapabilitiesDocs
Pendra glm-4.7-flash 200K 131.072K ReasoningTools Docs ↗
Z.AI glm-4.7-flash 200K 131.072K ReasoningTools Docs ↗
Zhipu AI glm-4.7-flash 200K 131.072K ReasoningTools Docs ↗
CrofAI glm-4.7-flash 202.752K 131.072K $0.04 $0.3 $0.008 ReasoningTools Docs ↗
Deep Infra zai-org/GLM-4.7-Flash 202.752K 16.384K $0.06 $0.4 $0.01 ReasoningToolsJSON Docs ↗
Eden AI deepinfra/zai-org/GLM-4.7-Flash 202.752K 131.072K $0.06 $0.4 $0.01 ReasoningToolsJSON Docs ↗
DevPass (LLM Gateway) glm-4.7-flash 200K 131.072K $0.06 $0.4 $0.01 ReasoningTools Docs ↗
Cloudflare AI Gateway workers-ai/@cf/zai-org/glm-4.7-flash 131.072K 131.072K $0.06 $0.4 ReasoningToolsJSON Docs ↗
Cloudflare Workers AI @cf/zai-org/glm-4.7-flash 131.072K 131.072K $0.06 $0.4 ReasoningToolsJSON Docs ↗
Eden AI cloudflare/@cf/zai-org/glm-4.7-flash 131.072K 131.072K $0.06 $0.4 ReasoningToolsJSON Docs ↗
OpenRouter z-ai/glm-4.7-flash 200K 117.964K $0.06 $0.4 ReasoningToolsJSON Docs ↗
Kilo Gateway z-ai/glm-4.7-flash 131.072K 117.964K $0.06 $0.4 ReasoningToolsJSON Docs ↗
Eden AI amazon/zai.glm-4.7-flash 200K 131.072K $0.07 $0.4 ReasoningTools Docs ↗
Amazon Bedrock zai.glm-4.7-flash 200K 131.072K $0.07 $0.4 ReasoningToolsJSON Docs ↗
Cortecs glm-4.7-flash 203K 203K $0.08 $0.478 ReasoningToolsJSON Docs ↗
Synthetic hf:zai-org/GLM-4.7-Flash 196.608K 65.536K $0.1 $0.5 $0.1 ReasoningTools Docs ↗

Capability badges appear only where the provider catalog explicitly lists support. A blank cell means the source is silent, not that the feature is absent.

Listed rates

Price across providers

Input price per million tokens as published by each provider. Bars are drawn from listed rates only — no traffic weighting, since the catalog observes no requests.

Lowest input $0.04/M

Across 13 priced providers

Median input $0.06/M

Midpoint of listed rates

Highest input $0.1/M

2.5× the lowest listed rate

Output range $0.3 – $0.5

Per million output tokens

CrofAI $0.04/MLowest
Deep Infra $0.06/M
Eden AI $0.06/M
DevPass (LLM Gateway) $0.06/M
Cloudflare AI Gateway $0.06/M
Cloudflare Workers AI $0.06/M
Eden AI $0.06/M
OpenRouter $0.06/M
Kilo Gateway $0.06/M
Eden AI $0.07/M
Amazon Bedrock $0.07/M
Cortecs $0.08/M
Synthetic $0.1/M
Cost calculator

Estimate a workload

$0.00
Excludes taxes, non-token charges, and tiered discounts.

Context limits also differ by provider, from 131.072K to 203K tokens. Compare the provider table above before choosing on price alone.

Specification

Capabilities

Recorded from the source catalog and provider listings.

Reasoning Yes
Tool calling Yes
? Structured output Unknown
× Attachments No
× Vision input No
Open weights Yes
Creator
Zhipu AI
Model family
glm-flash
Knowledge cutoff
2025-04
License
Not documented
Release date
2026-01-19
Model ID
zhipuai/glm-4.7-flash

Weights: Hugging Face ↗

Published evaluations

Benchmarks

Every result stays attached to its source, version, metric and harness. Scores from different versions are never merged, and the profile below plots each benchmark against its own population rather than on a shared scale.

Benchmark registry
SWE-Bench Verified resolved Source ↗
SWE-Bench Verified #31 of 32

Each strip shows every published score for that benchmark, with this model marked. The lighter shape behind the ticks is the density of results, and the dashed line is the median.

Cost against capability

Price and performance

Listed input price plotted against SWE-Bench Verified, the benchmark with the widest published coverage that this model appears in.

0 51 102 $0.1 $1 $10 Qwen3.5 122B-A10B — $0.4/M, score 72.0 Qwen3.5 27B — $0.3/M, score 72.4 Qwen3.5 397B-A17B — $0.6/M, score 76.4 Qwen3.6 27B — $0.6/M, score 77.2 Qwen3.6 35B-A3B — $0.248/M, score 73.4 Qwen3.7 Max — $2.5/M, score 80.4 Claude Fable 5 — $10/M, score 95.0 Claude Opus 4.8 — $5/M, score 88.6 Claude Opus 5 — $5/M, score 96.0 Claude Sonnet 5 — $2/M, score 85.2 DeepSeek V4 Flash — $0.15/M, score 79.0 DeepSeek V4 Pro — $0.435/M, score 80.6 MAI-Code-1-Flash — $0.75/M, score 71.6 MiniMax-M2 — $0.3/M, score 69.4 MiniMax-M2.1 — $0.3/M, score 74.0 MiniMax-M2.5 — $0.3/M, score 75.8 Devstral Medium — $0.4/M, score 61.6 Devstral Small — $0.1/M, score 53.6 Mistral Medium 3.5 — $1.5/M, score 77.6 Mistral Medium (latest) — $1.5/M, score 77.6 Kimi K2 Thinking — $0.4/M, score 71.3 Kimi K2.5 — $0.3/M, score 70.8 Kimi K2.6 — $0.95/M, score 80.2 Nemotron 3 Ultra 550B A55B — $0.5/M, score 70.7 Step 3.5 Flash — $0.1/M, score 74.4 Step 3.7 Flash — $0.185/M, score 76.5 Hy3 — $0.066/M, score 78.0 Hy3 preview — $0.066/M, score 74.4 MiMo-V2.5-Pro — $0.435/M, score 78.9 GLM-4.7 — $0.6/M, score 73.8 GLM-4.7-Flash — $0.06/M, score 59.2 GLM-5 — $1/M, score 72.8 GLM-4.7-Flash Input price per million tokens (log scale) Resolved

The stepped line is the efficient frontier: at each price, the best score available for that money or less. A model sitting on it is not being beaten by anything cheaper. Price is log-scaled because listed rates span four orders of magnitude. Only models with both a listed price and a score on this benchmark can appear.

Catalog activity

Change log

Field-level changes detected between successful source imports.

Full change log
Price Completion0.3 → 0.4
Price Prompt0.04 → 0.06
Provenance

Sources & verification

Every figure on this page traces back to one of these records.

Methodology
Public API

Use this record

Fetch the complete source-linked model record. No key, no account, no rate-limited tier.

API documentation
Endpoint
GET https://model.kyssta.lol/api/v1/models/zhipuai/glm-4.7-flash
curl
curl "https://model.kyssta.lol/api/v1/models/zhipuai/glm-4.7-flash"
Common questions

Frequently asked questions

Answered directly from the stored record — nothing here is generated beyond the catalog's own fields.

What is GLM-4.7-Flash?

Budget GLM lane for fast coding help, routing, and everyday automation. It is published by Zhipu AI and catalogued here from Models.dev.

How much does GLM-4.7-Flash cost?

Listed input pricing starts at $0.04 per million tokens from CrofAI, rising to $0.1 across 13 listed providers.

What is the context length of GLM-4.7-Flash?

GLM-4.7-Flash accepts up to 200K tokens of context and returns up to 131.072K output tokens.

Does GLM-4.7-Flash support tool calling and structured output?

Provider catalogs list support for tool calling, and reasoning.

Which providers serve GLM-4.7-Flash?

14 providers list this model: Pendra, Z.AI, Zhipu AI, CrofAI, Deep Infra, Eden AI and 8 more.

Are the weights for GLM-4.7-Flash open?

Yes. The weights are published and downloadable from Hugging Face.

When was GLM-4.7-Flash released?

The catalog records a release date of 2026-01-19, last verified Sep 11, 2026.

More models from Zhipu AI

View all →