Open flagship GLM for long-horizon coding agents and million-token context work
Model catalog
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
72 providers
Largest Nemotron 3 model for maximum open-weight reasoning and agent accuracy
15 providers
Newer StepFun flash model for faster agents, coding, and multimodal prompts
17 providers
Fast DeepSeek V4 lane for economical reasoning, coding, and long-context work
62 providers
Open MoE flagship with million-token context for coding and long agent runs
62 providers
Multimodal Kimi workhorse for agent loops, coding tasks, and visual context
64 providers
Largest Gemma 4 instruction model for open, self-hosted chat and reasoning
33 providers
Open MiniMax flagship for coding agents, office automation, and complex environments
35 providers
Large open Qwen multimodal MoE for visual agents and long technical tasks
18 providers
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| GLM-5.2zhipuai/glm-5.2 | 1M | $1.4 | $4.4 | 2026-06-13 | |||
| Nemotron 3 Ultra 550B A55Bnvidia/nemotron-3-ultra-550b-a55b | 1M | $0.5 | $2.5 | 2026-06-04 | |||
| Step 3.7 Flashstepfun/step-3.7-flash | 256K | $0.185 | $1.11 | 2026-05-29 | |||
| DeepSeek V4 Flashdeepseek/deepseek-v4-flash | 1M | $0.15 | $0.6 | 2026-04-24 | |||
| DeepSeek V4 Prodeepseek/deepseek-v4-pro | 1M | $0.435 | $0.87 | 2026-04-24 | |||
| Kimi K2.6moonshotai/kimi-k2.6 | 262.144K | $0.95 | $4 | 2026-04-21 | |||
| Gemma 4 31B ITgoogle/gemma-4-31b-it | 262.144K | $0.09 | $0.34 | 2026-04-02 | |||
| MiniMax-M2.7minimax/MiniMax-M2.7 | 204.8K | $0.3 | $1.2 | 2026-03-18 | |||
| Qwen3.5 397B-A17Balibaba/qwen3.5-397b-a17b | 262.144K | $0.6 | $3.6 | 2026-02-15 |