Open flagship GLM for long-horizon coding agents and million-token context work
72 providers
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
Open flagship GLM for long-horizon coding agents and million-token context work
Open Nemotron omni model combining reasoning with text, vision, and audio
Multimodal Kimi workhorse for agent loops, coding tasks, and visual context
Strong GLM coding model for agentic engineering, terminals, and repository generation
Largest Gemma 4 instruction model for open, self-hosted chat and reasoning
Nemotron middle tier for collaborative agents and high-volume reasoning workloads
Small Nemotron 3 MoE for efficient coding, math, and long-context agents
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| GLM-5.2zhipuai/glm-5.2 | 1M | $1.4 | $4.4 | 2026-06-13 | |||
| Nemotron 3 Nano Omni 30B A3B Reasoningnvidia/nemotron-3-nano-omni-30b-a3b-reasoning | 256K | $0.2 | $0.8 | 2026-04-28 | |||
| Kimi K2.6moonshotai/kimi-k2.6 | 262.144K | $0.95 | $4 | 2026-04-21 | |||
| GLM-5.1zhipuai/glm-5.1 | 200K | $1.4 | $4.4 | 2026-04-07 | |||
| Gemma 4 31B ITgoogle/gemma-4-31b-it | 262.144K | $0.09 | $0.34 | 2026-04-02 | |||
| Nemotron 3 Super 120B A12Bnvidia/nemotron-3-super-120b-a12b | 262.144K | $0.2 | $0.8 | 2026-03-11 | |||
| Nemotron 3 Nano 30B A3Bnvidia/nemotron-3-nano-30b-a3b | 262.144K | $0.05 | $0.2 | 2025-12-15 |