Qwen frontier model tuned for agent frameworks, coding assistants, and long tasks
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
Google's proven reasoning model for coding, math, and multimodal analysis
Fast Gemini workhorse for multimodal apps where latency and price matter
Late GLM-4 workhorse for coding agents, reasoning, and structured tasks
Flagship Qwen3 model for coding agents, complex reasoning, and tool use
Hybrid-reasoning GLM release that made the 4.5 line broadly useful
Lighter GLM-4.5 variant for fast coding assistance and cheaper agents
Popular open Llama workhorse for multilingual chat, coding, and self-hosting
GLM vision model for visual reasoning, documents, and multimodal agents
Lean Gemini 2.5 lane for cheap multimodal traffic and quick agents
| Model | Creator | Score | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|---|
| Qwen3.7 Maxalibaba/qwen3.7-max | 53.5 | 1M | $2.5 | $7.5 | 2026-05-21 | |||
| Gemini 2.5 Progoogle/gemini-2.5-pro | 42.8 | 1.04858M | $1.25 | $10 | 2025-06-17 | |||
| Gemini 2.5 Flashgoogle/gemini-2.5-flash | 39.4 | 1.04858M | $0.3 | $2.5 | 2025-06-17 | |||
| GLM-4.6zhipuai/glm-4.6 | 38.4 | 204.8K | $0.6 | $2.2 | 2025-09-30 | |||
| Qwen3 Maxalibaba/qwen3-max | 38.3 | 262.144K | $1.2 | $6 | 2025-09-23 | |||
| GLM-4.5zhipuai/glm-4.5 | 34.8 | 131.072K | $0.6 | $2.2 | 2025-07-28 | |||
| GLM-4.5-Airzhipuai/glm-4.5-air | 30.6 | 131.072K | $0.2 | $1.1 | 2025-07-28 | |||
| Llama-3.3-70B-Instructmeta/llama-3.3-70b-instruct | 26.0 | 128K | $0.1 | $0.32 | 2024-12-06 | |||
| GLM-4.5Vzhipuai/glm-4.5v | 22.1 | 64K | $0.6 | $1.8 | 2025-08-11 | |||
| Gemini 2.5 Flash-Litegoogle/gemini-2.5-flash-lite | 19.3 | 1.04858M | $0.1 | $0.4 | 2025-06-17 |