Claude Opus 5 is Anthropic’s flagship model for demanding reasoning, coding, and long-horizon agentic work. It is particularly strong at end-to-end software tasks, code review and bug finding, visual analysis...
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
Claude Fable 5 is a Mythos-class model from Anthropic, built for autonomous knowledge work and coding. It supports text, image, and file inputs with text output, with reasoning support and...
Top Claude Opus tier for the hardest reasoning, coding, and long-horizon agents
Sonnet 5 is Anthropic's most capable Sonnet-class model, with frontier performance across coding, agents, and professional work. It supports adaptive thinking with selectable reasoning effort levels (low, medium, high, max,...
DeepSeek V4 Pro is a large-scale Mixture-of-Experts model from DeepSeek with 1.6T total parameters and 49B activated parameters, supporting a 1M-token context window. It is designed for advanced reasoning, coding,...
MiniMax multimodal model for long-context coding, perception, and agent planning
Qwen frontier model tuned for agent frameworks, coding assistants, and long tasks
Kimi K2.6 is Moonshot AI's next-generation multimodal model, designed for long-horizon coding, coding-driven UI/UX generation, and multi-agent orchestration. It handles complex end-to-end coding tasks across Python, Rust, and Go, and...
Open MiniMax flagship for coding agents, office automation, and complex environments
DeepSeek V4 Flash is an efficiency-optimized Mixture-of-Experts model from DeepSeek with 284B total parameters and 13B activated parameters, supporting a 1M-token context window. It is designed for fast inference and...
MiMo-V2.5-Pro is Xiaomi’s flagship model, delivering strong performance in general agentic capabilities, complex software engineering, and long-horizon tasks, with top rankings on benchmarks such as ClawEval, GDPVal, and SWE-bench Pro....
Hy3 is a 295B-parameter Mixture-of-Experts model from Tencent (21B active, 192 experts with top-8 routing) built for reasoning, agentic workflows, and real-world production use. It supports a configurable reasoning effort:...
Large open Qwen multimodal MoE for visual agents and long technical tasks
Prior MiniMax coding model for agent workflows, office edits, and automation
Earlier MiniMax agent model for practical coding and productivity tasks
Mature GLM model for dependable coding, reasoning, and structured agent tasks
Open multimodal Qwen MoE for local agents that need vision, audio, and code
General GLM flagship for coding, analysis, and tool-heavy engineering workflows
Kimi K2 Thinking is Moonshot AI’s most advanced open reasoning model to date, extending the K2 series into agentic, long-horizon reasoning. Built on the trillion-parameter Mixture-of-Experts (MoE) architecture introduced in...
Kimi K2.5 is Moonshot AI's native multimodal model, delivering state-of-the-art visual coding capability and a self-directed agent swarm paradigm. Built on Kimi K2 with continued pretraining over approximately 15T mixed...
NVIDIA Nemotron 3 Ultra is an open frontier-reasoning and orchestration model from NVIDIA, with 55B active parameters out of 550B total (MoE). Built on a hybrid Transformer-Mamba mixture-of-experts architecture, it...
Efficient open MiniMax model built for coding agents and tool-heavy workflows
Budget GLM lane for fast coding help, routing, and everyday automation
Mistral coding agent model for repository tasks and software engineering workflows
| Model | Creator | Score | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|---|
| Claude Opus 5anthropic/claude-opus-5 | 96.0 | 1M | $5 | $25 | 2026-07-24 | |||
| Anthropic: Claude Fable 5anthropic/claude-fable-5 | 95.0 | 1M | $10 | $50 | 2026-06-09 | |||
| Claude Opus 4.8anthropic/claude-opus-4-8 | 88.6 | 1M | $5 | $25 | 2026-05-28 | |||
| Anthropic: Claude Sonnet 5anthropic/claude-sonnet-5 | 85.2 | 1M | $2 | $10 | 2026-06-30 | |||
| DeepSeek: DeepSeek V4 Pro 0423deepseek/deepseek-v4-pro | 80.6 | 1.024M | $0.948 | $1.896 | 2026-04-24 | |||
| MiniMax-M3minimax/MiniMax-M3 | 80.5 | 1.04858M | $0.3 | $1.2 | 2026-06-01 | |||
| Qwen3.7 Maxalibaba/qwen3.7-max | 80.4 | 1M | $2.5 | $7.5 | 2026-05-21 | |||
| MoonshotAI: Kimi K2.6moonshotai/kimi-k2.6 | 80.2 | 262.144K | $0.95 | $4 | 2026-04-21 | |||
| MiniMax-M2.7minimax/MiniMax-M2.7 | 79.9 | 204.8K | $0.3 | $1.2 | 2026-03-18 | |||
| DeepSeek: DeepSeek V4 Flash 0423deepseek/deepseek-v4-flash | 79.0 | 1.024M | $0.086 | $0.171 | 2026-04-24 | |||
| Xiaomi: MiMo-V2.5-Proxiaomi/mimo-v2.5-pro | 78.9 | 1.04858M | $0.435 | $0.87 | 2026-04-22 | |||
| Tencent: Hy3tencent/hy3 | 78.0 | 262.144K | $0.083 | $0.33 | 2026-07-06 | |||
| Qwen3.5 397B-A17Balibaba/qwen3.5-397b-a17b | 76.4 | 262.144K | $0.6 | $3.6 | 2026-02-15 | |||
| MiniMax-M2.5minimax/MiniMax-M2.5 | 75.8 | 204.8K | $0.3 | $1.2 | 2026-02-12 | |||
| MiniMax-M2.1minimax/MiniMax-M2.1 | 74.0 | 204.8K | $0.3 | $1.2 | 2025-12-23 | |||
| GLM-4.7zhipuai/glm-4.7 | 73.8 | 204.8K | $0.6 | $2.2 | 2025-12-22 | |||
| Qwen3.6 35B-A3Balibaba/qwen3.6-35b-a3b | 73.4 | 262.144K | $0.248 | $1.485 | 2026-04-17 | |||
| GLM-5zhipuai/glm-5 | 72.8 | 204.8K | $1 | $3.2 | 2026-02-12 | |||
| MoonshotAI: Kimi K2 Thinkingmoonshotai/kimi-k2-thinking | 71.3 | 262.144K | $0.6 | $2.5 | 2025-11-06 | |||
| MoonshotAI: Kimi K2.5moonshotai/kimi-k2.5 | 70.8 | 262.144K | $0.45 | $2.25 | 2026-01 | |||
| NVIDIA: Nemotron 3 Ultranvidia/nemotron-3-ultra-550b-a55b | 70.7 | 256K | $0.625 | $3.125 | 2026-06-04 | |||
| MiniMax-M2minimax/MiniMax-M2 | 69.4 | 204.8K | $0.3 | $1.2 | 2025-10-27 | |||
| GLM-4.7-Flashzhipuai/glm-4.7-flash | 59.2 | 200K | $0.06 | $0.4 | 2026-01-19 | |||
| Devstral Smallmistral/devstral-small-2507 | 53.6 | 128K | $0.1 | $0.3 | 2025-07-10 |