Claude Opus 5 is Anthropic’s flagship model for demanding reasoning, coding, and long-horizon agentic work. It is particularly strong at end-to-end software tasks, code review and bug finding, visual analysis...
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
Claude Fable 5 is a Mythos-class model from Anthropic, built for autonomous knowledge work and coding. It supports text, image, and file inputs with text output, with reasoning support and...
Top Claude Opus tier for the hardest reasoning, coding, and long-horizon agents
Sonnet 5 is Anthropic's most capable Sonnet-class model, with frontier performance across coding, agents, and professional work. It supports adaptive thinking with selectable reasoning effort levels (low, medium, high, max,...
DeepSeek V4 Pro is a large-scale Mixture-of-Experts model from DeepSeek with 1.6T total parameters and 49B activated parameters, supporting a 1M-token context window. It is designed for advanced reasoning, coding,...
MiniMax multimodal model for long-context coding, perception, and agent planning
Qwen frontier model tuned for agent frameworks, coding assistants, and long tasks
Kimi K2.6 is Moonshot AI's next-generation multimodal model, designed for long-horizon coding, coding-driven UI/UX generation, and multi-agent orchestration. It handles complex end-to-end coding tasks across Python, Rust, and Go, and...
Open MiniMax flagship for coding agents, office automation, and complex environments
DeepSeek V4 Flash is an efficiency-optimized Mixture-of-Experts model from DeepSeek with 284B total parameters and 13B activated parameters, supporting a 1M-token context window. It is designed for fast inference and...
MiMo-V2.5-Pro is Xiaomi’s flagship model, delivering strong performance in general agentic capabilities, complex software engineering, and long-horizon tasks, with top rankings on benchmarks such as ClawEval, GDPVal, and SWE-bench Pro....
Hy3 is a 295B-parameter Mixture-of-Experts model from Tencent (21B active, 192 experts with top-8 routing) built for reasoning, agentic workflows, and real-world production use. It supports a configurable reasoning effort:...
Balanced Mistral model for enterprise assistants, multilingual work, and tools
Step 3.7 Flash is StepFun's latest high-efficiency multimodal Mixture-of-Experts model. It pairs a 196B-parameter language backbone with a vision encoder for native image and video understanding, activating roughly 11B parameters...
Muse Glimmer 30B is a dense, open-weight multimodal model from Meta Superintelligence Labs, distilled from Muse Spark and optimized for autonomous agents on consumer hardware. It is suited for long-horizon...
Qwen vision-language model for visual reasoning, documents, and agent tasks
NVIDIA Nemotron 3 Ultra is an open frontier-reasoning and orchestration model from NVIDIA, with 55B active parameters out of 550B total (MoE). Built on a hybrid Transformer-Mamba mixture-of-experts architecture, it...
| Model | Creator | Score | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|---|
| Claude Opus 5anthropic/claude-opus-5 | 96.0 | 1M | $5 | $25 | 2026-07-24 | |||
| Anthropic: Claude Fable 5anthropic/claude-fable-5 | 95.0 | 1M | $10 | $50 | 2026-06-09 | |||
| Claude Opus 4.8anthropic/claude-opus-4-8 | 88.6 | 1M | $5 | $25 | 2026-05-28 | |||
| Anthropic: Claude Sonnet 5anthropic/claude-sonnet-5 | 85.2 | 1M | $2 | $10 | 2026-06-30 | |||
| DeepSeek: DeepSeek V4 Pro 0423deepseek/deepseek-v4-pro | 80.6 | 1.024M | $0.86 | $1.72 | 2026-04-24 | |||
| MiniMax-M3minimax/MiniMax-M3 | 80.5 | 1.04858M | $0.3 | $1.2 | 2026-06-01 | |||
| Qwen3.7 Maxalibaba/qwen3.7-max | 80.4 | 1M | $2.5 | $7.5 | 2026-05-21 | |||
| MoonshotAI: Kimi K2.6moonshotai/kimi-k2.6 | 80.2 | 262.144K | $0.95 | $4 | 2026-04-21 | |||
| MiniMax-M2.7minimax/MiniMax-M2.7 | 79.9 | 204.8K | $0.3 | $1.2 | 2026-03-18 | |||
| DeepSeek: DeepSeek V4 Flash 0423deepseek/deepseek-v4-flash | 79.0 | 1.024M | $0.086 | $0.172 | 2026-04-24 | |||
| Xiaomi: MiMo-V2.5-Proxiaomi/mimo-v2.5-pro | 78.9 | 1.04858M | $0.435 | $0.87 | 2026-04-22 | |||
| Tencent: Hy3tencent/hy3 | 78.0 | 262.144K | $0.083 | $0.33 | 2026-07-06 | |||
| Mistral Medium (latest)mistral/mistral-medium-latest | 77.6 | 262.144K | $1.5 | $7.5 | 2026-04-29 | |||
| StepFun: Step 3.7 Flashstepfun/step-3.7-flash | 76.5 | 256K | $0.2 | $1.15 | 2026-05-29 | |||
| Meta: Muse Glimmer 30Bmeta/muse-glimmer-30b | 76.0 | 131.072K | $0.3 | $1.1 | 2026-08-10 | |||
| Qwen3.5 27Balibaba/qwen3.5-27b | 72.4 | 262.144K | $0.3 | $2.4 | 2026-02-23 | |||
| NVIDIA: Nemotron 3 Ultranvidia/nemotron-3-ultra-550b-a55b | 70.7 | 256K | $0.625 | $3.125 | 2026-06-04 |