Claude Opus 5 is Anthropic’s flagship model for demanding reasoning, coding, and long-horizon agentic work. It is particularly strong at end-to-end software tasks, code review and bug finding, visual analysis...
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
Claude Fable 5 is a Mythos-class model from Anthropic, built for autonomous knowledge work and coding. It supports text, image, and file inputs with text output, with reasoning support and...
Top Claude Opus tier for the hardest reasoning, coding, and long-horizon agents
Sonnet 5 is Anthropic's most capable Sonnet-class model, with frontier performance across coding, agents, and professional work. It supports adaptive thinking with selectable reasoning effort levels (low, medium, high, max,...
DeepSeek V4 Pro is a large-scale Mixture-of-Experts model from DeepSeek with 1.6T total parameters and 49B activated parameters, supporting a 1M-token context window. It is designed for advanced reasoning, coding,...
Qwen frontier model tuned for agent frameworks, coding assistants, and long tasks
Kimi K2.6 is Moonshot AI's next-generation multimodal model, designed for long-horizon coding, coding-driven UI/UX generation, and multi-agent orchestration. It handles complex end-to-end coding tasks across Python, Rust, and Go, and...
DeepSeek V4 Flash is an efficiency-optimized Mixture-of-Experts model from DeepSeek with 284B total parameters and 13B activated parameters, supporting a 1M-token context window. It is designed for fast inference and...
MiMo-V2.5-Pro is Xiaomi’s flagship model, delivering strong performance in general agentic capabilities, complex software engineering, and long-horizon tasks, with top rankings on benchmarks such as ClawEval, GDPVal, and SWE-bench Pro....
Hy3 is a 295B-parameter Mixture-of-Experts model from Tencent (21B active, 192 experts with top-8 routing) built for reasoning, agentic workflows, and real-world production use. It supports a configurable reasoning effort:...
Balanced Mistral model for enterprise assistants, multilingual work, and tools
Qwen vision-language model for visual reasoning, documents, and agent tasks
Muse Glimmer 30B is a dense, open-weight multimodal model from Meta Superintelligence Labs, distilled from Muse Spark and optimized for autonomous agents on consumer hardware. It is suited for long-horizon...
Hy3 preview is a high-efficiency Mixture-of-Experts model from Tencent designed for agentic workflows and production use. It supports configurable reasoning levels across disabled, low, and high modes, allowing it to...
Step 3.5 Flash is StepFun's most capable open-source foundation model. Built on a sparse Mixture of Experts (MoE) architecture, it selectively activates only 11B of its 196B parameters per token....
Qwen vision-language model for visual reasoning, documents, and agent tasks
Kimi K2 Thinking is Moonshot AI’s most advanced open reasoning model to date, extending the K2 series into agentic, long-horizon reasoning. Built on the trillion-parameter Mixture-of-Experts (MoE) architecture introduced in...
Laguna XS 2.1 is the latest coding agent model in the 33B-A3B category from [Poolside](https://poolside.ai/) and a step forward from their Laguna XS.2 model (released in April 2026). It combines...
Kimi K2.5 is Moonshot AI's native multimodal model, delivering state-of-the-art visual coding capability and a self-directed agent swarm paradigm. Built on Kimi K2 with continued pretraining over approximately 15T mixed...
NVIDIA Nemotron 3 Ultra is an open frontier-reasoning and orchestration model from NVIDIA, with 55B active parameters out of 550B total (MoE). Built on a hybrid Transformer-Mamba mixture-of-experts architecture, it...
| Model | Creator | Score | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|---|
| Claude Opus 5anthropic/claude-opus-5 | 96.0 | 1M | $5 | $25 | 2026-07-24 | |||
| Anthropic: Claude Fable 5anthropic/claude-fable-5 | 95.0 | 1M | $10 | $50 | 2026-06-09 | |||
| Claude Opus 4.8anthropic/claude-opus-4-8 | 88.6 | 1M | $5 | $25 | 2026-05-28 | |||
| Anthropic: Claude Sonnet 5anthropic/claude-sonnet-5 | 85.2 | 1M | $2 | $10 | 2026-06-30 | |||
| DeepSeek: DeepSeek V4 Pro 0423deepseek/deepseek-v4-pro | 80.6 | 1.024M | $0.948 | $1.896 | 2026-04-24 | |||
| Qwen3.7 Maxalibaba/qwen3.7-max | 80.4 | 1M | $2.5 | $7.5 | 2026-05-21 | |||
| MoonshotAI: Kimi K2.6moonshotai/kimi-k2.6 | 80.2 | 262.144K | $0.95 | $4 | 2026-04-21 | |||
| DeepSeek: DeepSeek V4 Flash 0423deepseek/deepseek-v4-flash | 79.0 | 1.024M | $0.085 | $0.17 | 2026-04-24 | |||
| Xiaomi: MiMo-V2.5-Proxiaomi/mimo-v2.5-pro | 78.9 | 1.04858M | $0.435 | $0.87 | 2026-04-22 | |||
| Tencent: Hy3tencent/hy3 | 78.0 | 262.144K | $0.083 | $0.33 | 2026-07-06 | |||
| Mistral Medium 3.5mistral/mistral-medium-2604 | 77.6 | 262.144K | $1.5 | $7.5 | 2026-04-29 | |||
| Qwen3.6 27Balibaba/qwen3.6-27b | 77.2 | 262.144K | $0.6 | $3.6 | 2026-04-22 | |||
| Meta: Muse Glimmer 30Bmeta/muse-glimmer-30b | 76.0 | 131.072K | $0.3 | $1.1 | 2026-08-10 | |||
| Tencent: Hy3 previewtencent/hy3-preview | 74.4 | 262.144K | $0.18 | $0.6 | 2026-04-20 | |||
| StepFun: Step 3.5 Flashstepfun/step-3.5-flash | 74.4 | 262.144K | $0.1 | $0.3 | 2026-01-29 | |||
| Qwen3.5 27Balibaba/qwen3.5-27b | 72.4 | 262.144K | $0.3 | $2.4 | 2026-02-23 | |||
| MoonshotAI: Kimi K2 Thinkingmoonshotai/kimi-k2-thinking | 71.3 | 262.144K | $0.6 | $2.5 | 2025-11-06 | |||
| Poolside: Laguna XS 2.1poolside/laguna-xs-2.1 | 70.9 | 262.144K | $0.06 | $0.12 | 2026-07-02 | |||
| MoonshotAI: Kimi K2.5moonshotai/kimi-k2.5 | 70.8 | 262.144K | $0.45 | $2.25 | 2026-01 | |||
| NVIDIA: Nemotron 3 Ultranvidia/nemotron-3-ultra-550b-a55b | 70.7 | 256K | $0.625 | $3.125 | 2026-06-04 |