Claude Opus 5 is Anthropic’s flagship model for demanding reasoning, coding, and long-horizon agentic work. It is particularly strong at end-to-end software tasks, code review and bug finding, visual analysis...
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
Claude Fable 5 is a Mythos-class model from Anthropic, built for autonomous knowledge work and coding. It supports text, image, and file inputs with text output, with reasoning support and...
Top Claude Opus tier for the hardest reasoning, coding, and long-horizon agents
Sonnet 5 is Anthropic's most capable Sonnet-class model, with frontier performance across coding, agents, and professional work. It supports adaptive thinking with selectable reasoning effort levels (low, medium, high, max,...
DeepSeek V4 Pro is a large-scale Mixture-of-Experts model from DeepSeek with 1.6T total parameters and 49B activated parameters, supporting a 1M-token context window. It is designed for advanced reasoning, coding,...
Qwen frontier model tuned for agent frameworks, coding assistants, and long tasks
Kimi K2.6 is Moonshot AI's next-generation multimodal model, designed for long-horizon coding, coding-driven UI/UX generation, and multi-agent orchestration. It handles complex end-to-end coding tasks across Python, Rust, and Go, and...
DeepSeek V4 Flash is an efficiency-optimized Mixture-of-Experts model from DeepSeek with 284B total parameters and 13B activated parameters, supporting a 1M-token context window. It is designed for fast inference and...
MiMo-V2.5-Pro is Xiaomi’s flagship model, delivering strong performance in general agentic capabilities, complex software engineering, and long-horizon tasks, with top rankings on benchmarks such as ClawEval, GDPVal, and SWE-bench Pro....
Hy3 is a 295B-parameter Mixture-of-Experts model from Tencent (21B active, 192 experts with top-8 routing) built for reasoning, agentic workflows, and real-world production use. It supports a configurable reasoning effort:...
Qwen vision-language model for visual reasoning, documents, and agent tasks
Step 3.7 Flash is StepFun's latest high-efficiency multimodal Mixture-of-Experts model. It pairs a 196B-parameter language backbone with a vision encoder for native image and video understanding, activating roughly 11B parameters...
Muse Glimmer 30B is a dense, open-weight multimodal model from Meta Superintelligence Labs, distilled from Muse Spark and optimized for autonomous agents on consumer hardware. It is suited for long-horizon...
Step 3.5 Flash is StepFun's most capable open-source foundation model. Built on a sparse Mixture of Experts (MoE) architecture, it selectively activates only 11B of its 196B parameters per token....
General GLM flagship for coding, analysis, and tool-heavy engineering workflows
Kimi K2 Thinking is Moonshot AI’s most advanced open reasoning model to date, extending the K2 series into agentic, long-horizon reasoning. Built on the trillion-parameter Mixture-of-Experts (MoE) architecture introduced in...
Kimi K2.5 is Moonshot AI's native multimodal model, delivering state-of-the-art visual coding capability and a self-directed agent swarm paradigm. Built on Kimi K2 with continued pretraining over approximately 15T mixed...
NVIDIA Nemotron 3 Ultra is an open frontier-reasoning and orchestration model from NVIDIA, with 55B active parameters out of 550B total (MoE). Built on a hybrid Transformer-Mamba mixture-of-experts architecture, it...
| Model | Creator | Score | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|---|
| Claude Opus 5anthropic/claude-opus-5 | 96.0 | 1M | $5 | $25 | 2026-07-24 | |||
| Anthropic: Claude Fable 5anthropic/claude-fable-5 | 95.0 | 1M | $10 | $50 | 2026-06-09 | |||
| Claude Opus 4.8anthropic/claude-opus-4-8 | 88.6 | 1M | $5 | $25 | 2026-05-28 | |||
| Anthropic: Claude Sonnet 5anthropic/claude-sonnet-5 | 85.2 | 1M | $2 | $10 | 2026-06-30 | |||
| DeepSeek: DeepSeek V4 Pro 0423deepseek/deepseek-v4-pro | 80.6 | 1.024M | $0.86 | $1.72 | 2026-04-24 | |||
| Qwen3.7 Maxalibaba/qwen3.7-max | 80.4 | 1M | $2.5 | $7.5 | 2026-05-21 | |||
| MoonshotAI: Kimi K2.6moonshotai/kimi-k2.6 | 80.2 | 262.144K | $0.95 | $4 | 2026-04-21 | |||
| DeepSeek: DeepSeek V4 Flash 0423deepseek/deepseek-v4-flash | 79.0 | 1.024M | $0.086 | $0.172 | 2026-04-24 | |||
| Xiaomi: MiMo-V2.5-Proxiaomi/mimo-v2.5-pro | 78.9 | 1.04858M | $0.435 | $0.87 | 2026-04-22 | |||
| Tencent: Hy3tencent/hy3 | 78.0 | 262.144K | $0.083 | $0.33 | 2026-07-06 | |||
| Qwen3.6 27Balibaba/qwen3.6-27b | 77.2 | 262.144K | $0.6 | $3.6 | 2026-04-22 | |||
| StepFun: Step 3.7 Flashstepfun/step-3.7-flash | 76.5 | 256K | $0.2 | $1.15 | 2026-05-29 | |||
| Meta: Muse Glimmer 30Bmeta/muse-glimmer-30b | 76.0 | 131.072K | $0.3 | $1.1 | 2026-08-10 | |||
| StepFun: Step 3.5 Flashstepfun/step-3.5-flash | 74.4 | 262.144K | $0.1 | $0.3 | 2026-01-29 | |||
| GLM-5zhipuai/glm-5 | 72.8 | 204.8K | $1 | $3.2 | 2026-02-12 | |||
| MoonshotAI: Kimi K2 Thinkingmoonshotai/kimi-k2-thinking | 71.3 | 262.144K | $0.6 | $2.5 | 2025-11-06 | |||
| MoonshotAI: Kimi K2.5moonshotai/kimi-k2.5 | 70.8 | 262.144K | $0.45 | $2.25 | 2026-01 | |||
| NVIDIA: Nemotron 3 Ultranvidia/nemotron-3-ultra-550b-a55b | 70.7 | 256K | $0.625 | $3.125 | 2026-06-04 |