Strongest Claude Opus model for coding, agents, and professional work
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
Large coding-reasoning model for agentic software tasks and RL search
Everyday Claude agent model for coding, planning, browsing, and general work
Qwen frontier model tuned for agent frameworks, coding assistants, and long tasks
xAI's Grok model for chat, coding, agentic tools, and lower hallucination risk
Meituan LongCat-2.0, a reasoning model with tool calling and a 1M-token context window
Large coding-reasoning model for agentic software tasks and RL search
Largest Nemotron 3 model for maximum open-weight reasoning and agent accuracy
Agentic coding model from Poolside in the XS size class for local deployment
Open coding-reasoning model for repository tasks and self-improving agents
| Model | Creator | Score | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|---|
| Claude Opus 5anthropic/claude-opus-5 | 89.5 | 1M | $5 | $25 | 2026-07-24 | |||
| Ornith 1.0 397Bdeepreinforce/ornith-1.0-397b | 78.9 | 262.144K | — | — | 2026-06-25 | |||
| Claude Sonnet 5anthropic/claude-sonnet-5 | 78.3 | 1M | $2 | $10 | 2026-06-30 | |||
| Qwen3.7 Maxalibaba/qwen3.7-max | 78.3 | 1M | $2.5 | $7.5 | 2026-05-21 | |||
| Grok 4.5xai/grok-4.5 | 78.0 | 500K | $2 | $6 | 2026-07-08 | |||
| LongCat-2.0meituan/longcat-2.0 | 77.3 | 1M | $0.3 | $1.2 | 2026-06-30 | |||
| Ornith 1.0 35Bdeepreinforce/ornith-1.0-35b | 69.3 | 262.144K | — | — | 2026-06-25 | |||
| Nemotron 3 Ultra 550B A55Bnvidia/nemotron-3-ultra-550b-a55b | 67.7 | 1M | $0.5 | $2.5 | 2026-06-04 | |||
| Laguna XS 2.1poolside/laguna-xs-2.1 | 63.1 | 262.144K | $0.06 | $0.12 | 2026-07-02 | |||
| Ornith 1.0 9Bdeepreinforce/ornith-1.0-9b | 52.0 | 262.144K | — | — | 2026-06-25 |