7 models
Ranked by SWE-Bench Verified
Reasoning Tools 96.0

Strongest Claude Opus model for coding, agents, and professional work

anthropic/claude-opus-5 2026-07-24 1M context $5/M input $25/M output
32 providers
Reasoning Tools 95.0

Claude model for creative writing, analysis, and controlled agent workflows

anthropic/claude-fable-5 2026-06-09 1M context $10/M input $50/M output
37 providers
Reasoning Tools 88.6

Top Claude Opus tier for the hardest reasoning, coding, and long-horizon agents

anthropic/claude-opus-4-8 2026-05-28 1M context $5/M input $25/M output
44 providers
Tools JSON 85.2

Everyday Claude agent model for coding, planning, browsing, and general work

anthropic/claude-sonnet-5 2026-06-30 1M context $2/M input $10/M output
35 providers
Reasoning Tools Open weights 73.8

Mature GLM model for dependable coding, reasoning, and structured agent tasks

zhipuai/glm-4.7 2025-12-22 204.8K context $0.6/M input $2.2/M output
30 providers
Reasoning Tools Open weights 72.8

General GLM flagship for coding, analysis, and tool-heavy engineering workflows

zhipuai/glm-5 2026-02-12 204.8K context $1/M input $3.2/M output
37 providers
Tools Open weights 71.3

Thinking Kimi model for slower research passes, planning, and hard technical questions

moonshotai/kimi-k2-thinking 2025-11-06 262.144K context $0.4/M input $2.5/M output
22 providers