12 models
Ranked by SWE-Bench Verified
Reasoning Tools 96.0

Strongest Claude Opus model for coding, agents, and professional work

anthropic/claude-opus-5 2026-07-24 1M context $5/M input $25/M output
32 providers
Reasoning Tools 95.0

Claude model for creative writing, analysis, and controlled agent workflows

anthropic/claude-fable-5 2026-06-09 1M context $10/M input $50/M output
37 providers
Reasoning Tools 88.6

Top Claude Opus tier for the hardest reasoning, coding, and long-horizon agents

anthropic/claude-opus-4-8 2026-05-28 1M context $5/M input $25/M output
44 providers
Tools JSON 85.2

Everyday Claude agent model for coding, planning, browsing, and general work

anthropic/claude-sonnet-5 2026-06-30 1M context $2/M input $10/M output
35 providers
Reasoning Tools Open weights 80.5

MiniMax multimodal model for long-context coding, perception, and agent planning

minimax/MiniMax-M3 2026-06-01 1.04858M context $0.3/M input $1.2/M output
37 providers
Reasoning Tools 80.4

Qwen frontier model tuned for agent frameworks, coding assistants, and long tasks

alibaba/qwen3.7-max 2026-05-21 1M context $2.5/M input $7.5/M output
31 providers
Reasoning Tools JSON Open weights 80.2

Multimodal Kimi workhorse for agent loops, coding tasks, and visual context

moonshotai/kimi-k2.6 2026-04-21 262.144K context $0.95/M input $4/M output
64 providers
Reasoning Tools Open weights 79.9

Open MiniMax flagship for coding agents, office automation, and complex environments

minimax/MiniMax-M2.7 2026-03-18 204.8K context $0.3/M input $1.2/M output
35 providers
Reasoning Tools Open weights 75.8

Prior MiniMax coding model for agent workflows, office edits, and automation

minimax/MiniMax-M2.5 2026-02-12 204.8K context $0.3/M input $1.2/M output
35 providers
Reasoning Tools Open weights 74.0

Earlier MiniMax agent model for practical coding and productivity tasks

minimax/MiniMax-M2.1 2025-12-23 204.8K context $0.3/M input $1.2/M output
16 providers
Reasoning Tools JSON Open weights 73.4

Open multimodal Qwen MoE for local agents that need vision, audio, and code

alibaba/qwen3.6-35b-a3b 2026-04-17 262.144K context $0.248/M input $1.485/M output
19 providers
Tools Open weights 70.8

Earlier Kimi frontier model for long-context agents, coding, and multimodal work

moonshotai/kimi-k2.5 2026-01 262.144K context $0.3/M input $1.9/M output
46 providers