3 models
Ranked by SWE-Bench Pro
Reasoning Tools 67.7

Preview Qwen flagship for million-token multimodal reasoning and long-horizon agentic workflows

alibaba/qwen3.8-max-preview 2026-07-19 1M context $2/M input $6/M output
6 providers
Reasoning Tools 60.6

Qwen frontier model tuned for agent frameworks, coding assistants, and long tasks

alibaba/qwen3.7-max 2026-05-21 1M context $2.5/M input $7.5/M output
31 providers
Tools Open weights 38.7

Open Qwen coding heavyweight for repository reasoning and agentic engineering

alibaba/qwen3-coder-480b-a35b-instruct 2025-04 262.144K context $1.5/M input $7.5/M output
7 providers