4 models
Ranked by SWE-Bench Verified
Reasoning Tools Open weights 78.9

Stronger MiMo Pro tier for multimodal reasoning and coding-agent execution

xiaomi/mimo-v2.5-pro 2026-04-22 1.04858M context $0.435/M input $0.87/M output
24 providers
Reasoning Tools JSON Open weights 76.4

Large open Qwen multimodal MoE for visual agents and long technical tasks

alibaba/qwen3.5-397b-a17b 2026-02-15 262.144K context $0.6/M input $3.6/M output
18 providers
Reasoning Tools JSON Open weights 73.4

Open multimodal Qwen MoE for local agents that need vision, audio, and code

alibaba/qwen3.6-35b-a3b 2026-04-17 262.144K context $0.248/M input $1.485/M output
19 providers
Reasoning Tools Open weights 70.9

Agentic coding model from Poolside in the XS size class for local deployment

poolside/laguna-xs-2.1 2026-07-02 262.144K context $0.06/M input $0.12/M output
6 providers