2 models
Ranked by SWE-Bench Multilingual
Reasoning Tools Open weights 67.7

Largest Nemotron 3 model for maximum open-weight reasoning and agent accuracy

nvidia/nemotron-3-ultra-550b-a55b 2026-06-04 1M context $0.5/M input $2.5/M output
15 providers
Reasoning Tools Open weights 63.1

Agentic coding model from Poolside in the XS size class for local deployment

poolside/laguna-xs-2.1 2026-07-02 262.144K context $0.06/M input $0.12/M output
6 providers