Open MoE flagship with million-token context for coding and long agent runs
62 providers
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
Open MoE flagship with million-token context for coding and long agent runs
MiniMax multimodal model for long-context coding, perception, and agent planning
Fast DeepSeek V4 lane for economical reasoning, coding, and long-context work
Stronger MiMo Pro tier for multimodal reasoning and coding-agent execution
Largest Nemotron 3 model for maximum open-weight reasoning and agent accuracy
| Model | Creator | Score | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|---|
| DeepSeek V4 Prodeepseek/deepseek-v4-pro | 80.6 | 1M | $0.435 | $0.87 | 2026-04-24 | |||
| MiniMax-M3minimax/MiniMax-M3 | 80.5 | 1.04858M | $0.3 | $1.2 | 2026-06-01 | |||
| DeepSeek V4 Flashdeepseek/deepseek-v4-flash | 79.0 | 1M | $0.15 | $0.6 | 2026-04-24 | |||
| MiMo-V2.5-Proxiaomi/mimo-v2.5-pro | 78.9 | 1.04858M | $0.435 | $0.87 | 2026-04-22 | |||
| Nemotron 3 Ultra 550B A55Bnvidia/nemotron-3-ultra-550b-a55b | 70.7 | 1M | $0.5 | $2.5 | 2026-06-04 |