Large coding-reasoning model for agentic software tasks and RL search
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
Open MoE flagship with million-token context for coding and long agent runs
MiniMax multimodal model for long-context coding, perception, and agent planning
Multimodal Kimi workhorse for agent loops, coding tasks, and visual context
Open MiniMax flagship for coding agents, office automation, and complex environments
Fast DeepSeek V4 lane for economical reasoning, coding, and long-context work
Stronger MiMo Pro tier for multimodal reasoning and coding-agent execution
Tencent Hy reasoning model for coding, instruction following, and agent tasks
Balanced Mistral model for enterprise assistants, multilingual work, and tools
Balanced Mistral model for enterprise assistants, multilingual work, and tools
Qwen vision-language model for visual reasoning, documents, and agent tasks
Newer StepFun flash model for faster agents, coding, and multimodal prompts
Large open Qwen multimodal MoE for visual agents and long technical tasks
Muse Glimmer is a 30-billion-parameter open-weight multimodal model from Meta Superintelligence Labs, distilled from Muse Spark for always-on local agents, tool use, coding, and image understanding.
Prior MiniMax coding model for agent workflows, office edits, and automation
Large coding-reasoning model for agentic software tasks and RL search
StepFun flash lane for quick multimodal reasoning and coding assistance
Tencent Hy reasoning model for coding, instruction following, and agent tasks
Earlier MiniMax agent model for practical coding and productivity tasks
Mature GLM model for dependable coding, reasoning, and structured agent tasks
Open multimodal Qwen MoE for local agents that need vision, audio, and code
General GLM flagship for coding, analysis, and tool-heavy engineering workflows
Qwen vision-language model for visual reasoning, documents, and agent tasks
Qwen vision-language model for visual reasoning, documents, and agent tasks
Thinking Kimi model for slower research passes, planning, and hard technical questions
Agentic coding model from Poolside in the XS size class for local deployment
Earlier Kimi frontier model for long-context agents, coding, and multimodal work
Largest Nemotron 3 model for maximum open-weight reasoning and agent accuracy
Efficient open MiniMax model built for coding agents and tool-heavy workflows
Open coding-reasoning model for repository tasks and self-improving agents
Cohere coding model for practical software engineering and agentic edits
Budget GLM lane for fast coding help, routing, and everyday automation
Mistral coding agent model for repository tasks and software engineering workflows
| Model | Creator | Score | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|---|
| Ornith 1.0 397Bdeepreinforce/ornith-1.0-397b | 82.4 | 262.144K | — | — | 2026-06-25 | |||
| DeepSeek V4 Prodeepseek/deepseek-v4-pro | 80.6 | 1M | $0.435 | $0.87 | 2026-04-24 | |||
| MiniMax-M3minimax/MiniMax-M3 | 80.5 | 1.04858M | $0.3 | $1.2 | 2026-06-01 | |||
| Kimi K2.6moonshotai/kimi-k2.6 | 80.2 | 262.144K | $0.95 | $4 | 2026-04-21 | |||
| MiniMax-M2.7minimax/MiniMax-M2.7 | 79.9 | 204.8K | $0.3 | $1.2 | 2026-03-18 | |||
| DeepSeek V4 Flashdeepseek/deepseek-v4-flash | 79.0 | 1M | $0.15 | $0.6 | 2026-04-24 | |||
| MiMo-V2.5-Proxiaomi/mimo-v2.5-pro | 78.9 | 1.04858M | $0.435 | $0.87 | 2026-04-22 | |||
| Hy3tencent/hy3 | 78.0 | 256K | $0.066 | $0.26 | 2026-07-06 | |||
| Mistral Medium 3.5mistral/mistral-medium-2604 | 77.6 | 262.144K | $1.5 | $7.5 | 2026-04-29 | |||
| Mistral Medium (latest)mistral/mistral-medium-latest | 77.6 | 262.144K | $1.5 | $7.5 | 2026-04-29 | |||
| Qwen3.6 27Balibaba/qwen3.6-27b | 77.2 | 262.144K | $0.6 | $3.6 | 2026-04-22 | |||
| Step 3.7 Flashstepfun/step-3.7-flash | 76.5 | 256K | $0.185 | $1.11 | 2026-05-29 | |||
| Qwen3.5 397B-A17Balibaba/qwen3.5-397b-a17b | 76.4 | 262.144K | $0.6 | $3.6 | 2026-02-15 | |||
| Muse Glimmer 30Bmeta/muse-glimmer-30b | 76.0 | 131.072K | $0.2 | $0.8 | 2026-08-10 | |||
| MiniMax-M2.5minimax/MiniMax-M2.5 | 75.8 | 204.8K | $0.3 | $1.2 | 2026-02-12 | |||
| Ornith 1.0 35Bdeepreinforce/ornith-1.0-35b | 75.6 | 262.144K | — | — | 2026-06-25 | |||
| Step 3.5 Flashstepfun/step-3.5-flash | 74.4 | 256K | $0.1 | $0.3 | 2026-01-29 | |||
| Hy3 previewtencent/hy3-preview | 74.4 | 256K | $0.066 | $0.26 | 2026-04-20 | |||
| MiniMax-M2.1minimax/MiniMax-M2.1 | 74.0 | 204.8K | $0.3 | $1.2 | 2025-12-23 | |||
| GLM-4.7zhipuai/glm-4.7 | 73.8 | 204.8K | $0.6 | $2.2 | 2025-12-22 | |||
| Qwen3.6 35B-A3Balibaba/qwen3.6-35b-a3b | 73.4 | 262.144K | $0.248 | $1.485 | 2026-04-17 | |||
| GLM-5zhipuai/glm-5 | 72.8 | 204.8K | $1 | $3.2 | 2026-02-12 | |||
| Qwen3.5 27Balibaba/qwen3.5-27b | 72.4 | 262.144K | $0.3 | $2.4 | 2026-02-23 | |||
| Qwen3.5 122B-A10Balibaba/qwen3.5-122b-a10b | 72.0 | 262.144K | $0.4 | $3.2 | 2026-02-23 | |||
| Kimi K2 Thinkingmoonshotai/kimi-k2-thinking | 71.3 | 262.144K | $0.4 | $2.5 | 2025-11-06 | |||
| Laguna XS 2.1poolside/laguna-xs-2.1 | 70.9 | 262.144K | $0.06 | $0.12 | 2026-07-02 | |||
| Kimi K2.5moonshotai/kimi-k2.5 | 70.8 | 262.144K | $0.3 | $1.9 | 2026-01 | |||
| Nemotron 3 Ultra 550B A55Bnvidia/nemotron-3-ultra-550b-a55b | 70.7 | 1M | $0.5 | $2.5 | 2026-06-04 | |||
| MiniMax-M2minimax/MiniMax-M2 | 69.4 | 204.8K | $0.3 | $1.2 | 2025-10-27 | |||
| Ornith 1.0 9Bdeepreinforce/ornith-1.0-9b | 69.4 | 262.144K | — | — | 2026-06-25 | |||
| North Mini Codecohere/north-mini-code-1-0 | 67.6 | 256K | — | — | 2026-06-09 | |||
| GLM-4.7-Flashzhipuai/glm-4.7-flash | 59.2 | 200K | $0.06 | $0.4 | 2026-01-19 | |||
| Devstral Smallmistral/devstral-small-2507 | 53.6 | 128K | $0.1 | $0.3 | 2025-07-10 |