Open MoE flagship with million-token context for coding and long agent runs
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
MiniMax multimodal model for long-context coding, perception, and agent planning
Multimodal Kimi workhorse for agent loops, coding tasks, and visual context
Open MiniMax flagship for coding agents, office automation, and complex environments
Fast DeepSeek V4 lane for economical reasoning, coding, and long-context work
Stronger MiMo Pro tier for multimodal reasoning and coding-agent execution
Tencent Hy reasoning model for coding, instruction following, and agent tasks
Balanced Mistral model for enterprise assistants, multilingual work, and tools
Balanced Mistral model for enterprise assistants, multilingual work, and tools
Qwen vision-language model for visual reasoning, documents, and agent tasks
Newer StepFun flash model for faster agents, coding, and multimodal prompts
Large open Qwen multimodal MoE for visual agents and long technical tasks
Muse Glimmer is a 30-billion-parameter open-weight multimodal model from Meta Superintelligence Labs, distilled from Muse Spark for always-on local agents, tool use, coding, and image understanding.
Prior MiniMax coding model for agent workflows, office edits, and automation
Tencent Hy reasoning model for coding, instruction following, and agent tasks
StepFun flash lane for quick multimodal reasoning and coding assistance
Earlier MiniMax agent model for practical coding and productivity tasks
Mature GLM model for dependable coding, reasoning, and structured agent tasks
Open multimodal Qwen MoE for local agents that need vision, audio, and code
General GLM flagship for coding, analysis, and tool-heavy engineering workflows
Qwen vision-language model for visual reasoning, documents, and agent tasks
Qwen vision-language model for visual reasoning, documents, and agent tasks
Thinking Kimi model for slower research passes, planning, and hard technical questions
Earlier Kimi frontier model for long-context agents, coding, and multimodal work
Largest Nemotron 3 model for maximum open-weight reasoning and agent accuracy
Efficient open MiniMax model built for coding agents and tool-heavy workflows
Cohere coding model for practical software engineering and agentic edits
Budget GLM lane for fast coding help, routing, and everyday automation
| Model | Creator | Score | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|---|
| DeepSeek V4 Prodeepseek/deepseek-v4-pro | 80.6 | 1M | $0.435 | $0.87 | 2026-04-24 | |||
| MiniMax-M3minimax/MiniMax-M3 | 80.5 | 1.04858M | $0.3 | $1.2 | 2026-06-01 | |||
| Kimi K2.6moonshotai/kimi-k2.6 | 80.2 | 262.144K | $0.95 | $4 | 2026-04-21 | |||
| MiniMax-M2.7minimax/MiniMax-M2.7 | 79.9 | 204.8K | $0.3 | $1.2 | 2026-03-18 | |||
| DeepSeek V4 Flashdeepseek/deepseek-v4-flash | 79.0 | 1M | $0.15 | $0.6 | 2026-04-24 | |||
| MiMo-V2.5-Proxiaomi/mimo-v2.5-pro | 78.9 | 1.04858M | $0.435 | $0.87 | 2026-04-22 | |||
| Hy3tencent/hy3 | 78.0 | 256K | $0.066 | $0.26 | 2026-07-06 | |||
| Mistral Medium (latest)mistral/mistral-medium-latest | 77.6 | 262.144K | $1.5 | $7.5 | 2026-04-29 | |||
| Mistral Medium 3.5mistral/mistral-medium-2604 | 77.6 | 262.144K | $1.5 | $7.5 | 2026-04-29 | |||
| Qwen3.6 27Balibaba/qwen3.6-27b | 77.2 | 262.144K | $0.6 | $3.6 | 2026-04-22 | |||
| Step 3.7 Flashstepfun/step-3.7-flash | 76.5 | 256K | $0.185 | $1.11 | 2026-05-29 | |||
| Qwen3.5 397B-A17Balibaba/qwen3.5-397b-a17b | 76.4 | 262.144K | $0.6 | $3.6 | 2026-02-15 | |||
| Muse Glimmer 30Bmeta/muse-glimmer-30b | 76.0 | 131.072K | $0.2 | $0.8 | 2026-08-10 | |||
| MiniMax-M2.5minimax/MiniMax-M2.5 | 75.8 | 204.8K | $0.3 | $1.2 | 2026-02-12 | |||
| Hy3 previewtencent/hy3-preview | 74.4 | 256K | $0.066 | $0.26 | 2026-04-20 | |||
| Step 3.5 Flashstepfun/step-3.5-flash | 74.4 | 256K | $0.1 | $0.3 | 2026-01-29 | |||
| MiniMax-M2.1minimax/MiniMax-M2.1 | 74.0 | 204.8K | $0.3 | $1.2 | 2025-12-23 | |||
| GLM-4.7zhipuai/glm-4.7 | 73.8 | 204.8K | $0.6 | $2.2 | 2025-12-22 | |||
| Qwen3.6 35B-A3Balibaba/qwen3.6-35b-a3b | 73.4 | 262.144K | $0.248 | $1.485 | 2026-04-17 | |||
| GLM-5zhipuai/glm-5 | 72.8 | 204.8K | $1 | $3.2 | 2026-02-12 | |||
| Qwen3.5 27Balibaba/qwen3.5-27b | 72.4 | 262.144K | $0.3 | $2.4 | 2026-02-23 | |||
| Qwen3.5 122B-A10Balibaba/qwen3.5-122b-a10b | 72.0 | 262.144K | $0.4 | $3.2 | 2026-02-23 | |||
| Kimi K2 Thinkingmoonshotai/kimi-k2-thinking | 71.3 | 262.144K | $0.4 | $2.5 | 2025-11-06 | |||
| Kimi K2.5moonshotai/kimi-k2.5 | 70.8 | 262.144K | $0.3 | $1.9 | 2026-01 | |||
| Nemotron 3 Ultra 550B A55Bnvidia/nemotron-3-ultra-550b-a55b | 70.7 | 1M | $0.5 | $2.5 | 2026-06-04 | |||
| MiniMax-M2minimax/MiniMax-M2 | 69.4 | 204.8K | $0.3 | $1.2 | 2025-10-27 | |||
| North Mini Codecohere/north-mini-code-1-0 | 67.6 | 256K | — | — | 2026-06-09 | |||
| GLM-4.7-Flashzhipuai/glm-4.7-flash | 59.2 | 200K | $0.06 | $0.4 | 2026-01-19 |