Kimi K2.6 is Moonshot AI's next-generation multimodal model, designed for long-horizon coding, coding-driven UI/UX generation, and multi-agent orchestration. It handles complex end-to-end coding tasks across Python, Rust, and Go, and...
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
Kimi K3 is a 2.8T parameter open-weight multimodal reasoning model from Moonshot AI. It is suited for complex coding, knowledge work, and long-horizon agentic workflows, and is particularly strong at...
DeepSeek V4 Pro 0813 is a large-scale mixture-of-experts model from DeepSeek. This is the GA release of DeepSeek V4 Pro.
Kimi K2.5 is Moonshot AI's native multimodal model, delivering state-of-the-art visual coding capability and a self-directed agent swarm paradigm. Built on Kimi K2 with continued pretraining over approximately 15T mixed...
DeepSeek V4 Flash 0731 is a sparse mixture-of-experts model from DeepSeek, with 13B active parameters out of 284B total. This re-post-trained revision is suited for coding, reasoning, and agent workflows....
MoonshotAI: Kimi K2.7 Code is a coding-focused model in Moonshot AI's Kimi K2 family, built to complete end-to-end programming tasks reliably over long contexts. It uses a native multimodal mixture-of-experts...
Inkling is an open-weight multimodal mixture-of-experts model from Thinking Machines Lab, with 41B active parameters out of 975B total. It is designed for general-purpose reasoning, coding, agentic and tool-use systems,...
NVIDIA Nemotron 3 Ultra is an open frontier-reasoning and orchestration model from NVIDIA, with 55B active parameters out of 550B total (MoE). Built on a hybrid Transformer-Mamba mixture-of-experts architecture, it...
gpt-oss-120b is an open-weight, 117B-parameter Mixture-of-Experts (MoE) language model from OpenAI designed for high-reasoning, agentic, and general-purpose production use cases. It activates 5.1B parameters per forward pass and is optimized...
DeepSeek R1 is here: Performance on par with [OpenAI o1](/openai/o1), but open-sourced and with fully open reasoning tokens. It's 671B parameters in size, with 37B active in an inference pass....
gpt-oss-20b is an open-weight 21B parameter model released by OpenAI under the Apache 2.0 license. It uses a Mixture-of-Experts (MoE) architecture with 3.6B active parameters per forward pass, optimized for...
| Model | Creator | Score | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|---|
| MoonshotAI: Kimi K2.6moonshotai/kimi-k2.6 | 45.1 | 262.144K | $0.95 | $4 | 2026-04-21 | |||
| MoonshotAI: Kimi K3moonshotai/kimi-k3 | 43.8 | 1.04858M | $2.303 | $11.55 | 2026-07-16 | |||
| DeepSeek: DeepSeek V4 Pro 0813deepseek/deepseek-v4-pro-0813 | 36.3 | 1.04858M | $0.578 | $1.734 | 2026-08-12 | |||
| MoonshotAI: Kimi K2.5moonshotai/kimi-k2.5 | 36.0 | 262.144K | $0.45 | $2.25 | 2026-01 | |||
| DeepSeek: DeepSeek V4 Flash 0731deepseek/deepseek-v4-flash-0731 | 34.5 | 1.04858M | $0.04 | $0.08 | 2026-07-31 | |||
| MoonshotAI: Kimi K2.7 Codemoonshotai/kimi-k2.7-code | 26.3 | 262.144K | $0.71 | $3.5 | 2026-06-12 | |||
| Thinking Machines: Inklingthinkingmachines/inkling | 25.5 | 524.288K | $1 | $4.05 | 2026-07-15 | |||
| NVIDIA: Nemotron 3 Ultranvidia/nemotron-3-ultra-550b-a55b | 23.4 | 256K | $0.625 | $3.125 | 2026-06-04 | |||
| OpenAI: gpt-oss-120bopenai/gpt-oss-120b | 12.3 | 131.072K | $0.037 | $0.17 | 2025-08-05 | |||
| DeepSeek: R1deepseek/deepseek-r1 | 11.4 | 64K | $0.7 | $2.5 | 2025-01-20 | |||
| OpenAI: gpt-oss-20bopenai/gpt-oss-20b | 9.0 | 131.072K | $0.03 | $0.13 | 2025-08-05 |