Kimi K3 is a 2.8T parameter open-weight multimodal reasoning model from Moonshot AI. It is suited for complex coding, knowledge work, and long-horizon agentic workflows, and is particularly strong at...
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
Kimi K2.6 is Moonshot AI's next-generation multimodal model, designed for long-horizon coding, coding-driven UI/UX generation, and multi-agent orchestration. It handles complex end-to-end coding tasks across Python, Rust, and Go, and...
DeepSeek V4 Pro is a large-scale Mixture-of-Experts model from DeepSeek with 1.6T total parameters and 49B activated parameters, supporting a 1M-token context window. It is designed for advanced reasoning, coding,...
MiMo-V2.5-Pro is Xiaomi’s flagship model, delivering strong performance in general agentic capabilities, complex software engineering, and long-horizon tasks, with top rankings on benchmarks such as ClawEval, GDPVal, and SWE-bench Pro....
MoonshotAI: Kimi K2.7 Code is a coding-focused model in Moonshot AI's Kimi K2 family, built to complete end-to-end programming tasks reliably over long contexts. It uses a native multimodal mixture-of-experts...
MiMo-V2.5 is a native omnimodal model by Xiaomi. It delivers Pro-level agentic performance at roughly half the inference cost, while surpassing MiMo-V2-Omni in multimodal perception across image and video understanding...
DeepSeek V4 Flash 0731 is a sparse mixture-of-experts model from DeepSeek, with 13B active parameters out of 284B total. This re-post-trained revision is suited for coding, reasoning, and agent workflows....
Kimi K2.5 is Moonshot AI's native multimodal model, delivering state-of-the-art visual coding capability and a self-directed agent swarm paradigm. Built on Kimi K2 with continued pretraining over approximately 15T mixed...
DeepSeek V4 Flash is an efficiency-optimized Mixture-of-Experts model from DeepSeek with 284B total parameters and 13B activated parameters, supporting a 1M-token context window. It is designed for fast inference and...
Hy3 is a 295B-parameter Mixture-of-Experts model from Tencent (21B active, 192 experts with top-8 routing) built for reasoning, agentic workflows, and real-world production use. It supports a configurable reasoning effort:...
Inkling is an open-weight multimodal mixture-of-experts model from Thinking Machines Lab, with 41B active parameters out of 975B total. It is designed for general-purpose reasoning, coding, agentic and tool-use systems,...
NVIDIA Nemotron 3 Ultra is an open frontier-reasoning and orchestration model from NVIDIA, with 55B active parameters out of 550B total (MoE). Built on a hybrid Transformer-Mamba mixture-of-experts architecture, it...
DeepSeek-V3.2 is a large language model designed to harmonize high computational efficiency with strong reasoning and agentic tool-use performance. It introduces DeepSeek Sparse Attention (DSA), a fine-grained sparse attention mechanism...
Step 3.7 Flash is StepFun's latest high-efficiency multimodal Mixture-of-Experts model. It pairs a 196B-parameter language backbone with a vision encoder for native image and video understanding, activating roughly 11B parameters...
DeepSeek-V3 is the latest model from the DeepSeek team, building upon the instruction following and coding abilities of the previous versions. Pre-trained on nearly 15 trillion tokens, the reported evaluations...
Trinity Large Thinking is a powerful open source reasoning model from the team at Arcee AI. It shows strong performance in PinchBench, agentic workloads, and reasoning tasks. Launch video: https://youtu.be/Gc82AXLa0Rg?si=4RLn6WBz33qT--B7...
gpt-oss-120b is an open-weight, 117B-parameter Mixture-of-Experts (MoE) language model from OpenAI designed for high-reasoning, agentic, and general-purpose production use cases. It activates 5.1B parameters per forward pass and is optimized...
| Model | Creator | Score | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|---|
| MoonshotAI: Kimi K3moonshotai/kimi-k3 | 1426.0 | 1.04858M | $2.34 | $11.7 | 2026-07-16 | |||
| MoonshotAI: Kimi K2.6moonshotai/kimi-k2.6 | 1297.0 | 262.144K | $0.95 | $4 | 2026-04-21 | |||
| DeepSeek: DeepSeek V4 Pro 0423deepseek/deepseek-v4-pro | 1283.0 | 1.024M | $0.86 | $1.72 | 2026-04-24 | |||
| Xiaomi: MiMo-V2.5-Proxiaomi/mimo-v2.5-pro | 1276.0 | 1.04858M | $0.435 | $0.87 | 2026-04-22 | |||
| MoonshotAI: Kimi K2.7 Codemoonshotai/kimi-k2.7-code | 1269.0 | 262.144K | $0.71 | $3.5 | 2026-06-12 | |||
| Xiaomi: MiMo-V2.5xiaomi/mimo-v2.5 | 1241.0 | 1.04858M | $0.14 | $0.28 | 2026-04-22 | |||
| DeepSeek: DeepSeek V4 Flash 0731deepseek/deepseek-v4-flash-0731 | 1238.0 | 1.04858M | $0.065 | $0.18 | 2026-07-31 | |||
| MoonshotAI: Kimi K2.5moonshotai/kimi-k2.5 | 1234.0 | 262.144K | $0.45 | $2.25 | 2026-01 | |||
| DeepSeek: DeepSeek V4 Flash 0423deepseek/deepseek-v4-flash | 1217.0 | 1.024M | $0.086 | $0.172 | 2026-04-24 | |||
| Tencent: Hy3tencent/hy3 | 1205.0 | 262.144K | $0.083 | $0.33 | 2026-07-06 | |||
| Thinking Machines: Inklingthinkingmachines/inkling | 1188.0 | 1.04858M | $1 | $4.05 | 2026-07-15 | |||
| NVIDIA: Nemotron 3 Ultranvidia/nemotron-3-ultra-550b-a55b | 1163.0 | 256K | $0.625 | $3.125 | 2026-06-04 | |||
| DeepSeek: DeepSeek V3.2deepseek/deepseek-v3.2 | 1161.0 | 163.84K | $0.269 | $0.4 | 2025-12-01 | |||
| StepFun: Step 3.7 Flashstepfun/step-3.7-flash | 1156.0 | 256K | $0.2 | $1.15 | 2026-05-29 | |||
| DeepSeek: DeepSeek V3deepseek/deepseek-chat | 1115.0 | 128K | $0.257 | $1.029 | 2025-12-01 | |||
| Arcee AI: Trinity Large Thinkingarcee-ai/trinity-large-thinking | 1110.0 | 262.144K | $0.25 | $0.8 | 2026-04-01 | |||
| OpenAI: gpt-oss-120bopenai/gpt-oss-120b | 930.0 | 131.072K | $0.037 | $0.17 | 2025-08-05 |