9 models
Ranked by Agentic Index
Reasoning Tools JSON Open weights 50.6

Multimodal Kimi model with 1M context and toggleable max-effort thinking for long-horizon agent work

moonshotai/kimi-k3 2026-07-16 1.04858M context $3/M input $15/M output
63 providers
Reasoning Tools JSON Open weights 42.3

DeepSeek V4 Pro snapshot with million-token context and support for thinking and non-thinking modes

deepseek/deepseek-v4-pro-0813 2026-08-12 1M context $0.442/M input $0.884/M output
30 providers
Reasoning Tools JSON Open weights 41.7

Official DeepSeek V4 Flash release with enhanced agentic capabilities and integrated DSpark speculative decoding

deepseek/deepseek-v4-flash-0731 2026-07-31 1M context $0.035/M input $0.07/M output
43 providers
Reasoning Tools JSON Open weights 27.9

Fast DeepSeek V4 lane for economical reasoning, coding, and long-context work

deepseek/deepseek-v4-flash 2026-04-24 1M context $0.15/M input $0.6/M output
62 providers
Reasoning Tools JSON Open weights 27.7

Open MoE flagship with million-token context for coding and long agent runs

deepseek/deepseek-v4-pro 2026-04-24 1M context $0.435/M input $0.87/M output
62 providers
Tools JSON Open weights 24.3

Multimodal MoE reasoning model (975B total, 41B active) for text, image, and audio

thinkingmachines/inkling 2026-07-15 1.04858M context $1.87/M input $4.68/M output
21 providers
Reasoning Tools JSON Open weights 22.5

Coding-focused Kimi model, stronger on long-horizon repo work with less overthinking

moonshotai/kimi-k2.7-code 2026-06-12 262.144K context $0.95/M input $4/M output
61 providers
Reasoning Tools Open weights 21.7

Largest Nemotron 3 model for maximum open-weight reasoning and agent accuracy

nvidia/nemotron-3-ultra-550b-a55b 2026-06-04 1M context $0.5/M input $2.5/M output
15 providers
Reasoning Tools Open weights 6.1

Fast NVIDIA Nemotron MoE for reliable agentic tasks across enterprise workloads

nvidia/nemotron-3.5-lightning 2026-08-11 262.144K context $0.05/M input $0.2/M output
10 providers