Reasoning Tools JSON Open weights

DeepSeek V4.1 Flash model for reasoning and agentic coding

deepseek/deepseek-v4.1-flash 2026-09-10 1M context $0.15/M input $0.6/M output
23 providers
Reasoning Tools JSON Open weights

Native multimodal GLM model for efficient coding and long-horizon agent tasks

zhipuai/glm-5.3-flash 2026-08-26 1M context $0.075/M input $0.25/M output
36 providers
Reasoning Tools JSON Open weights

Dense 27B vision-language model for coding, agent tasks, and image and video understanding

alibaba/qwen3.8-27b 2026-08-14 262.144K context $0.1/M input $0.4/M output
27 providers
Reasoning Tools JSON Open weights

Multimodal Kimi model with 1M context and toggleable max-effort thinking for long-horizon agent work

moonshotai/kimi-k3 2026-07-16 1.04858M context $3/M input $15/M output
63 providers
Tools JSON Open weights

Multimodal MoE reasoning model (975B total, 41B active) for text, image, and audio

thinkingmachines/inkling 2026-07-15 1.04858M context $1.87/M input $4.68/M output
21 providers
Reasoning Tools JSON Open weights

Coding-focused Kimi model, stronger on long-horizon repo work with less overthinking

moonshotai/kimi-k2.7-code 2026-06-12 262.144K context $0.95/M input $4/M output
61 providers
Reasoning Tools Open weights

MiniMax multimodal model for long-context coding, perception, and agent planning

minimax/MiniMax-M3 2026-06-01 1.04858M context $0.3/M input $1.2/M output
37 providers
Reasoning Tools JSON Open weights

Multimodal Kimi workhorse for agent loops, coding tasks, and visual context

moonshotai/kimi-k2.6 2026-04-21 262.144K context $0.95/M input $4/M output
64 providers
Reasoning Tools JSON Open weights

Open Gemma instruction model for efficient chat and self-hosted deployments

google/gemma-4-26b-a4b-it 2026-04-02 262.144K context $0.042/M input $0.22/M output
22 providers
Tools Open weights

Earlier Kimi frontier model for long-context agents, coding, and multimodal work

moonshotai/kimi-k2.5 2026-01 262.144K context $0.3/M input $1.9/M output
46 providers
Tools Open weights

Open multimodal Llama for strong reasoning with efficient everyday serving

meta/llama-4-maverick-17b-instruct 2025-04-05 1M context $0.14/M input $0.59/M output
6 providers