Reasoning Tools JSON Open weights

DeepSeek V4.1 Flash model for reasoning and agentic coding

deepseek/deepseek-v4.1-flash 2026-09-10 1M context $0.15/M input $0.6/M output
23 providers
Reasoning Tools JSON

GPT-6 Astra is OpenAI's most capable model for complex reasoning, coding, computer use, research, and document creation.

openai/gpt-6-astra 2026-09-04 1.05M context $10/M input $50/M output
18 providers
Reasoning Tools JSON

Muse Spark 1.3 is a multimodal reasoning model from Meta for long-running agentic, multi-agent, and coding workflows. It improves long-horizon agent collaboration, instruction following, and coding efficiency relative to Muse Spark 1.2.

meta/muse-spark-1.3 2026-09-02 1.04858M context $1.25/M input $4.25/M output
9 providers
Reasoning Tools JSON

Google's most intelligent Flash model, engineered for long-horizon software engineering, autonomous agents, and complex enterprise workflows

google/gemini-3.8-flash 2026-09-02 1.04858M context $0.75/M input $3.75/M output
17 providers
Reasoning Tools JSON Open weights

Dense 27B vision-language model for coding, agent tasks, and image and video understanding

alibaba/qwen3.8-27b 2026-08-14 262.144K context $0.1/M input $0.4/M output
27 providers
Reasoning Tools JSON

High-efficiency Gemini model for agentic workflows, coding, and multimodal reasoning

google/gemini-3.7-flash 2026-08-13 1.04858M context $0.75/M input $3.75/M output
23 providers
Reasoning Tools JSON Open weights

DeepSeek V4 Pro snapshot with million-token context and support for thinking and non-thinking modes

deepseek/deepseek-v4-pro-0813 2026-08-12 1M context $0.442/M input $0.884/M output
30 providers
Reasoning Tools Open weights

Fast NVIDIA Nemotron MoE for reliable agentic tasks across enterprise workloads

nvidia/nemotron-3.5-lightning 2026-08-11 262.144K context $0.05/M input $0.2/M output
10 providers
Reasoning Tools JSON Open weights

Muse Glimmer is a 30-billion-parameter open-weight multimodal model from Meta Superintelligence Labs, distilled from Muse Spark for always-on local agents, tool use, coding, and image understanding.

meta/muse-glimmer-30b 2026-08-10 131.072K context $0.2/M input $0.8/M output
9 providers
Tools JSON

Upstage's flagship model, specialized for agentic use

upstage/solar-pro4 2026-08-06 524.288K context $0.3/M input $1.2/M output
5 providers
Reasoning Tools JSON

Muse Spark 1.2 is a coding-focused update to Muse Spark 1.1 with improvements in code generation, complex debugging, codebase understanding, and end-to-end developer workflows.

meta/muse-spark-1.2 2026-08-05 1.04858M context $1.25/M input $4.25/M output
14 providers
Reasoning Tools JSON

Japanese-specialized reasoning model based on Kimi K2.6 and tuned for Japanese language, culture, and business workflows

sakana/sakana-namazu 2026-08-03 262.144K context $0.95/M input $4/M output
4 providers
Reasoning Tools JSON Open weights

Official DeepSeek V4 Flash release with enhanced agentic capabilities and integrated DSpark speculative decoding

deepseek/deepseek-v4-flash-0731 2026-07-31 1M context $0.035/M input $0.07/M output
43 providers
Tools Open weights

Multimodal MoE reasoning model (276B total, 12B active) for text, image, and audio

thinkingmachines/inkling-small 2026-07-30 1.04858M context $0.45/M input $1.2/M output
11 providers
Reasoning Tools

Strongest Claude Opus model for coding, agents, and professional work

anthropic/claude-opus-5 2026-07-24 1M context $5/M input $25/M output
32 providers
Reasoning Tools JSON

Fast Gemini model balancing multimodal reasoning, tool use, and cost

google/gemini-3.5-flash-lite 2026-07-21 1.04858M context $0.3/M input $2.5/M output
23 providers
Reasoning Tools JSON

Fast Gemini model balancing multimodal reasoning, tool use, and cost

google/gemini-3.6-flash 2026-07-21 1.04858M context $0.75/M input $3.75/M output
24 providers
Reasoning Tools JSON Open weights

Multimodal Kimi model with 1M context and toggleable max-effort thinking for long-horizon agent work

moonshotai/kimi-k3 2026-07-16 1.04858M context $3/M input $15/M output
63 providers
Tools JSON Open weights

Multimodal MoE reasoning model (975B total, 41B active) for text, image, and audio

thinkingmachines/inkling 2026-07-15 1.04858M context $1.87/M input $4.68/M output
21 providers
Reasoning Tools JSON

Frontier GPT-5.6 model for complex professional work, coding, and agentic workflows

openai/gpt-5.6-sol 2026-07-09 1.05M context $4/M input $20/M output
36 providers
Reasoning Tools JSON

Balanced GPT-5.6 model for capable, cost-efficient everyday work

openai/gpt-5.6-terra 2026-07-09 1.05M context $2/M input $12/M output
35 providers
Reasoning Tools JSON

Cost-efficient GPT-5.6 model for fast, high-volume workloads

openai/gpt-5.6-luna 2026-07-09 1.05M context $0.2/M input $1.2/M output
36 providers
Reasoning Tools JSON

xAI's Grok model for chat, coding, agentic tools, and lower hallucination risk

xai/grok-4.5 2026-07-08 500K context $2/M input $6/M output
26 providers
Reasoning Tools JSON Open weights

Tencent Hy reasoning model for coding, instruction following, and agent tasks

tencent/hy3 2026-07-06 256K context $0.066/M input $0.26/M output
17 providers
Reasoning Tools Open weights

Agentic coding model from Poolside in the XS size class for local deployment

poolside/laguna-xs-2.1 2026-07-02 262.144K context $0.06/M input $0.12/M output
6 providers
Tools JSON

Everyday Claude agent model for coding, planning, browsing, and general work

anthropic/claude-sonnet-5 2026-06-30 1M context $2/M input $10/M output
35 providers
Reasoning Tools

Meituan LongCat-2.0, a reasoning model with tool calling and a 1M-token context window

meituan/longcat-2.0 2026-06-30 1M context $0.3/M input $1.2/M output
5 providers
Reasoning Tools JSON

Quality-first multi-agent model for hard research, analysis, and competitions

sakana/fugu-ultra 2026-06-15 1M context $5/M input $30/M output
11 providers
Reasoning Tools JSON Open weights

Open flagship GLM for long-horizon coding agents and million-token context work

zhipuai/glm-5.2 2026-06-13 1M context $1.4/M input $4.4/M output
72 providers
Reasoning Tools JSON Open weights

Coding-focused Kimi model, stronger on long-horizon repo work with less overthinking

moonshotai/kimi-k2.7-code 2026-06-12 262.144K context $0.95/M input $4/M output
61 providers
Reasoning Tools

Claude model for creative writing, analysis, and controlled agent workflows

anthropic/claude-fable-5 2026-06-09 1M context $10/M input $50/M output
37 providers
Reasoning Tools Open weights

Largest Nemotron 3 model for maximum open-weight reasoning and agent accuracy

nvidia/nemotron-3-ultra-550b-a55b 2026-06-04 1M context $0.5/M input $2.5/M output
15 providers
Reasoning Tools Open weights

MiniMax multimodal model for long-context coding, perception, and agent planning

minimax/MiniMax-M3 2026-06-01 1.04858M context $0.3/M input $1.2/M output
37 providers
Reasoning Tools JSON Open weights

Newer StepFun flash model for faster agents, coding, and multimodal prompts

stepfun/step-3.7-flash 2026-05-29 256K context $0.185/M input $1.11/M output
17 providers
Reasoning Tools

Top Claude Opus tier for the hardest reasoning, coding, and long-horizon agents

anthropic/claude-opus-4-8 2026-05-28 1M context $5/M input $25/M output
44 providers
Reasoning Tools

Qwen frontier model tuned for agent frameworks, coding assistants, and long tasks

alibaba/qwen3.7-max 2026-05-21 1M context $2.5/M input $7.5/M output
31 providers
Reasoning Tools JSON

Fast Gemini model balancing multimodal reasoning, tool use, and cost

google/gemini-3.5-flash 2026-05-19 1.04858M context $1.5/M input $9/M output
31 providers
Reasoning Tools JSON Open weights

Fast DeepSeek V4 lane for economical reasoning, coding, and long-context work

deepseek/deepseek-v4-flash 2026-04-24 1M context $0.15/M input $0.6/M output
62 providers
Reasoning Tools JSON Open weights

Open MoE flagship with million-token context for coding and long agent runs

deepseek/deepseek-v4-pro 2026-04-24 1M context $0.435/M input $0.87/M output
62 providers
Reasoning Tools JSON

Default frontier GPT for coding, computer use, research, and knowledge work

openai/gpt-5.5 2026-04-23 1.05M context $5/M input $30/M output
46 providers
Reasoning Tools JSON

Highest-accuracy GPT-5.5 tier for slower, precision-heavy reasoning and coding

openai/gpt-5.5-pro 2026-04-23 1.05M context $30/M input $180/M output
18 providers
Reasoning Tools JSON Open weights

Open MiMo model for multimodal coding agents and long-context automation

xiaomi/mimo-v2.5 2026-04-22 1.04858M context $0.14/M input $0.28/M output
23 providers
Tools Open weights

Qwen vision-language model for visual reasoning, documents, and agent tasks

alibaba/qwen3.6-27b 2026-04-22 262.144K context $0.6/M input $3.6/M output
24 providers
Reasoning Tools Open weights

Stronger MiMo Pro tier for multimodal reasoning and coding-agent execution

xiaomi/mimo-v2.5-pro 2026-04-22 1.04858M context $0.435/M input $0.87/M output
24 providers
Reasoning Tools JSON Open weights

Multimodal Kimi workhorse for agent loops, coding tasks, and visual context

moonshotai/kimi-k2.6 2026-04-21 262.144K context $0.95/M input $4/M output
64 providers
Reasoning Tools Open weights

Tencent Hy reasoning model for coding, instruction following, and agent tasks

tencent/hy3-preview 2026-04-20 256K context $0.066/M input $0.26/M output
7 providers
Reasoning Tools JSON

xAI's default Grok for chat, coding, agentic tools, and lower hallucination risk

xai/grok-4.3 2026-04-17 1M context $1.25/M input $2.5/M output
28 providers
Reasoning Tools JSON Open weights

Open multimodal Qwen MoE for local agents that need vision, audio, and code

alibaba/qwen3.6-35b-a3b 2026-04-17 262.144K context $0.248/M input $1.485/M output
19 providers
Reasoning Tools

Stronger Opus tier for advanced software work and high-stakes reasoning

anthropic/claude-opus-4-7 2026-04-16 1M context $5/M input $25/M output
40 providers
Reasoning

Muse Spark is a natively multimodal reasoning model with support for tool-use, visual chain of thought, and multi-agent orchestration.

meta/muse-spark-1.1 2026-04-08 1.04858M context $1.25/M input $4.25/M output
12 providers