Reasoning Tools JSON Open weights

DeepSeek V4.1 Flash model for reasoning and agentic coding

deepseek/deepseek-v4.1-flash 2026-09-10 1M context $0.15/M input $0.6/M output
17 providers
Reasoning Tools JSON

Muse Spark 1.3 is a multimodal reasoning model from Meta for long-running agentic, multi-agent, and coding workflows. It improves long-horizon agent collaboration, instruction following, and coding efficiency relative to Muse Spark 1.2.

meta/muse-spark-1.3 2026-09-02 1.04858M context $1.25/M input $4.25/M output
9 providers
Reasoning Tools JSON

Claude model for demanding reasoning and long-horizon agentic work

anthropic/claude-fable-5-1 2026-09-01 1M context $10/M input $50/M output
24 providers
Reasoning Tools JSON Open weights

Native multimodal GLM model for efficient coding and long-horizon agent tasks

zhipuai/glm-5.3-flash 2026-08-26 1M context $0.075/M input $0.25/M output
35 providers
Reasoning Tools JSON

Qwen vision-language model for visual reasoning, documents, and agent tasks

alibaba/qwen3.8-flash 2026-08-26 1M context $0.15/M input $0.47/M output
21 providers
Reasoning Tools JSON

Experimental multimodal DeepSeek V4 Flash model for image understanding, coding, and agentic work

deepseek/deepseek-v4-flash-vision-exp 2026-08-21 1M context $0.15/M input $0.6/M output
19 providers
Reasoning Tools JSON Open weights

Dense 27B vision-language model for coding, agent tasks, and image and video understanding

alibaba/qwen3.8-27b 2026-08-14 262.144K context $0.1/M input $0.4/M output
27 providers
Reasoning Tools JSON Open weights

Flagship GLM model for long-horizon coding, agents, and complex project delivery

zhipuai/glm-5.3 2026-08-14 1M context $1.4/M input $4.4/M output
37 providers
Reasoning Tools JSON

High-efficiency Gemini model for agentic workflows, coding, and multimodal reasoning

google/gemini-3.7-flash 2026-08-13 1.04858M context $0.75/M input $3.75/M output
23 providers
Reasoning Tools JSON

xAI's frontier model for long-running agents, coding, knowledge work, and visual projects

xai/grok-4.6 2026-08-12 500K context $2/M input $6/M output
27 providers
Reasoning Tools JSON

Muse Spark 1.2 is a coding-focused update to Muse Spark 1.1 with improvements in code generation, complex debugging, codebase understanding, and end-to-end developer workflows.

meta/muse-spark-1.2 2026-08-05 1.04858M context $1.25/M input $4.25/M output
14 providers
Tools JSON

2.4-trillion-parameter MoE flagship for coding, professional work, multimodal understanding, and long-horizon agentic workflows

alibaba/qwen3.8-max 2026-08-03 1M context $2/M input $6/M output
28 providers
Reasoning Tools

Strongest Claude Opus model for coding, agents, and professional work

anthropic/claude-opus-5 2026-07-24 1M context $5/M input $25/M output
32 providers
Reasoning Tools JSON

Fast Gemini model balancing multimodal reasoning, tool use, and cost

google/gemini-3.6-flash 2026-07-21 1.04858M context $0.75/M input $3.75/M output
24 providers
Reasoning Tools JSON

Fast Gemini model balancing multimodal reasoning, tool use, and cost

google/gemini-3.5-flash-lite 2026-07-21 1.04858M context $0.3/M input $2.5/M output
23 providers
Reasoning Tools

Preview Qwen flagship for million-token multimodal reasoning and long-horizon agentic workflows

alibaba/qwen3.8-max-preview 2026-07-19 1M context $2/M input $6/M output
6 providers
Reasoning Tools JSON Open weights

Multimodal Kimi model with 1M context and toggleable max-effort thinking for long-horizon agent work

moonshotai/kimi-k3 2026-07-16 1.04858M context $3/M input $15/M output
62 providers
Reasoning Tools JSON

Cost-efficient GPT-5.6 model for fast, high-volume workloads

openai/gpt-5.6-luna 2026-07-09 1.05M context $0.2/M input $1.2/M output
36 providers
Reasoning Tools JSON

Balanced GPT-5.6 model for capable, cost-efficient everyday work

openai/gpt-5.6-terra 2026-07-09 1.05M context $2/M input $12/M output
35 providers
Reasoning Tools JSON

Frontier GPT-5.6 model for complex professional work, coding, and agentic workflows

openai/gpt-5.6-sol 2026-07-09 1.05M context $4/M input $20/M output
36 providers
Reasoning Tools JSON

xAI's Grok model for chat, coding, agentic tools, and lower hallucination risk

xai/grok-4.5 2026-07-08 500K context $2/M input $6/M output
26 providers
Reasoning Tools JSON Open weights

Tencent Hy reasoning model for coding, instruction following, and agent tasks

tencent/hy3 2026-07-06 256K context $0.066/M input $0.26/M output
17 providers
Tools JSON

Everyday Claude agent model for coding, planning, browsing, and general work

anthropic/claude-sonnet-5 2026-06-30 1M context $2/M input $10/M output
35 providers
Reasoning Tools JSON

Faster ByteDance Seed 2.1 model for multimodal reasoning and latency-sensitive agent workflows

bytedance-seed/seed-2.1-turbo 2026-06-23 256K context $0.354/M input $1.77/M output
7 providers
Reasoning Tools JSON Open weights

Open flagship GLM for long-horizon coding agents and million-token context work

zhipuai/glm-5.2 2026-06-13 1M context $1.4/M input $4.4/M output
72 providers
Reasoning Tools JSON Open weights

Coding-focused Kimi model, stronger on long-horizon repo work with less overthinking

moonshotai/kimi-k2.7-code 2026-06-12 262.144K context $0.95/M input $4/M output
61 providers
Reasoning Tools Open weights

Lower-latency Kimi Code variant for interactive edits and coding-agent loops

moonshotai/kimi-k2.7-code-highspeed 2026-06-12 262.144K context $1.9/M input $8/M output
9 providers
Reasoning Tools

Claude model for creative writing, analysis, and controlled agent workflows

anthropic/claude-fable-5 2026-06-09 1M context $10/M input $50/M output
37 providers
Reasoning Tools Open weights

Largest Nemotron 3 model for maximum open-weight reasoning and agent accuracy

nvidia/nemotron-3-ultra-550b-a55b 2026-06-04 1M context $0.5/M input $2.5/M output
15 providers
Reasoning Tools

Multimodal Qwen workhorse for long-context agents, visual inputs, and coding

alibaba/qwen3.7-plus 2026-06-02 1M context $0.5/M input $3/M output
29 providers
Reasoning Tools Open weights

MiniMax multimodal model for long-context coding, perception, and agent planning

minimax/MiniMax-M3 2026-06-01 1.04858M context $0.3/M input $1.2/M output
37 providers
Reasoning Tools

Top Claude Opus tier for the hardest reasoning, coding, and long-horizon agents

anthropic/claude-opus-4-8 2026-05-28 1M context $5/M input $25/M output
44 providers
Reasoning Tools

Qwen frontier model tuned for agent frameworks, coding assistants, and long tasks

alibaba/qwen3.7-max 2026-05-21 1M context $2.5/M input $7.5/M output
31 providers
Reasoning Tools JSON

Fast Gemini model balancing multimodal reasoning, tool use, and cost

google/gemini-3.5-flash 2026-05-19 1.04858M context $1.5/M input $9/M output
31 providers
Reasoning Tools JSON

Low-latency Gemini model for high-volume multimodal and agent workloads

google/gemini-3.1-flash-lite 2026-05-07 1.04858M context $0.25/M input $1.5/M output
23 providers

Qwen vision-language model for visual reasoning, documents, and agent tasks

alibaba/qwen3.6-flash 2026-04-27 1M context $0.188/M input $1.125/M output
22 providers
Reasoning Tools JSON Open weights

Fast DeepSeek V4 lane for economical reasoning, coding, and long-context work

deepseek/deepseek-v4-flash 2026-04-24 1M context $0.15/M input $0.6/M output
62 providers
Reasoning Tools JSON Open weights

Open MoE flagship with million-token context for coding and long agent runs

deepseek/deepseek-v4-pro 2026-04-24 1M context $0.435/M input $0.87/M output
62 providers
Reasoning Tools JSON

Highest-accuracy GPT-5.5 tier for slower, precision-heavy reasoning and coding

openai/gpt-5.5-pro 2026-04-23 1.05M context $30/M input $180/M output
18 providers
Reasoning Tools JSON

Default frontier GPT for coding, computer use, research, and knowledge work

openai/gpt-5.5 2026-04-23 1.05M context $5/M input $30/M output
46 providers
Reasoning Tools JSON Open weights

Open MiMo model for multimodal coding agents and long-context automation

xiaomi/mimo-v2.5 2026-04-22 1.04858M context $0.14/M input $0.28/M output
23 providers
Reasoning Tools Open weights

Stronger MiMo Pro tier for multimodal reasoning and coding-agent execution

xiaomi/mimo-v2.5-pro 2026-04-22 1.04858M context $0.435/M input $0.87/M output
24 providers
Reasoning Tools JSON Open weights

Multimodal Kimi workhorse for agent loops, coding tasks, and visual context

moonshotai/kimi-k2.6 2026-04-21 262.144K context $0.95/M input $4/M output
64 providers

Flagship Qwen model for complex reasoning, coding, and agentic workflows

alibaba/qwen3.6-max-preview 2026-04-20 262.144K context $1.3/M input $7.8/M output
12 providers
Reasoning Tools JSON Open weights

Open multimodal Qwen MoE for local agents that need vision, audio, and code

alibaba/qwen3.6-35b-a3b 2026-04-17 262.144K context $0.248/M input $1.485/M output
19 providers
Reasoning Tools JSON

xAI's default Grok for chat, coding, agentic tools, and lower hallucination risk

xai/grok-4.3 2026-04-17 1M context $1.25/M input $2.5/M output
28 providers
Reasoning Tools

Stronger Opus tier for advanced software work and high-stakes reasoning

anthropic/claude-opus-4-7 2026-04-16 1M context $5/M input $25/M output
40 providers
Reasoning Tools JSON

Fast Grok coding model tuned for agentic engineering and iterative edits

xai/grok-build-0.1 2026-04-16 256K context $1/M input $2/M output
19 providers
Reasoning

Muse Spark is a natively multimodal reasoning model with support for tool-use, visual chain of thought, and multi-agent orchestration.

meta/muse-spark-1.1 2026-04-08 1.04858M context $1.25/M input $4.25/M output
12 providers
Reasoning Tools JSON Open weights

Strong GLM coding model for agentic engineering, terminals, and repository generation

zhipuai/glm-5.1 2026-04-07 200K context $1.4/M input $4.4/M output
45 providers