129 models
Reasoning Tools JSON

Highest-accuracy GPT-5.5 tier for slower, precision-heavy reasoning and coding

openai/gpt-5.5-pro 2026-04-23 1.05M context $30/M input $180/M output
18 providers
Reasoning Tools JSON

Default frontier GPT for coding, computer use, research, and knowledge work

openai/gpt-5.5 2026-04-23 1.05M context $5/M input $30/M output
46 providers
Reasoning Tools Open weights

Stronger MiMo Pro tier for multimodal reasoning and coding-agent execution

xiaomi/mimo-v2.5-pro 2026-04-22 1.04858M context $0.435/M input $0.87/M output
24 providers
Reasoning Tools JSON Open weights

Open MiMo model for multimodal coding agents and long-context automation

xiaomi/mimo-v2.5 2026-04-22 1.04858M context $0.14/M input $0.28/M output
23 providers
Tools Open weights

Qwen vision-language model for visual reasoning, documents, and agent tasks

alibaba/qwen3.6-27b 2026-04-22 262.144K context $0.6/M input $3.6/M output
24 providers
Reasoning Tools JSON Open weights

Multimodal Kimi workhorse for agent loops, coding tasks, and visual context

moonshotai/kimi-k2.6 2026-04-21 262.144K context $0.95/M input $4/M output
64 providers
Reasoning Tools JSON

xAI's default Grok for chat, coding, agentic tools, and lower hallucination risk

xai/grok-4.3 2026-04-17 1M context $1.25/M input $2.5/M output
28 providers
Reasoning Tools JSON

Fast Grok coding model tuned for agentic engineering and iterative edits

xai/grok-build-0.1 2026-04-16 256K context $1/M input $2/M output
19 providers
Reasoning Tools

Stronger Opus tier for advanced software work and high-stakes reasoning

anthropic/claude-opus-4-7 2026-04-16 1M context $5/M input $25/M output
40 providers
Reasoning

Muse Spark is a natively multimodal reasoning model with support for tool-use, visual chain of thought, and multi-agent orchestration.

meta/muse-spark-1.1 2026-04-08 1.04858M context $1.25/M input $4.25/M output
12 providers
Reasoning Tools JSON Open weights

Open Gemma instruction model for efficient chat and self-hosted deployments

google/gemma-4-26b-a4b-it 2026-04-02 262.144K context $0.042/M input $0.22/M output
22 providers
Reasoning Tools JSON Open weights

Largest Gemma 4 instruction model for open, self-hosted chat and reasoning

google/gemma-4-31b-it 2026-04-02 262.144K context $0.09/M input $0.34/M output
33 providers
Reasoning Tools

Earlier Qwen multimodal workhorse for million-token agent and document tasks

alibaba/qwen3.6-plus 2026-04-02 1M context $0.5/M input $3/M output
23 providers
Reasoning Tools Open weights

Reasoning-optimized 398B MoE agent model with extended thinking for long-horizon and multi-turn tool use

arcee-ai/trinity-large-thinking 2026-04-01 524.288K context $0.25/M input $0.9/M output
5 providers
Reasoning Tools

Earlier MiMo Pro model for multimodal agents, reasoning, and code tasks

xiaomi/mimo-v2-pro 2026-03-18 1.04858M context $0.435/M input $0.87/M output
10 providers
Reasoning Tools JSON

Strong small GPT for coding subagents, quick tool use, and high-volume work

openai/gpt-5.4-mini 2026-03-17 400K context $0.75/M input $4.5/M output
31 providers
Reasoning Tools JSON

Cheapest GPT-5.4 lane for simple routing, extraction, and bulk automation

openai/gpt-5.4-nano 2026-03-17 400K context $0.2/M input $1.25/M output
27 providers
Tools Open weights

Efficient Mistral model for fast chat, extraction, and production assistants

mistral/mistral-small-latest 2026-03-16 256K context $0.15/M input $0.6/M output
5 providers
Reasoning Open weights

Nemotron middle tier for collaborative agents and high-volume reasoning workloads

nvidia/nemotron-3-super-120b-a12b 2026-03-11 262.144K context $0.2/M input $0.8/M output
15 providers
Reasoning Tools

More exact GPT-5.4 tier for demanding professional reasoning and agent tasks

openai/gpt-5.4-pro 2026-03-05 1.05M context $30/M input $180/M output
19 providers
Reasoning Tools JSON

Agent-ready GPT for coding and computer-use workflows at a lower cost

openai/gpt-5.4 2026-03-05 1.05M context $2.5/M input $15/M output
42 providers
Reasoning Tools JSON

Low-latency Gemini model for high-volume multimodal and agent workloads

google/gemini-3.1-flash-lite-preview 2026-03-03 1.04858M context $0.25/M input $1.5/M output
13 providers

Qwen vision-language model for visual reasoning, documents, and agent tasks

alibaba/qwen3.5-flash 2026-02-23 1M context $0.029/M input $0.287/M output
7 providers
Reasoning Tools JSON

Reasoning-first Gemini preview for agentic coding and complex problem solving

google/gemini-3.1-pro-preview 2026-02-19 1.04858M context $2/M input $12/M output
31 providers
Reasoning Tools

Claude workhorse for coding agents, careful analysis, and production cost control

anthropic/claude-sonnet-4-6 2026-02-17 1M context $3/M input $15/M output
45 providers

Qwen vision-language model for visual reasoning, documents, and agent tasks

alibaba/qwen3.5-plus 2026-02-16 1M context $0.4/M input $2.4/M output
13 providers
Reasoning Tools Open weights

General GLM flagship for coding, analysis, and tool-heavy engineering workflows

zhipuai/glm-5 2026-02-12 204.8K context $1/M input $3.2/M output
37 providers
Reasoning Tools

High-end Claude for difficult coding, planning, and slower expert reasoning

anthropic/claude-opus-4-6 2026-02-05 1M context $5/M input $25/M output
37 providers
Reasoning Tools JSON

Coding-optimized GPT model for repository edits, reviews, and agentic software work

openai/gpt-5.3-codex 2026-02-05 400K context $1.75/M input $14/M output
28 providers
Tools JSON Open weights

Open-weight Qwen coding model for agents, repository edits, and multi-turn tool use

alibaba/qwen3-coder-next 2026-02-03 262.144K context $0.108/M input $0.675/M output
13 providers
Open weights

StepFun flash lane for quick multimodal reasoning and coding assistance

stepfun/step-3.5-flash 2026-01-29 256K context $0.1/M input $0.3/M output
14 providers
Tools Open weights

Earlier Kimi frontier model for long-context agents, coding, and multimodal work

moonshotai/kimi-k2.5 2026-01 262.144K context $0.3/M input $1.9/M output
46 providers
Reasoning Tools JSON

ByteDance Seed model for multimodal reasoning, long-context analysis, and agent workflows

bytedance-seed/seed-1-8 2025-12-28 256K context $0.119/M input $1.187/M output
3 providers
Reasoning Tools Open weights

MiMo flash model for fast multimodal assistance and agent workflows

xiaomi/mimo-v2-flash 2025-12-16 262.144K context $0.14/M input $0.28/M output
7 providers
Open weights

Small Nemotron 3 MoE for efficient coding, math, and long-context agents

nvidia/nemotron-3-nano-30b-a3b 2025-12-15 262.144K context $0.05/M input $0.2/M output
11 providers
Reasoning Tools

Higher-accuracy GPT-5.2 variant for tougher reasoning and review workflows

openai/gpt-5.2-pro 2025-12-11 400K context $21/M input $168/M output
13 providers
Reasoning Tools

Reliable GPT generation for broad coding, writing, and tool-assisted product work

openai/gpt-5.2 2025-12-11 400K context $1.75/M input $14/M output
33 providers
Reasoning Tools JSON

Code-specialist GPT for repository edits, reviews, and long-running software agents

openai/gpt-5.2-codex 2025-12-11 400K context $0.14/M input $1.14/M output
23 providers
Tools Open weights

Compact multimodal coding model for repository exploration, file editing, and software agents

mistral/devstral-small-2 2025-12-09 262.144K context $0.1/M input $0.3/M output
3 providers
Tools Open weights

Mistral's coding-agent model for repository work, terminal tasks, and software fixes

mistral/devstral-2512 2025-12-09 262.144K context $0.4/M input $2/M output
13 providers
Reasoning

Multimodal reasoning model for visual analysis, planning, and tool use

amazon/nova-2-lite 2025-12-02 1M context $0.3/M input $2.5/M output
4 providers
Open weights

Compact multimodal Mistral model for local assistants, edge agents, and efficient tool use

mistral/ministral-14b 2025-12-02 262.144K context $0.2/M input $0.2/M output
3 providers
Reasoning Tools JSON

Preview Gemini flagship for complex reasoning, coding, and rich multimodal prompts

google/gemini-3-pro-preview 2025-11-18 1.04858M context $0.57/M input $3.43/M output
10 providers
Reasoning

Codex GPT for repository edits, code review, and practical software agents

openai/gpt-5.1-codex 2025-11-13 400K context $1.07/M input $8.5/M output
19 providers
Reasoning Tools JSON

Coding-optimized GPT model for repository edits, reviews, and agentic software work

openai/gpt-5.1-codex-max 2025-11-13 400K context $1.1/M input $9/M output
15 providers
Reasoning

Coding-optimized GPT model for repository edits, reviews, and agentic software work

openai/gpt-5.1-codex-mini 2025-11-13 400K context $0.22/M input $1.8/M output
15 providers
Tools Open weights

Thinking Kimi model for slower research passes, planning, and hard technical questions

moonshotai/kimi-k2-thinking 2025-11-06 262.144K context $0.4/M input $2.5/M output
22 providers
Reasoning Tools JSON

Flagship Claude model for deep reasoning, coding, and long-horizon agents

anthropic/claude-opus-4-5-20251101 2025-11-01 200K context $5/M input $25/M output
17 providers
Tools JSON

Fast Claude model for responsive assistance, classification, and lightweight agents

anthropic/claude-haiku-4-5-20251001 2025-10-15 200K context $1/M input $5/M output
20 providers
Reasoning Tools JSON

ByteDance Seed model for long-context reasoning, instruction following, and tool-assisted tasks

bytedance-seed/seed-1-6 2025-10-15 256K context $0.119/M input $1.187/M output
4 providers