Tools Open weights

Efficient Mistral model for fast chat, extraction, and production assistants

mistral/mistral-small-latest 2026-03-16 256K context $0.15/M input $0.6/M output
5 providers
Reasoning Tools JSON

Faster GLM-5 lane for coding agents that need lower latency

zhipuai/glm-5-turbo 2026-03-16 200K context $0.9/M input $3.7/M output
12 providers
Reasoning Open weights

Nemotron middle tier for collaborative agents and high-volume reasoning workloads

nvidia/nemotron-3-super-120b-a12b 2026-03-11 262.144K context $0.2/M input $0.8/M output
15 providers
Reasoning Tools JSON

Reasoning Grok for document-heavy analysis and long-horizon tool use

xai/grok-4.20-0309-reasoning 2026-03-09 1M context $1.25/M input $2.5/M output
9 providers
Tools JSON

Grok model for agentic tool use, reasoning, coding, and live assistance

xai/grok-4.20-0309-non-reasoning 2026-03-09 1M context $1.25/M input $2.5/M output
8 providers
Reasoning Tools

More exact GPT-5.4 tier for demanding professional reasoning and agent tasks

openai/gpt-5.4-pro 2026-03-05 1.05M context $30/M input $180/M output
19 providers
Reasoning Tools JSON

Agent-ready GPT for coding and computer-use workflows at a lower cost

openai/gpt-5.4 2026-03-05 1.05M context $2.5/M input $15/M output
42 providers
Reasoning Tools JSON

Low-latency Gemini model for high-volume multimodal and agent workloads

google/gemini-3.1-flash-lite-preview 2026-03-03 1.04858M context $0.25/M input $1.5/M output
13 providers
Reasoning Tools JSON Open weights

Qwen vision-language model for visual reasoning, documents, and agent tasks

alibaba/qwen3.5-35b-a3b 2026-02-23 262.144K context $0.25/M input $2/M output
11 providers

Qwen vision-language model for visual reasoning, documents, and agent tasks

alibaba/qwen3.5-flash 2026-02-23 1M context $0.029/M input $0.287/M output
7 providers
Reasoning Tools Open weights

Qwen instruction model for multilingual chat, reasoning, and tool use

alibaba/qwen3.5-9b 2026-02-23 262.144K context $0.04/M input $0.15/M output
16 providers
Tools Open weights

Qwen vision-language model for visual reasoning, documents, and agent tasks

alibaba/qwen3.5-27b 2026-02-23 262.144K context $0.3/M input $2.4/M output
14 providers
Reasoning Tools JSON Open weights

Qwen vision-language model for visual reasoning, documents, and agent tasks

alibaba/qwen3.5-122b-a10b 2026-02-23 262.144K context $0.4/M input $3.2/M output
16 providers
Reasoning Tools JSON

Advanced Gemini model for complex reasoning, coding, and multimodal analysis

google/gemini-3.1-pro-preview-customtools 2026-02-19 1.04858M context $2/M input $12/M output
11 providers
Reasoning Tools JSON

Reasoning-first Gemini preview for agentic coding and complex problem solving

google/gemini-3.1-pro-preview 2026-02-19 1.04858M context $2/M input $12/M output
31 providers
Reasoning Tools

Claude workhorse for coding agents, careful analysis, and production cost control

anthropic/claude-sonnet-4-6 2026-02-17 1M context $3/M input $15/M output
45 providers

Qwen vision-language model for visual reasoning, documents, and agent tasks

alibaba/qwen3.5-plus 2026-02-16 1M context $0.4/M input $2.4/M output
13 providers
Reasoning Tools JSON Open weights

Large open Qwen multimodal MoE for visual agents and long technical tasks

alibaba/qwen3.5-397b-a17b 2026-02-15 262.144K context $0.6/M input $3.6/M output
18 providers
Reasoning Tools JSON

Flagship ByteDance Seed 2.0 model for complex multimodal reasoning and long-horizon agent workflows

bytedance-seed/seed-2.0-pro 2026-02-14 256K context $0.475/M input $2.375/M output
6 providers
Reasoning Tools JSON

Lightweight ByteDance Seed 2.0 model for low-latency multimodal reasoning and high-volume tasks

bytedance-seed/seed-2.0-mini 2026-02-14 256K context $0.03/M input $0.297/M output
9 providers
JSON

Cost-efficient ByteDance Seed 2.0 model for production chat, analysis, and structured generation

bytedance-seed/seed-2.0-lite 2026-02-14 256K context $0.089/M input $0.534/M output
8 providers
Reasoning Tools JSON

ByteDance Seed coding model for multimodal software engineering and long-running agents

bytedance-seed/seed-2.0-code 2026-02-14 262.144K context $0.4/M input $2.4/M output
10 providers
Reasoning Tools Open weights

Prior MiniMax coding model for agent workflows, office edits, and automation

minimax/MiniMax-M2.5 2026-02-12 204.8K context $0.3/M input $1.2/M output
35 providers
Reasoning Tools Open weights

General GLM flagship for coding, analysis, and tool-heavy engineering workflows

zhipuai/glm-5 2026-02-12 204.8K context $1/M input $3.2/M output
37 providers
Reasoning Tools

High-end Claude for difficult coding, planning, and slower expert reasoning

anthropic/claude-opus-4-6 2026-02-05 1M context $5/M input $25/M output
37 providers
Reasoning Tools JSON

Coding-optimized GPT model for repository edits, reviews, and agentic software work

openai/gpt-5.3-codex 2026-02-05 400K context $1.75/M input $14/M output
28 providers
Tools JSON Open weights

Open-weight Qwen coding model for agents, repository edits, and multi-turn tool use

alibaba/qwen3-coder-next 2026-02-03 262.144K context $0.108/M input $0.675/M output
13 providers
Open weights

StepFun flash lane for quick multimodal reasoning and coding assistance

stepfun/step-3.5-flash 2026-01-29 256K context $0.1/M input $0.3/M output
14 providers
Reasoning Tools Open weights

Budget GLM lane for fast coding help, routing, and everyday automation

zhipuai/glm-4.7-flash 2026-01-19 200K context $0.06/M input $0.4/M output
14 providers
Tools Open weights

Earlier Kimi frontier model for long-context agents, coding, and multimodal work

moonshotai/kimi-k2.5 2026-01 262.144K context $0.3/M input $1.9/M output
46 providers
Reasoning Tools JSON

ByteDance Seed model for multimodal reasoning, long-context analysis, and agent workflows

bytedance-seed/seed-1-8 2025-12-28 256K context $0.119/M input $1.187/M output
3 providers
Reasoning Tools Open weights

Earlier MiniMax agent model for practical coding and productivity tasks

minimax/MiniMax-M2.1 2025-12-23 204.8K context $0.3/M input $1.2/M output
16 providers
Reasoning Tools Open weights

Mature GLM model for dependable coding, reasoning, and structured agent tasks

zhipuai/glm-4.7 2025-12-22 204.8K context $0.6/M input $2.2/M output
30 providers
Reasoning Tools JSON

New Gemini flash lane bringing frontier-style multimodal reasoning to cheaper runs

google/gemini-3-flash-preview 2025-12-17 1.04858M context $0.5/M input $3/M output
25 providers
Open weights

Small Nemotron 3 MoE for efficient coding, math, and long-context agents

nvidia/nemotron-3-nano-30b-a3b 2025-12-15 262.144K context $0.05/M input $0.2/M output
11 providers
Reasoning Tools JSON

Code-specialist GPT for repository edits, reviews, and long-running software agents

openai/gpt-5.2-codex 2025-12-11 400K context $0.14/M input $1.14/M output
23 providers
Reasoning Tools

Higher-accuracy GPT-5.2 variant for tougher reasoning and review workflows

openai/gpt-5.2-pro 2025-12-11 400K context $21/M input $168/M output
13 providers
Reasoning Tools

Reliable GPT generation for broad coding, writing, and tool-assisted product work

openai/gpt-5.2 2025-12-11 400K context $1.75/M input $14/M output
33 providers
Tools Open weights

Mistral's coding-agent model for repository work, terminal tasks, and software fixes

mistral/devstral-2512 2025-12-09 262.144K context $0.4/M input $2/M output
13 providers
Reasoning

Multimodal reasoning model for visual analysis, planning, and tool use

amazon/nova-2-lite 2025-12-02 1M context $0.3/M input $2.5/M output
4 providers
Tools Open weights

Mistral coding agent model for repository tasks and software engineering workflows

mistral/devstral-medium-latest 2025-12-02 262.144K context $0.4/M input $2/M output
3 providers
Tools Open weights

DeepSeek chat model for instruction following, coding, and analysis

deepseek/deepseek-chat 2025-12-01 1M context $0.147/M input $0.295/M output
8 providers
Open weights

DeepSeek reasoning model for multi-step analysis, math, coding, and tools

deepseek/deepseek-reasoner 2025-12-01 1M context $0.147/M input $0.295/M output
6 providers
Reasoning Tools

Flagship Claude model for deep reasoning, coding, and long-horizon agents

anthropic/claude-opus-4-5 2025-11-24 200K context $5/M input $25/M output
17 providers
Tools JSON

xAI's fast agentic tool-calling model with a 2M context window; non-reasoning variant for low-latency responses

xai/grok-4.1-fast 2025-11-19 2M context $0.2/M input $0.5/M output
2 providers
Reasoning Tools JSON

xAI's fast agentic tool-calling model with a 2M context window and built-in reasoning

xai/grok-4.1-fast-reasoning 2025-11-19 2M context $0.2/M input $0.5/M output
4 providers
Reasoning Tools JSON

Preview Gemini flagship for complex reasoning, coding, and rich multimodal prompts

google/gemini-3-pro-preview 2025-11-18 1.04858M context $0.57/M input $3.43/M output
10 providers
Reasoning Tools JSON

Coding-optimized GPT model for repository edits, reviews, and agentic software work

openai/gpt-5.1-codex-max 2025-11-13 400K context $1.1/M input $9/M output
15 providers
Reasoning

Codex GPT for repository edits, code review, and practical software agents

openai/gpt-5.1-codex 2025-11-13 400K context $1.07/M input $8.5/M output
19 providers
Reasoning

Coding-optimized GPT model for repository edits, reviews, and agentic software work

openai/gpt-5.1-codex-mini 2025-11-13 400K context $0.22/M input $1.8/M output
15 providers