117 models
Reasoning Tools

Earlier Qwen multimodal workhorse for million-token agent and document tasks

alibaba/qwen3.6-plus 2026-04-02 1M context $0.5/M input $3/M output
23 providers
Reasoning Tools

Fast GLM vision model for screenshots, documents, and multimodal agent tasks

zhipuai/glm-5v-turbo 2026-04-01 200K context $5/M input $22/M output
11 providers
Reasoning Tools Open weights

Low-latency M2.7 variant for interactive coding plans and agent loops

minimax/MiniMax-M2.7-highspeed 2026-03-18 204.8K context $0.6/M input $2.4/M output
13 providers
Reasoning Tools Open weights

Open MiniMax flagship for coding agents, office automation, and complex environments

minimax/MiniMax-M2.7 2026-03-18 204.8K context $0.3/M input $1.2/M output
35 providers
Reasoning Tools JSON

Cheapest GPT-5.4 lane for simple routing, extraction, and bulk automation

openai/gpt-5.4-nano 2026-03-17 400K context $0.2/M input $1.25/M output
27 providers
Reasoning Tools JSON

Strong small GPT for coding subagents, quick tool use, and high-volume work

openai/gpt-5.4-mini 2026-03-17 400K context $0.75/M input $4.5/M output
31 providers
Reasoning Tools JSON

Faster GLM-5 lane for coding agents that need lower latency

zhipuai/glm-5-turbo 2026-03-16 200K context $0.9/M input $3.7/M output
12 providers
Reasoning Tools JSON

Reasoning Grok for document-heavy analysis and long-horizon tool use

xai/grok-4.20-0309-reasoning 2026-03-09 1M context $1.25/M input $2.5/M output
9 providers
Reasoning Tools

More exact GPT-5.4 tier for demanding professional reasoning and agent tasks

openai/gpt-5.4-pro 2026-03-05 1.05M context $30/M input $180/M output
19 providers
Reasoning Tools JSON

Agent-ready GPT for coding and computer-use workflows at a lower cost

openai/gpt-5.4 2026-03-05 1.05M context $2.5/M input $15/M output
42 providers
Tools Open weights

Qwen vision-language model for visual reasoning, documents, and agent tasks

alibaba/qwen3.5-27b 2026-02-23 262.144K context $0.3/M input $2.4/M output
14 providers

Qwen vision-language model for visual reasoning, documents, and agent tasks

alibaba/qwen3.5-flash 2026-02-23 1M context $0.029/M input $0.287/M output
7 providers
Reasoning Tools JSON Open weights

Qwen vision-language model for visual reasoning, documents, and agent tasks

alibaba/qwen3.5-122b-a10b 2026-02-23 262.144K context $0.4/M input $3.2/M output
16 providers
Reasoning Tools JSON Open weights

Qwen vision-language model for visual reasoning, documents, and agent tasks

alibaba/qwen3.5-35b-a3b 2026-02-23 262.144K context $0.25/M input $2/M output
11 providers
Reasoning Tools JSON

Reasoning-first Gemini preview for agentic coding and complex problem solving

google/gemini-3.1-pro-preview 2026-02-19 1.04858M context $2/M input $12/M output
31 providers
Reasoning Tools

Claude workhorse for coding agents, careful analysis, and production cost control

anthropic/claude-sonnet-4-6 2026-02-17 1M context $3/M input $15/M output
45 providers

Qwen vision-language model for visual reasoning, documents, and agent tasks

alibaba/qwen3.5-plus 2026-02-16 1M context $0.4/M input $2.4/M output
13 providers
Reasoning Tools JSON Open weights

Large open Qwen multimodal MoE for visual agents and long technical tasks

alibaba/qwen3.5-397b-a17b 2026-02-15 262.144K context $0.6/M input $3.6/M output
18 providers
JSON

Cost-efficient ByteDance Seed 2.0 model for production chat, analysis, and structured generation

bytedance-seed/seed-2.0-lite 2026-02-14 256K context $0.089/M input $0.534/M output
8 providers
Reasoning Tools JSON

Flagship ByteDance Seed 2.0 model for complex multimodal reasoning and long-horizon agent workflows

bytedance-seed/seed-2.0-pro 2026-02-14 256K context $0.475/M input $2.375/M output
6 providers
Reasoning Tools JSON

Lightweight ByteDance Seed 2.0 model for low-latency multimodal reasoning and high-volume tasks

bytedance-seed/seed-2.0-mini 2026-02-14 256K context $0.03/M input $0.297/M output
9 providers
Reasoning Tools JSON

ByteDance Seed coding model for multimodal software engineering and long-running agents

bytedance-seed/seed-2.0-code 2026-02-14 262.144K context $0.4/M input $2.4/M output
10 providers
Reasoning Tools Open weights

General GLM flagship for coding, analysis, and tool-heavy engineering workflows

zhipuai/glm-5 2026-02-12 204.8K context $1/M input $3.2/M output
37 providers
Reasoning Tools Open weights

Prior MiniMax coding model for agent workflows, office edits, and automation

minimax/MiniMax-M2.5 2026-02-12 204.8K context $0.3/M input $1.2/M output
35 providers
Reasoning Tools JSON

Coding-optimized GPT model for repository edits, reviews, and agentic software work

openai/gpt-5.3-codex 2026-02-05 400K context $1.75/M input $14/M output
28 providers
Reasoning Tools

High-end Claude for difficult coding, planning, and slower expert reasoning

anthropic/claude-opus-4-6 2026-02-05 1M context $5/M input $25/M output
37 providers
Tools JSON Open weights

Open-weight Qwen coding model for agents, repository edits, and multi-turn tool use

alibaba/qwen3-coder-next 2026-02-03 262.144K context $0.108/M input $0.675/M output
13 providers

MiniMax M2 variant tuned for conversational and character-driven agent interactions

minimax/MiniMax-M2-Her 2026-01-23 65.536K context $0.3/M input $1.2/M output
3 providers
Reasoning Tools Open weights

Efficient GLM model for fast reasoning, coding, and agent workflows

zhipuai/glm-4.7-flashx 2026-01-19 200K context $0.07/M input $0.4/M output
5 providers
Tools Open weights

Earlier Kimi frontier model for long-context agents, coding, and multimodal work

moonshotai/kimi-k2.5 2026-01 262.144K context $0.3/M input $1.9/M output
46 providers
Reasoning Tools JSON

ByteDance Seed model for multimodal reasoning, long-context analysis, and agent workflows

bytedance-seed/seed-1-8 2025-12-28 256K context $0.119/M input $1.187/M output
3 providers
Reasoning Tools Open weights

Earlier MiniMax agent model for practical coding and productivity tasks

minimax/MiniMax-M2.1 2025-12-23 204.8K context $0.3/M input $1.2/M output
16 providers
Reasoning Tools Open weights

Mature GLM model for dependable coding, reasoning, and structured agent tasks

zhipuai/glm-4.7 2025-12-22 204.8K context $0.6/M input $2.2/M output
30 providers
Reasoning Tools JSON

New Gemini flash lane bringing frontier-style multimodal reasoning to cheaper runs

google/gemini-3-flash-preview 2025-12-17 1.04858M context $0.5/M input $3/M output
25 providers
Reasoning Tools

Reliable GPT generation for broad coding, writing, and tool-assisted product work

openai/gpt-5.2 2025-12-11 400K context $1.75/M input $14/M output
33 providers
Reasoning Tools JSON

Code-specialist GPT for repository edits, reviews, and long-running software agents

openai/gpt-5.2-codex 2025-12-11 400K context $0.14/M input $1.14/M output
23 providers
Reasoning Tools JSON Open weights

Hybrid-reasoning DeepSeek model with thinking and non-thinking modes, sparse attention, and tool-use

deepseek/deepseek-v3.2 2025-12-01 128K context $0.18/M input $0.35/M output
24 providers
Reasoning Tools

Flagship Claude model for deep reasoning, coding, and long-horizon agents

anthropic/claude-opus-4-5 2025-11-24 200K context $5/M input $25/M output
17 providers
Tools JSON

xAI's fast agentic tool-calling model with a 2M context window; non-reasoning variant for low-latency responses

xai/grok-4.1-fast 2025-11-19 2M context $0.2/M input $0.5/M output
2 providers
Reasoning

Coding-optimized GPT model for repository edits, reviews, and agentic software work

openai/gpt-5.1-codex-mini 2025-11-13 400K context $0.22/M input $1.8/M output
15 providers
Reasoning Tools JSON

Sharper GPT-5 generation for coding, product work, and tool-assisted tasks

openai/gpt-5.1 2025-11-13 400K context $1.25/M input $10/M output
28 providers
Reasoning Tools JSON

Coding-optimized GPT model for repository edits, reviews, and agentic software work

openai/gpt-5.1-codex-max 2025-11-13 400K context $1.1/M input $9/M output
15 providers
Reasoning Tools Open weights

Efficient open MiniMax model built for coding agents and tool-heavy workflows

minimax/MiniMax-M2 2025-10-27 204.8K context $0.3/M input $1.2/M output
14 providers
Reasoning Tools JSON

ByteDance Seed model for long-context reasoning, instruction following, and tool-assisted tasks

bytedance-seed/seed-1-6 2025-10-15 256K context $0.119/M input $1.187/M output
4 providers
Reasoning Tools

Fast Claude lane for lightweight agents, office tasks, and responsive chat

anthropic/claude-haiku-4-5 2025-10-15 200K context $1/M input $5/M output
23 providers
Reasoning Tools Open weights

Late GLM-4 workhorse for coding agents, reasoning, and structured tasks

zhipuai/glm-4.6 2025-09-30 204.8K context $0.6/M input $2.2/M output
18 providers
Reasoning Tools

Balanced Claude model for coding, analysis, agent workflows, and cost control

anthropic/claude-sonnet-4-5 2025-09-29 200K context $3/M input $15/M output
20 providers

Flagship Qwen3 model for coding agents, complex reasoning, and tool use

alibaba/qwen3-max 2025-09-23 262.144K context $1.2/M input $6/M output
20 providers
Reasoning Tools JSON

Low-latency ByteDance Seed model for high-throughput chat, extraction, and lightweight tool use

bytedance-seed/seed-1-6-flash 2025-08-28 256K context $0.022/M input $0.223/M output
3 providers
Reasoning Tools JSON

ByteDance Seed multimodal model for image understanding, visual reasoning, and tool-assisted tasks

bytedance-seed/seed-1-6-vision 2025-08-15 256K context $0.119/M input $1.187/M output
2 providers