Reasoning Tools

Balanced Claude model for coding, analysis, agent workflows, and cost control

anthropic/claude-sonnet-4-5 2025-09-29 200K context $3/M input $15/M output
20 providers

Flagship Qwen3 model for coding agents, complex reasoning, and tool use

alibaba/qwen3-max 2025-09-23 262.144K context $1.2/M input $6/M output
20 providers
Reasoning Tools JSON

Coding-optimized GPT model for repository edits, reviews, and agentic software work

openai/gpt-5-codex 2025-09-15 400K context $1.1/M input $9/M output
13 providers

Nano Banana image model for fast generation, edits, and character-consistent assets

google/gemini-2.5-flash-image 2025-08-26 32.768K context $0.3/M input $30/M output
8 providers
Reasoning

Small GPT-5 for responsive agents, coding help, and everyday automation

openai/gpt-5-mini 2025-08-07 400K context $0.25/M input $2/M output
29 providers
Reasoning

Tiny GPT-5 lane for routing, extraction, classification, and bulk jobs

openai/gpt-5-nano 2025-08-07 400K context $0.05/M input $0.4/M output
26 providers
Tools

Original GPT-5 workhorse for reasoning, coding, writing, and tool workflows

openai/gpt-5 2025-08-07 400K context $1.25/M input $10/M output
30 providers
Tools

Mistral coding agent model for repository tasks and software engineering workflows

mistral/devstral-medium-2507 2025-07-10 128K context $0.4/M input $2/M output
2 providers
Reasoning Tools JSON

Fast Gemini workhorse for multimodal apps where latency and price matter

google/gemini-2.5-flash 2025-06-17 1.04858M context $0.3/M input $2.5/M output
30 providers
Reasoning Tools JSON

Google's proven reasoning model for coding, math, and multimodal analysis

google/gemini-2.5-pro 2025-06-17 1.04858M context $1.25/M input $10/M output
29 providers
Reasoning Tools JSON

Lean Gemini 2.5 lane for cheap multimodal traffic and quick agents

google/gemini-2.5-flash-lite 2025-06-17 1.04858M context $0.1/M input $0.4/M output
18 providers
Reasoning Tools

High-effort o3 tier for difficult technical reasoning and careful answers

openai/o3-pro 2025-06-10 200K context $20/M input $80/M output
8 providers
Reasoning Tools

Balanced Claude model for coding, analysis, agent workflows, and cost control

anthropic/claude-sonnet-4-20250514 2025-05-22 200K context $3/M input $15/M output
11 providers
Reasoning Tools

Flagship Claude model for deep reasoning, coding, and long-horizon agents

anthropic/claude-opus-4-20250514 2025-05-22 200K context $15/M input $75/M output
8 providers

Flagship Claude model for deep reasoning, coding, and long-horizon agents

anthropic/claude-opus-4-0 2025-05-22 200K context Input not listed Output not listed
2 providers
Reasoning Tools JSON

Balanced Claude model for coding, analysis, agent workflows, and cost control

anthropic/claude-sonnet-4-0 2025-05-22 200K context $2.898/M input $14.493/M output
4 providers

Mistral model for multilingual chat, reasoning, and tool-assisted workflows

mistral/mistral-medium-2505 2025-05-07 131.072K context $0.4/M input $2/M output
10 providers
Reasoning

Deliberate o-series reasoner for hard math, coding, and multi-step analysis

openai/o3 2025-04-16 200K context $2/M input $8/M output
20 providers
Reasoning Tools JSON

Fast o-series model for compact reasoning, coding, and tool use

openai/o4-mini 2025-04-16 200K context $1.1/M input $4.4/M output
20 providers

Tiny GPT-4.1 option for classification, routing, and very high-volume tasks

openai/gpt-4.1-nano 2025-04-14 1.04758M context $0.1/M input $0.4/M output
20 providers

Affordable GPT-4.1 lane for fast coding help and structured extraction

openai/gpt-4.1-mini 2025-04-14 1.04758M context $0.4/M input $1.6/M output
25 providers
Tools JSON

Long-lived GPT workhorse for coding, instruction following, and production apps

openai/gpt-4.1 2025-04-14 1.04758M context $2/M input $8/M output
28 providers
Reasoning Tools

Balanced Claude model for coding, analysis, agent workflows, and cost control

anthropic/claude-3-7-sonnet-20250219 2025-02-19 200K context $3/M input $15/M output
4 providers
Reasoning Tools JSON

Smaller o-series reasoner for economical coding, math, and planning tasks

openai/o3-mini 2024-12-20 200K context $1.1/M input $4.4/M output
19 providers
Reasoning

O-series reasoning model for hard analysis, math, coding, and planning

openai/o1 2024-12-05 200K context $15/M input $60/M output
16 providers
Tools

GPT model for general reasoning, writing, coding, and tool-assisted tasks

openai/gpt-4o-2024-11-20 2024-11-20 128K context $2.5/M input $10/M output
8 providers

Balanced Claude model for coding, analysis, agent workflows, and cost control

anthropic/claude-3-5-sonnet-20241022 2024-10-22 200K context Input not listed Output not listed
1 provider
Tools

Fast Claude model for responsive assistance, classification, and lightweight agents

anthropic/claude-3-5-haiku-20241022 2024-10-22 200K context $0.8/M input $4/M output
3 providers
Tools

GPT model for general reasoning, writing, coding, and tool-assisted tasks

openai/gpt-4o-2024-08-06 2024-08-06 128K context $2.5/M input $10/M output
7 providers

Small omni GPT for cheap multimodal assistance and production-scale traffic

openai/gpt-4o-mini 2024-07-18 128K context $0.15/M input $0.6/M output
23 providers
Tools JSON

GPT model for general reasoning, writing, coding, and tool-assisted tasks

openai/gpt-4o-2024-05-13 2024-05-13 128K context $5/M input $15/M output
5 providers
Tools

Omni-era GPT for multimodal chat, practical coding, and general assistants

openai/gpt-4o 2024-05-13 128K context $2.5/M input $10/M output
23 providers
Tools

Flagship Qwen model for complex reasoning, coding, and agentic workflows

alibaba/qwen-max 2024-04-03 32.768K context $1.6/M input $6.4/M output
6 providers
JSON

Deeper Sonar search model with broader retrieval and stronger synthesis

perplexity/sonar-pro 2024-01-01 200K context $3/M input $15/M output
9 providers
JSON

Fast web-grounded Sonar for current answers, citations, and lightweight retrieval

perplexity/sonar 2024-01-01 128K context $1/M input $1/M output
9 providers

Compact GPT model for low-latency assistance and high-volume workloads

openai/gpt-4-turbo 2023-11-06 128K context $10/M input $30/M output
15 providers
Tools

GPT model for general reasoning, writing, coding, and tool-assisted tasks

openai/gpt-4 2023-11-06 8.192K context $30/M input $60/M output
11 providers

Compact GPT model for low-latency assistance and high-volume workloads

openai/gpt-3.5-turbo 2023-03-01 16.385K context $0.5/M input $1.5/M output
13 providers

Mistral Medium 3.1 is an updated version of Mistral Medium 3, which is a high-performance enterprise-grade language model designed to deliver frontier-level capabilities at significantly reduced operational cost. It balances...

mistralai/mistral-medium-3.1:batch 131.072K context $0.2/M input $1/M output

GLM 5.2 is a large-scale reasoning model from Z.ai. It supports text input and output with a 1M-token context window, and is suited for long-horizon agent workflows, project-level software engineering,...

z-ai/glm-5.2 202.752K context $0.6/M input $2/M output

Opus 4.6 is Anthropic’s strongest model for coding and long-running professional tasks. It is built for agents that operate across entire workflows rather than single prompts, making it especially effective...

anthropic/claude-opus-4.6 1M context $5/M input $25/M output

Claude Opus 4.1 is an updated version of Anthropic’s flagship model, offering improved performance in coding, reasoning, and agentic tasks. It achieves 74.5% on SWE-bench Verified and shows notable gains...

anthropic/claude-opus-4.1:batch 200K context $7.5/M input $37.5/M output

Grok 4.5 is a model from SpaceXAI with frontier performance on coding, knowledge work, and STEM.

x-ai/grok-4.5 500K context $2/M input $6/M output

The Meta Llama 3.3 multilingual large language model (LLM) is a pretrained and instruction tuned generative model in 70B (text in/text out). The Llama 3.3 instruction tuned text only model...

meta-llama/llama-3.3-70b-instruct:free 65.536K context Free input Free output

Mistral's cutting-edge language model for coding released end of July 2025. Codestral specializes in low-latency, high-frequency tasks such as fill-in-the-middle (FIM), code correction and test generation. [Blog Post](https://mistral.ai/news/codestral-25-08)

mistralai/codestral-2508 256K context $0.3/M input $0.9/M output

Command A is an open-weights 111B parameter model with a 256k context window focused on delivering great performance across agentic, multilingual, and coding use cases. Compared to other leading proprietary...

cohere/command-a 256K context $2.5/M input $10/M output

Gemma 4 31B Instruct is Google DeepMind's 30.7B dense multimodal model supporting text and image input with text output. Features a 256K token context window, configurable thinking/reasoning mode, native function...

google/gemma-4-31b-it:free 262.144K context Free input Free output

Nex-N2-Pro is an agentic mixture-of-experts model from Nex AGI, with 17B active parameters out of 397B total. Built on the Qwen3.5 architecture, it accepts text and image input and produces...

nex-agi/nex-n2-pro 262.144K context $0.25/M input $1/M output

GPT-5 Image Mini combines OpenAI's advanced language capabilities, powered by [GPT-5 Mini](https://openrouter.ai/openai/gpt-5-mini), with GPT Image 1 Mini for efficient image generation. This natively multimodal model features superior instruction following, text...

openai/gpt-5-image-mini 400K context $2.5/M input $2/M output

North Mini Code is Cohere's first agentic coding model and the debut of its North family. A sparse mixture-of-experts model with 30B total parameters and 3B active, it is optimized...

cohere/north-mini-code:free 256K context Free input Free output