209 models
Reasoning Tools

Flagship Claude model for deep reasoning, coding, and long-horizon agents

anthropic/claude-opus-4-5 2025-11-24 200K context $5/M input $25/M output
17 providers
Reasoning Tools JSON

Preview Gemini flagship for complex reasoning, coding, and rich multimodal prompts

google/gemini-3-pro-preview 2025-11-18 1.04858M context $0.57/M input $3.43/M output
10 providers
Reasoning Tools JSON

Flagship Claude model for deep reasoning, coding, and long-horizon agents

anthropic/claude-opus-4-5-20251101 2025-11-01 200K context $5/M input $25/M output
17 providers
Tools JSON

Fast Claude model for responsive assistance, classification, and lightweight agents

anthropic/claude-haiku-4-5-20251001 2025-10-15 200K context $1/M input $5/M output
20 providers
Reasoning Tools

Fast Claude lane for lightweight agents, office tasks, and responsive chat

anthropic/claude-haiku-4-5 2025-10-15 200K context $1/M input $5/M output
23 providers
Reasoning Tools

Balanced Claude model for coding, analysis, agent workflows, and cost control

anthropic/claude-sonnet-4-5 2025-09-29 200K context $3/M input $15/M output
20 providers
Reasoning Tools JSON

Balanced Claude model for coding, analysis, agent workflows, and cost control

anthropic/claude-sonnet-4-5-20250929 2025-09-29 200K context $3/M input $15/M output
18 providers
Reasoning Tools JSON

Flagship Claude model for deep reasoning, coding, and long-horizon agents

anthropic/claude-opus-4-1-20250805 2025-08-05 200K context $15/M input $75/M output
13 providers
Reasoning Tools

Flagship Claude model for deep reasoning, coding, and long-horizon agents

anthropic/claude-opus-4-1 2025-08-05 200K context $15/M input $75/M output
9 providers
Reasoning Tools JSON

Fast Gemini workhorse for multimodal apps where latency and price matter

google/gemini-2.5-flash 2025-06-17 1.04858M context $0.3/M input $2.5/M output
30 providers
Reasoning Tools JSON

Lean Gemini 2.5 lane for cheap multimodal traffic and quick agents

google/gemini-2.5-flash-lite 2025-06-17 1.04858M context $0.1/M input $0.4/M output
18 providers
Reasoning Tools JSON

Google's proven reasoning model for coding, math, and multimodal analysis

google/gemini-2.5-pro 2025-06-17 1.04858M context $1.25/M input $10/M output
29 providers
Reasoning Tools JSON

Balanced Claude model for coding, analysis, agent workflows, and cost control

anthropic/claude-sonnet-4-0 2025-05-22 200K context $2.898/M input $14.493/M output
4 providers
Reasoning Tools

Balanced Claude model for coding, analysis, agent workflows, and cost control

anthropic/claude-sonnet-4-20250514 2025-05-22 200K context $3/M input $15/M output
11 providers
Reasoning Tools

Flagship Claude model for deep reasoning, coding, and long-horizon agents

anthropic/claude-opus-4-20250514 2025-05-22 200K context $15/M input $75/M output
8 providers

Flagship Claude model for deep reasoning, coding, and long-horizon agents

anthropic/claude-opus-4-0 2025-05-22 200K context Input not listed Output not listed
2 providers

Multimodal model for complex analysis, long-context understanding, tool use, and model distillation

amazon/nova-premier 2025-04-30 1M context Input not listed Output not listed
Reasoning

Deliberate o-series reasoner for hard math, coding, and multi-step analysis

openai/o3 2025-04-16 200K context $2/M input $8/M output
20 providers
Tools JSON

Long-lived GPT workhorse for coding, instruction following, and production apps

openai/gpt-4.1 2025-04-14 1.04758M context $2/M input $8/M output
28 providers

Affordable GPT-4.1 lane for fast coding help and structured extraction

openai/gpt-4.1-mini 2025-04-14 1.04758M context $0.4/M input $1.6/M output
25 providers
Reasoning Tools

Balanced Claude model for coding, analysis, agent workflows, and cost control

anthropic/claude-3-7-sonnet-20250219 2025-02-19 200K context $3/M input $15/M output
4 providers
Tools

Low-latency Gemini model for high-volume multimodal and agent workloads

google/gemini-2.0-flash-lite 2024-12-11 1.04858M context $0.052/M input $0.21/M output
2 providers
Tools

Earlier Gemini Flash workhorse for responsive multimodal apps and tool use

google/gemini-2.0-flash 2024-12-11 1.04858M context $0.1/M input $0.42/M output
2 providers
Reasoning

O-series reasoning model for hard analysis, math, coding, and planning

openai/o1 2024-12-05 200K context $15/M input $60/M output
16 providers
Tools

Efficient model for low-latency assistance, extraction, and routine automation

amazon/nova-lite 2024-12-03 300K context $0.06/M input $0.24/M output
3 providers
Tools

Flagship model for demanding analysis, coding, and production agent workflows

amazon/nova-pro 2024-12-03 300K context $0.8/M input $3.2/M output
3 providers
Tools

Fast Claude model for responsive assistance, classification, and lightweight agents

anthropic/claude-3-5-haiku-20241022 2024-10-22 200K context $0.8/M input $4/M output
3 providers

Balanced Claude model for coding, analysis, agent workflows, and cost control

anthropic/claude-3-5-sonnet-20241022 2024-10-22 200K context Input not listed Output not listed
1 provider

Small omni GPT for cheap multimodal assistance and production-scale traffic

openai/gpt-4o-mini 2024-07-18 128K context $0.15/M input $0.6/M output
23 providers
Tools

Omni-era GPT for multimodal chat, practical coding, and general assistants

openai/gpt-4o 2024-05-13 128K context $2.5/M input $10/M output
23 providers
Tools

Legacy model retained for compatibility with older integrations

anthropic/claude-3-haiku-20240307 2024-03-13 200K context $0.25/M input $1.25/M output
2 providers

GPT Chat Latest points to OpenAI's stable API alias `chat-latest` that always resolves to the latest Instant chat model used in ChatGPT. As OpenAI rolls out new Instant model updates...

openai/gpt-chat-latest 400K context $5/M input $30/M output

Muse Spark 1.2 contributor tier is a reasoning model from Meta designed for developers who want to start building at an even lower cost. It’s meaningfully cheaper than Muse Spark...

meta/muse-spark-1.2-contributor 1.04858M context $0.1/M input $0.2/M output

Mistral Medium 3.5 is a dense 128B instruction-following model from Mistral AI. It supports text and image inputs with text output, and is designed for agentic workflows, coding, and complex...

mistralai/mistral-medium-3-5 262.144K context $1.5/M input $7.5/M output

Nova 2 Lite is a fast, cost-effective reasoning model for everyday workloads that can process text, images, and videos to generate text. Nova 2 Lite demonstrates standout capabilities in processing...

amazon/nova-2-lite-v1 1M context $0.3/M input $2.5/M output

Claude Opus 4.8 is Anthropic's most capable generally available model in the Opus family. It supports text, image, and file inputs with text output, with reasoning support and a 1M-token...

anthropic/claude-opus-4.8 1M context $5/M input $25/M output

Sonnet 5 is Anthropic's most capable Sonnet-class model, with frontier performance across coding, agents, and professional work. It supports adaptive thinking with selectable reasoning effort levels (low, medium, high, max,...

anthropic/claude-sonnet-5:batch 1M context $1/M input $5/M output

This model always redirects to the latest model in the OpenAI GPT family.

~openai/gpt-latest 1.05M context $2/M input $10/M output

Grok 4.5 is a model from SpaceXAI with frontier performance on coding, knowledge work, and STEM.

x-ai/grok-4.5 500K context $2/M input $6/M output

OpenAI o3-mini-high is the same model as [o3-mini](/openai/o3-mini) with reasoning_effort set to high. o3-mini is a cost-efficient language model optimized for STEM reasoning tasks, particularly excelling in science, mathematics, and...

openai/o3-mini-high:batch 200K context $0.55/M input $2.2/M output

GPT-5.2 Chat (AKA Instant) is the fast, lightweight member of the 5.2 family, optimized for low-latency chat while retaining strong general intelligence. It uses adaptive reasoning to selectively “think” on...

openai/gpt-5.2-chat 128K context $1.75/M input $14/M output

Auto Router (Beta) is a task-aware router from OpenRouter. It classifies each request, then routes it the [most popular model](/rankings#task-spend) for that task based on aggregate spend, filtered by your...

openrouter/auto-beta 2M context Input not listed Output not listed

Mistral Medium 3 is a high-performance enterprise-grade language model designed to deliver frontier-level capabilities at significantly reduced operational cost. It balances state-of-the-art reasoning and multimodal performance with 8× lower cost...

mistralai/mistral-medium-3 131.072K context $0.4/M input $2/M output

Gemini 2.5 Pro is Google’s state-of-the-art AI model designed for advanced reasoning, coding, mathematics, and scientific tasks. It employs “thinking” capabilities, enabling it to reason through responses with enhanced accuracy...

google/gemini-2.5-pro:batch 1.04858M context $0.625/M input $5/M output

This model always redirects to the latest model in the Google Gemini Flash family.

~google/gemini-flash-latest 1.04858M context $0.75/M input $3.75/M output

Claude Opus 4.1 is an updated version of Anthropic’s flagship model, offering improved performance in coding, reasoning, and agentic tasks. It achieves 74.5% on SWE-bench Verified and shows notable gains...

anthropic/claude-opus-4.1:batch 200K context $7.5/M input $37.5/M output

GPT-5 Mini is a compact version of GPT-5, designed to handle lighter-weight reasoning tasks. It provides the same instruction-following and safety-tuning benefits as GPT-5, but with reduced latency and cost....

openai/gpt-5-mini:batch 400K context $0.125/M input $1/M output

This model always redirects to the latest model in the Google Gemini Pro family.

~google/gemini-pro-latest 1.04858M context $2/M input $12/M output