Deeper Sonar search model with broader retrieval and stronger synthesis
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
Compact GPT model for low-latency assistance and high-volume workloads
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Qwen-Plus, based on the Qwen2.5 foundation model, is a 131K context model with a balanced performance, speed, and cost combination.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
gpt-oss-120b is an open-weight, 117B-parameter Mixture-of-Experts (MoE) language model from OpenAI designed for high-reasoning, agentic, and general-purpose production use cases. It activates 5.1B parameters per forward pass and is optimized...
Hermes 4 is a large-scale reasoning model built on Meta-Llama-3.1-405B and released by Nous Research. It introduces a hybrid reasoning mode, where the model can choose to deliberate internally with...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
This model always redirects to the latest model in the GLM Flash family.
No provider description is available for this model yet.
No provider description is available for this model yet.
GPT-5 Pro is OpenAI’s most advanced model, offering major improvements in reasoning, code quality, and user experience. It is optimized for complex tasks that require step-by-step reasoning, instruction following, and...
Amazon Nova Pro 1.0 is a capable multimodal model from Amazon focused on providing a combination of accuracy, speed, and cost for a wide range of tasks. As of December...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
OpenAI o4-mini is a compact reasoning model in the o-series, optimized for fast, cost-efficient performance while retaining strong multimodal and agentic capabilities. It supports tool use and demonstrates competitive reasoning...
This model always redirects to the latest model in the GPT Luna family.
GLM-5.3-Flash is a native multimodal model from Z.ai. It is suited for efficient coding and long-horizon agent tasks. Its hybrid sparse and linear attention architecture maintains accurate long-context behavior while...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Gemini 2.5 Flash-Lite is a lightweight reasoning model in the Gemini 2.5 family, optimized for ultra-low latency and cost efficiency. It offers improved throughput, faster token generation, and better performance...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
OpenAI o3-mini is a cost-efficient language model optimized for STEM reasoning tasks, particularly excelling in science, mathematics, and coding. This model supports the `reasoning_effort` parameter, which can be set to...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
The Auto Router automatically selects the best model for your prompt, powered by the wisdom of the market. It routes you based on what the OpenRouter community collectively spends on...
Qwen3-Coder-30B-A3B-Instruct is a 30.5B parameter Mixture-of-Experts (MoE) model with 128 experts (8 active per forward pass), designed for advanced code generation, repository-scale understanding, and agentic tool use. Built on the...
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| Sonar Properplexity/sonar-pro | 200K | $3 | $15 | 2024-01-01 | |||
| GPT-4 Turboopenai/gpt-4-turbo | 128K | $10 | $30 | 2023-11-06 | |||
| amazon.nova-micro-v1:0bedrock_converse/amazon.nova-micro-v1:0 | 128K | $0.035 | $0.14 | — | |||
| us.amazon.nova-2-pro-preview-20251202-v1:0bedrock_converse/us.amazon.nova-2-pro-preview-20251202-v1:0 | 1M | $2.188 | $17.5 | — | |||
| glm-5-2mistral/glm-5-2 | 1.04858M | $1.4 | $4.4 | — | |||
| us.amazon.nova-2-lite-v1:0bedrock_converse/us.amazon.nova-2-lite-v1:0 | 1M | $0.33 | $2.75 | — | |||
| eu.amazon.nova-2-pro-preview-20251202-v1:0bedrock_converse/eu.amazon.nova-2-pro-preview-20251202-v1:0 | 1M | $2.188 | $17.5 | — | |||
| zai-glm-5-2mistral/zai-glm-5-2 | 1.04858M | $1.4 | $4.4 | — | |||
| eu.amazon.nova-2-lite-v1:0bedrock_converse/eu.amazon.nova-2-lite-v1:0 | 1M | $0.33 | $2.75 | — | |||
| apac.amazon.nova-2-pro-preview-20251202-v1:0bedrock_converse/apac.amazon.nova-2-pro-preview-20251202-v1:0 | 1M | $2.188 | $17.5 | — | |||
| chat-latestopenai/chat-latest | 400K | $5 | $30 | — | |||
| apac.amazon.nova-2-lite-v1:0bedrock_converse/apac.amazon.nova-2-lite-v1:0 | 1M | $0.33 | $2.75 | — | |||
| daybreak-blue-latestopenai/daybreak-blue-latest | 1.05M | $5 | $30 | — | |||
| Qwen: Qwen-Plusqwen/qwen-plus | 1M | $0.26 | $0.78 | — | |||
| amazon.nova-2-pro-preview-20251202-v1:0bedrock_converse/amazon.nova-2-pro-preview-20251202-v1:0 | 1M | $2.188 | $17.5 | — | |||
| amazon.nova-2-lite-v1:0bedrock_converse/amazon.nova-2-lite-v1:0 | 1M | $0.3 | $2.5 | — | |||
| daybreak-red-latestopenai/daybreak-red-latest | 400K | $12.5 | $75 | — | |||
| amazon.nova-lite-v1:0bedrock_converse/amazon.nova-lite-v1:0 | 300K | $0.06 | $0.24 | — | |||
| OpenAI: gpt-oss-120b (free)openai/gpt-oss-120b:free | 131.072K | Free | Free | — | |||
| Nous: Hermes 4 405Bnousresearch/hermes-4-405b | 131.072K | $1 | $3 | — | |||
| writer.palmyra-x5-v1:0bedrock_converse/writer.palmyra-x5-v1:0 | 1M | $0.6 | $6 | — | |||
| writer.palmyra-x4-v1:0bedrock_converse/writer.palmyra-x4-v1:0 | 128K | $2.5 | $10 | — | |||
| sa-east-1/qwen.qwen3-coder-nextbedrock/sa-east-1/qwen.qwen3-coder-next | 262.144K | $0.6 | $1.44 | — | |||
| Z.ai: GLM Flash Latest~z-ai/glm-flash-latest | 1.04858M | $0.075 | $0.25 | — | |||
| us.writer.palmyra-x5-v1:0bedrock_converse/us.writer.palmyra-x5-v1:0 | 1M | $0.6 | $6 | — | |||
| sa-east-1/moonshotai.kimi-k2.5bedrock/sa-east-1/moonshotai.kimi-k2.5 | 262.144K | $0.72 | $3.6 | — | |||
| OpenAI: GPT-5 Pro (batch)openai/gpt-5-pro:batch | 400K | $7.5 | $60 | — | |||
| Amazon: Nova Pro 1.0amazon/nova-pro-v1 | 300K | $0.8 | $3.2 | — | |||
| us.writer.palmyra-x4-v1:0bedrock_converse/us.writer.palmyra-x4-v1:0 | 128K | $2.5 | $10 | — | |||
| gpt-5.6-cyberopenai/gpt-5.6-cyber | 400K | $12.5 | $75 | — | |||
| ai21.jamba-1-5-mini-v1:0bedrock/ai21.jamba-1-5-mini-v1:0 | 256K | $0.2 | $0.4 | — | |||
| ai21.jamba-1-5-large-v1:0bedrock/ai21.jamba-1-5-large-v1:0 | 256K | $2 | $8 | — | |||
| OpenAI: o4 Mini (batch)openai/o4-mini:batch | 200K | $0.55 | $2.2 | — | |||
| OpenAI: GPT Luna Latest~openai/gpt-luna-latest | 1.05M | $0.2 | $1.2 | — | |||
| Z.ai: GLM 5.3 Flash (batch)z-ai/glm-5.3-flash:batch | 1.04858M | $0.075 | $0.25 | — | |||
| sa-east-1/moonshotai.kimi-k2-thinkingbedrock/sa-east-1/moonshotai.kimi-k2-thinking | 262.144K | $0.73 | $3.03 | — | |||
| us-east-2/qwen.qwen3-coder-nextbedrock/us-east-2/qwen.qwen3-coder-next | 262.144K | $0.5 | $1.2 | — | |||
| sa-east-1/minimax.minimax-m2.5bedrock/sa-east-1/minimax.minimax-m2.5 | 1M | $0.36 | $1.44 | — | |||
| Google: Gemini 2.5 Flash Lite (batch)google/gemini-2.5-flash-lite:batch | 1.04858M | $0.05 | $0.2 | — | |||
| us-east-2/moonshotai.kimi-k2.5bedrock/us-east-2/moonshotai.kimi-k2.5 | 262.144K | $0.6 | $3 | — | |||
| us-east-2/moonshotai.kimi-k2-thinkingbedrock/us-east-2/moonshotai.kimi-k2-thinking | 262.144K | $0.6 | $2.5 | — | |||
| gpt-6-astraazure_ai/gpt-6-astra | 922K | $10 | $50 | — | |||
| us-east-2/minimax.minimax-m2.5bedrock/us-east-2/minimax.minimax-m2.5 | 1M | $0.3 | $1.2 | — | |||
| eu-north-1/moonshotai.kimi-k2.5bedrock/eu-north-1/moonshotai.kimi-k2.5 | 262.144K | $0.72 | $3.6 | — | |||
| OpenAI: o3 Mini (batch)openai/o3-mini:batch | 200K | $0.55 | $2.2 | — | |||
| gpt-5-minigithub_copilot/gpt-5-mini | 128K | — | — | — | |||
| gpt-5github_copilot/gpt-5 | 128K | — | — | — | |||
| eu-north-1/minimax.minimax-m2.5bedrock/eu-north-1/minimax.minimax-m2.5 | 1M | $0.36 | $1.44 | — | |||
| Auto Routeropenrouter/auto | 2M | — | — | — | |||
| Qwen: Qwen3 Coder 30B A3B Instructqwen/qwen3-coder-30b-a3b-instruct | 262.144K | $0.07 | $0.28 | — |