No provider description is available for this model yet.
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
No provider description is available for this model yet.
No provider description is available for this model yet.
Voxtral Small is an enhancement of Mistral Small 3, incorporating state-of-the-art audio input capabilities while retaining best-in-class text performance. It excels at speech transcription, translation and audio understanding. Input audio...
GLM-4.5 is our latest flagship foundation model, purpose-built for agent-based applications. It leverages a Mixture-of-Experts (MoE) architecture and supports a context length of up to 128k tokens. GLM-4.5 delivers significantly...
Ling-2.6-1T is an instant (instruct) model from inclusionAI and the company’s trillion-parameter flagship, designed for real-world agents that require fast execution and high efficiency at scale. It uses a “fast...
No provider description is available for this model yet.
No provider description is available for this model yet.
GLM-4.5V is a vision-language foundation model for multimodal agent applications. Built on a Mixture-of-Experts (MoE) architecture with 106B parameters and 12B activated parameters, it achieves state-of-the-art results in video understanding,...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
The preview GPT-4 model with improved instruction following, JSON mode, reproducible outputs, parallel function calling, and more. Training data: up to Dec 2023. **Note:** heavily rate limited by OpenAI while...
No provider description is available for this model yet.
GPT-5.6 Terra Pro is the same underlying model as [GPT-5.6 Terra](https://openrouter.ai/openai/gpt-5.6-terra), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. Learn more in OpenAI's docs: https://developers.openai.com/api/docs/guides/reasoning#reasoning-mode
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Qwen3 Coder Flash is Alibaba's fast and cost efficient version of their proprietary Qwen3 Coder Plus. It is a powerful coding agent model specializing in autonomous programming via tool calling...
gpt-oss-20b is an open-weight 21B parameter model released by OpenAI under the Apache 2.0 license. It uses a Mixture-of-Experts (MoE) architecture with 3.6B active parameters per forward pass, optimized for...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Claude Opus 4.5 is Anthropic’s frontier reasoning model optimized for complex software engineering, agentic workflows, and long-horizon computer use. It offers strong multimodal capabilities, competitive performance across real-world coding and...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
The largest model in the Ministral 3 family, Ministral 3 14B offers frontier capabilities and performance comparable to its larger Mistral Small 3.2 24B counterpart. A powerful and efficient language...
DeepSeek-V3.1 is a large hybrid reasoning model (671B parameters, 37B active) that supports both thinking and non-thinking modes via prompt templates. It extends the DeepSeek-V3 base with a two-phase long-context...
Fast-mode variant of [Opus 4.7](/anthropic/claude-opus-4.7) - identical capabilities with higher output speed at premium 6x pricing. Learn more in Anthropic's docs: https://platform.claude.com/docs/en/build-with-claude/fast-mode
Llama 3.2 3B is a 3-billion-parameter multilingual large language model, optimized for advanced natural language processing tasks like dialogue generation, reasoning, and summarization. Designed with the latest transformer architecture, it...
Qwen3-Coder-480B-A35B-Instruct is a Mixture-of-Experts (MoE) code generation model developed by the Qwen team. It is optimized for agentic coding tasks such as function calling, tool use, and long-context reasoning over...
No provider description is available for this model yet.
No provider description is available for this model yet.
Mistral Small 4 is the next major release in the Mistral Small family, unifying the capabilities of several flagship Mistral models into a single system. It combines strong reasoning from...
NVIDIA Nemotron 3 Nano 30B A3B is a small language MoE model with highest compute efficiency and accuracy for developers to build specialized agentic AI systems. The model is fully...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
The simplest way to get free inference. openrouter/free is a router that selects free models at random from the models available on OpenRouter. The router smartly filters for models that...
Qwen3-30B-A3B-Instruct-2507 is a 30.5B-parameter mixture-of-experts language model from Qwen, with 3.3B active parameters per inference. It operates in non-thinking mode and is designed for high-quality instruction following, multilingual understanding, and...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Qwen3-Coder-Next is an open-weight causal language model optimized for coding agents and local development workflows. It uses a sparse MoE design with 80B total parameters and only 3B activated per...
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| anthropic.claude-opus-5bedrock_converse/anthropic.claude-opus-5 | 1M | $5 | $25 | — | |||
| eu.anthropic.claude-fable-5bedrock_converse/eu.anthropic.claude-fable-5 | 1M | $11 | $55 | — | |||
| us.anthropic.claude-fable-5bedrock_converse/us.anthropic.claude-fable-5 | 1M | $11 | $55 | — | |||
| Mistral: Voxtral Small 24B 2507mistralai/voxtral-small-24b-2507 | 32.768K | $0.1 | $0.3 | — | |||
| Z.ai: GLM 4.5z-ai/glm-4.5 | 131.072K | $0.6 | $2.2 | — | |||
| inclusionAI: Ling-2.6-1Tinclusionai/ling-2.6-1t | 262.144K | $0.075 | $0.625 | — | |||
| @cf/aisingapore/gemma-sea-lion-v4-27b-itcloudflare/@cf/aisingapore/gemma-sea-lion-v4-27b-it | 128K | $0.351 | $0.555 | — | |||
| global.anthropic.claude-fable-5bedrock_converse/global.anthropic.claude-fable-5 | 1M | $10 | $50 | — | |||
| Z.ai: GLM 4.5Vz-ai/glm-4.5v | 65.536K | $0.6 | $1.8 | — | |||
| global.openai.gpt-5.6-lunabedrock_converse/global.openai.gpt-5.6-luna | 1M | $0.2 | $1.2 | — | |||
| us.openai.gpt-5.6-lunabedrock_converse/us.openai.gpt-5.6-luna | 1M | $0.22 | $1.32 | — | |||
| llama3ollama/llama3 | 8.192K | — | — | — | |||
| OpenAI: GPT-4 Turbo Previewopenai/gpt-4-turbo-preview | 128K | $10 | $30 | — | |||
| glm-5.1dashscope/glm-5.1 | 202.745K | $1.4 | $4.4 | — | |||
| OpenAI: GPT-5.6 Terra Pro (batch)openai/gpt-5.6-terra-pro:batch | 1.05M | $1 | $6 | — | |||
| sa-east-1/deepseek.v3.2bedrock/sa-east-1/deepseek.v3.2 | 163.84K | $0.74 | $2.22 | — | |||
| deepseek-v4-prodashscope/deepseek-v4-pro | 1M | $2.4 | $4.8 | — | |||
| us-gov-west-1/anthropic.claude-3-7-sonnet-20250219-v1:0bedrock/us-gov-west-1/anthropic.claude-3-7-sonnet-20250219-v1:0 | 200K | $3.6 | $18 | — | |||
| Qwen: Qwen3 Coder Flashqwen/qwen3-coder-flash | 1M | $0.195 | $0.975 | — | |||
| OpenAI: gpt-oss-20b (free)openai/gpt-oss-20b:free | 131.072K | Free | Free | — | |||
| global.openai.gpt-5.6-terrabedrock_converse/global.openai.gpt-5.6-terra | 1M | $2 | $12 | — | |||
| deepseek-v4-flash-0731dashscope/deepseek-v4-flash-0731 | 1M | $0.2 | $0.4 | — | |||
| sa-east-1/meta.llama3-8b-instruct-v1:0bedrock/sa-east-1/meta.llama3-8b-instruct-v1:0 | 8.192K | $0.5 | $1.01 | — | |||
| deepseek-v4-flashdashscope/deepseek-v4-flash | 1M | $0.2 | $0.4 | — | |||
| us.openai.gpt-5.6-terrabedrock_converse/us.openai.gpt-5.6-terra | 1M | $2.2 | $13.2 | — | |||
| sa-east-1/meta.llama3-70b-instruct-v1:0bedrock/sa-east-1/meta.llama3-70b-instruct-v1:0 | 8.192K | $4.45 | $5.88 | — | |||
| Anthropic: Claude Opus 4.5anthropic/claude-opus-4.5 | 200K | $5 | $25 | — | |||
| @cf/nvidia/nemotron-3-120b-a12bcloudflare/@cf/nvidia/nemotron-3-120b-a12b | 256K | $0.5 | $1.5 | — | |||
| global.openai.gpt-5.6-solbedrock_converse/global.openai.gpt-5.6-sol | 1M | $5 | $30 | — | |||
| us-gov-west-1/amazon.titan-text-premier-v1:0bedrock/us-gov-west-1/amazon.titan-text-premier-v1:0 | 42K | $0.5 | $1.5 | — | |||
| Mistral: Ministral 3 14B 2512mistralai/ministral-14b-2512 | 262.144K | $0.2 | $0.2 | — | |||
| DeepSeek: DeepSeek V3.1deepseek/deepseek-chat-v3.1 | 163.84K | $0.25 | $0.95 | — | |||
| Anthropic: Claude Opus 4.7 (Fast)anthropic/claude-opus-4.7-fast | 1M | $30 | $150 | — | |||
| Meta: Llama 3.2 3B Instructmeta-llama/llama-3.2-3b-instruct | 131.072K | $0.05 | $0.33 | — | |||
| Qwen: Qwen3 Coder 480B A35B (free)qwen/qwen3-coder:free | 262K | Free | Free | — | |||
| us-gov-west-1/amazon.titan-text-lite-v1bedrock/us-gov-west-1/amazon.titan-text-lite-v1 | 42K | $0.3 | $0.4 | — | |||
| us.openai.gpt-5.6-solbedrock_converse/us.openai.gpt-5.6-sol | 1M | $5.5 | $33 | — | |||
| Mistral: Mistral Small 4 (batch)mistralai/mistral-small-2603:batch | 262.144K | $0.075 | $0.3 | — | |||
| NVIDIA: Nemotron 3 Nano 30B A3B (free)nvidia/nemotron-3-nano-30b-a3b:free | 256K | Free | Free | — | |||
| us-gov-west-1/amazon.titan-text-express-v1bedrock/us-gov-west-1/amazon.titan-text-express-v1 | 42K | $1.3 | $1.7 | — | |||
| Qwen3.8-Maxscx-ai/qwen3.8-max | 1M | $1.65 | $4.99 | — | |||
| anthropic.claude-fable-5bedrock_converse/anthropic.claude-fable-5 | 1M | $10 | $50 | — | |||
| GLM-5.2scx-ai/glm-5.2 | 1.04858M | $0.61 | $1.98 | — | |||
| Free Models Routeropenrouter/free | 200K | Free | Free | — | |||
| Qwen: Qwen3 30B A3B Instruct 2507qwen/qwen3-30b-a3b-instruct-2507 | 128K | $0.048 | $0.193 | — | |||
| us-gov-west-1/amazon.nova-pro-v1:0bedrock/us-gov-west-1/amazon.nova-pro-v1:0 | 300K | $0.96 | $3.84 | — | |||
| gpt-audio-miniazure/gpt-audio-mini | 128K | $0.6 | $2.4 | — | |||
| au.anthropic.claude-opus-4-7bedrock_converse/au.anthropic.claude-opus-4-7 | 1M | $5.5 | $27.5 | — | |||
| google.gemma-3-12b-itbedrock_converse/google.gemma-3-12b-it | 128K | $0.09 | $0.29 | — | |||
| Qwen: Qwen3 Coder Nextqwen/qwen3-coder-next | 262.144K | $0.12 | $0.8 | — |