No provider description is available for this model yet.
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
No provider description is available for this model yet.
DeepSeek V3, a 685B-parameter, mixture-of-experts model, is the latest iteration of the flagship chat model family from the DeepSeek team. It succeeds the [DeepSeek V3](/deepseek/deepseek-chat-v3) model and performs really well...
OpenAI o3-mini-high is the same model as [o3-mini](/openai/o3-mini) with reasoning_effort set to high. o3-mini is a cost-efficient language model optimized for STEM reasoning tasks, particularly excelling in science, mathematics, and...
No provider description is available for this model yet.
No provider description is available for this model yet.
Llama 4 Scout 17B Instruct (16E) is a mixture-of-experts (MoE) language model developed by Meta, activating 17 billion parameters out of a total of 109B. It supports native multimodal input...
No provider description is available for this model yet.
No provider description is available for this model yet.
Qwen3-32B is a dense 32.8B parameter causal language model from the Qwen3 series, optimized for both complex reasoning and efficient dialogue. It supports seamless switching between a "thinking" mode for...
Mistral Saba is a 24B-parameter language model specifically designed for the Middle East and South Asia, delivering accurate and contextually relevant responses while maintaining efficient performance. Trained on curated regional...
Ling-2.6-flash is an instant (instruct) model from inclusionAI with 104B total parameters and 7.4B active parameters, designed for real-world agents that require fast responses, strong execution, and high token efficiency....
Qwen2.5 72B is the latest series of Qwen large language models. Qwen2.5 brings the following improvements upon Qwen2: - Significantly more knowledge and has greatly improved capabilities in coding and...
GPT-5.6 Luna Pro is the same underlying model as [GPT-5.6 Luna](https://openrouter.ai/openai/gpt-5.6-luna), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. Learn more in OpenAI's docs: https://developers.openai.com/api/docs/guides/reasoning#reasoning-mode
No provider description is available for this model yet.
No provider description is available for this model yet.
Qwen3-8B is a dense 8.2B parameter causal language model from the Qwen3 series, designed for both reasoning-heavy tasks and efficient dialogue. It supports seamless switching between "thinking" mode for math,...
No provider description is available for this model yet.
No provider description is available for this model yet.
Qwen3, the latest generation in the Qwen large language model series, features both dense and mixture-of-experts (MoE) architectures to excel in reasoning, multilingual support, and advanced agent tasks. Its unique...
Llama 4 Maverick 17B Instruct (128E) is a high-capacity multimodal language model from Meta, built on a mixture-of-experts (MoE) architecture with 128 experts and 17 billion active parameters per forward...
No provider description is available for this model yet.
No provider description is available for this model yet.
May 28th update to the [original DeepSeek R1](/deepseek/deepseek-r1) Performance on par with [OpenAI o1](/openai/o1), but open-sourced and with fully open reasoning tokens. It's 671B parameters in size, with 37B active...
No provider description is available for this model yet.
No provider description is available for this model yet.
ERNIE-4.5-VL-424B-A47B is a multimodal Mixture-of-Experts (MoE) model from Baidu’s ERNIE 4.5 series, featuring 424B total parameters with 47B active per token. It is trained jointly on text and image data...
Skyfall 36B v2 is an enhanced iteration of Mistral Small 2501, specifically fine-tuned for improved creativity, nuanced writing, role-playing, and coherent storytelling.
Qwen3.6 27B is a dense 27-billion-parameter language model from the Qwen Team at Alibaba, released in April 2026. It features hybrid multimodal capabilities — accepting text, image, and video inputs...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Qwen3-235B-A22B-Instruct-2507 is a multilingual, instruction-tuned mixture-of-experts language model based on the Qwen3-235B architecture, with 22B active parameters per forward pass. It is optimized for general-purpose text generation, including instruction following,...
No provider description is available for this model yet.
GPT-6 Astra Pro is the same underlying model as [GPT-6 Astra](https://openrouter.ai/openai/gpt-6-astra), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. Learn more in OpenAI's docs: https://developers.openai.com/api/docs/guides/reasoning#reasoning-mode
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
A cost-efficient version of GPT Audio. The new snapshot features an upgraded decoder for more natural sounding voices and maintains better voice consistency. Input is priced at $0.60 per million...
No provider description is available for this model yet.
No provider description is available for this model yet.
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| us-east-1/anthropic.claude-v2:1bedrock/us-east-1/anthropic.claude-v2:1 | 100K | $8 | $24 | — | |||
| us-east-1/anthropic.claude-v1bedrock/us-east-1/anthropic.claude-v1 | 100K | $8 | $24 | — | |||
| DeepSeek: DeepSeek V3 0324deepseek/deepseek-chat-v3-0324 | 163.84K | $0.25 | $1 | — | |||
| OpenAI: o3 Mini Highopenai/o3-mini-high | 200K | $1.1 | $4.4 | — | |||
| global.xai.grok-4.6bedrock_converse/global.xai.grok-4.6 | 500K | $2 | $6 | — | |||
| us.xai.grok-4.6bedrock_converse/us.xai.grok-4.6 | 500K | $2.2 | $6.6 | — | |||
| Meta: Llama 4 Scoutmeta-llama/llama-4-scout | 327.68K | $0.1 | $0.3 | — | |||
| us-east-1/anthropic.claude-instant-v1bedrock/us-east-1/anthropic.claude-instant-v1 | 100K | $0.8 | $2.4 | — | |||
| us-gov-west-1/amazon.nova-micro-v1:0bedrock/us-gov-west-1/amazon.nova-micro-v1:0 | 128K | $0.042 | $0.168 | — | |||
| Qwen: Qwen3 32Bqwen/qwen3-32b | 40.96K | $0.08 | $0.28 | — | |||
| Mistral: Sabamistralai/mistral-saba | 32.768K | $0.2 | $0.6 | — | |||
| inclusionAI: Ling-2.6-flashinclusionai/ling-2.6-flash | 262.144K | $0.01 | $0.03 | — | |||
| Qwen2.5 72B Instructqwen/qwen-2.5-72b-instruct | 32.768K | $0.36 | $0.4 | — | |||
| OpenAI: GPT-5.6 Luna Proopenai/gpt-5.6-luna-pro | 1.05M | $0.2 | $1.2 | — | |||
| us-gov-west-1/amazon.nova-lite-v1:0bedrock/us-gov-west-1/amazon.nova-lite-v1:0 | 300K | $0.072 | $0.288 | — | |||
| llama3.1ollama/llama3.1 | 8.192K | — | — | — | |||
| Qwen: Qwen3 8Bqwen/qwen3-8b | 131.072K | $0.117 | $0.455 | — | |||
| us-east-1/6-month-commitment/anthropic.claude-v2:1bedrock/us-east-1/6-month-commitment/anthropic.claude-v2:1 | 100K | — | — | — | |||
| xai.grok-4.6bedrock_mantle/xai.grok-4.6 | 500K | $2.2 | $6.6 | — | |||
| Qwen: Qwen3 30B A3Bqwen/qwen3-30b-a3b | 40.96K | $0.12 | $0.5 | — | |||
| Meta: Llama 4 Maverickmeta-llama/llama-4-maverick | 128K | $0.2 | $0.696 | — | |||
| deepseek/deepseek-v4-pro-0813openrouter/deepseek/deepseek-v4-pro-0813 | 1.04858M | $1.32 | $3.96 | — | |||
| us-east-1/6-month-commitment/anthropic.claude-v1bedrock/us-east-1/6-month-commitment/anthropic.claude-v1 | 100K | — | — | — | |||
| DeepSeek: R1 0528deepseek/deepseek-r1-0528 | 163.84K | $0.5 | $2.15 | — | |||
| deepseek/deepseek-v4-proopenrouter/deepseek/deepseek-v4-pro | 1.04858M | $1.32 | $3.96 | — | |||
| anthropic/claude-opus-5openrouter/anthropic/claude-opus-5 | 1M | $5 | $25 | — | |||
| Baidu: ERNIE 4.5 VL 424B A47B baidu/ernie-4.5-vl-424b-a47b | 123K | $0.42 | $1.25 | — | |||
| TheDrummer: Skyfall 36B V2thedrummer/skyfall-36b-v2 | 32.768K | $0.55 | $0.8 | — | |||
| Qwen: Qwen3.6 27Bqwen/qwen3.6-27b | 262.144K | $0.3 | $2 | — | |||
| microsoft/Dayhoff-170M-GRS-SS-134000microsoft/Dayhoff-170M-GRS-SS-134000 | Not documented | — | — | — | |||
| glm-5-2mistral/glm-5-2 | 1.04858M | $1.4 | $4.4 | — | |||
| google/gemma-4-E4B-it-qat-w4a16-ctgoogle/gemma-4-E4B-it-qat-w4a16-ct | Not documented | — | — | — | |||
| us-gov-east-1/anthropic.claude-sonnet-4-5-20250929-v1:0bedrock/us-gov-east-1/anthropic.claude-sonnet-4-5-20250929-v1:0 | 200K | $3.6 | $18 | — | |||
| zai-glm-5-2mistral/zai-glm-5-2 | 1.04858M | $1.4 | $4.4 | — | |||
| Qwen: Qwen3 235B A22B Instruct 2507qwen/qwen3-235b-a22b-2507 | 262.144K | $0.087 | $0.35 | — | |||
| microsoft/Dayhoff-3b-UR90-30000microsoft/Dayhoff-3b-UR90-30000 | Not documented | — | — | — | |||
| OpenAI: GPT-6 Astra Proopenai/gpt-6-astra-pro | 1.05M | $10 | $50 | — | |||
| us-gov-east-1/claude-sonnet-4-5-20250929-v1:0bedrock/us-gov-east-1/claude-sonnet-4-5-20250929-v1:0 | 200K | $3.6 | $18 | — | |||
| eu.anthropic.claude-opus-4-7bedrock_converse/eu.anthropic.claude-opus-4-7 | 1M | $5.5 | $27.5 | — | |||
| us-gov-east-1/meta.llama3-70b-instruct-v1:0bedrock/us-gov-east-1/meta.llama3-70b-instruct-v1:0 | 8K | $2.65 | $3.5 | — | |||
| microsoft/Dayhoff-170M-UR90-HL-16000microsoft/Dayhoff-170M-UR90-HL-16000 | Not documented | — | — | — | |||
| microsoft/Dayhoff-170M-UR90-HL-8000microsoft/Dayhoff-170M-UR90-HL-8000 | Not documented | — | — | — | |||
| microsoft/Dayhoff-170M-GR-1000microsoft/Dayhoff-170M-GR-1000 | Not documented | — | — | — | |||
| microsoft/Dayhoff-170M-UR90-HL-4000microsoft/Dayhoff-170M-UR90-HL-4000 | Not documented | — | — | — | |||
| google/gemma-4-E2B-it-qat-w4a16-ctgoogle/gemma-4-E2B-it-qat-w4a16-ct | Not documented | — | — | — | |||
| deepseek-ai/DeepSeek-V3.1baseten/deepseek-ai/deepseek-v3.1 | Not documented | $0.5 | $1.5 | — | |||
| deepseek-ai/DeepSeek-V3-0324baseten/deepseek-ai/deepseek-v3-0324 | Not documented | $0.77 | $0.77 | — | |||
| OpenAI: GPT Audio Miniopenai/gpt-audio-mini | 128K | $0.6 | $2.4 | — | |||
| us-gov-east-1/meta.llama3-8b-instruct-v1:0bedrock/us-gov-east-1/meta.llama3-8b-instruct-v1:0 | 8K | $0.3 | $2.65 | — | |||
| us.anthropic.claude-opus-4-7bedrock_converse/us.anthropic.claude-opus-4-7 | 1M | $5.5 | $27.5 | — |