No provider description is available for this model yet.
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
OpenAI o4-mini-high is the same model as [o4-mini](/openai/o4-mini) with reasoning_effort set to high. OpenAI o4-mini is a compact reasoning model in the o-series, optimized for fast, cost-efficient performance while retaining...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Qwen3.5-9B is a multimodal foundation model from the Qwen3.5 family, designed to deliver strong reasoning, coding, and visual understanding in an efficient 9B-parameter architecture. It uses a unified vision-language design...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
GPT-5.6 Sol Pro is the same underlying model as [GPT-5.6 Sol](https://openrouter.ai/openai/gpt-5.6-sol), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. Learn more in OpenAI's docs: https://developers.openai.com/api/docs/guides/reasoning#reasoning-mode
A balanced model in the Ministral 3 family, Ministral 3 8B is a powerful, efficient tiny language model with vision capabilities.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Qwen Plus 0728, based on the Qwen3 foundation model, is a 1 million context hybrid reasoning model with a balanced performance, speed, and cost combination.
GLM-4.5V is a vision-language foundation model for multimodal agent applications. Built on a Mixture-of-Experts (MoE) architecture with 106B parameters and 12B activated parameters, it achieves state-of-the-art results in video understanding,...
No provider description is available for this model yet.
The Meta Llama 3.3 multilingual large language model (LLM) is a pretrained and instruction tuned generative model in 70B (text in/text out). The Llama 3.3 instruction tuned text only model...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Inkling is an open-weight multimodal mixture-of-experts model from Thinking Machines Lab, with 41B active parameters out of 975B total. It is designed for general-purpose reasoning, coding, agentic and tool-use systems,...
Qwen2.5-VL is proficient in recognizing common objects such as flowers, birds, fish, and insects. It is also highly capable of analyzing texts, charts, icons, graphics, and layouts within images.
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| us-gov-west-1/amazon.titan-text-premier-v1:0bedrock/us-gov-west-1/amazon.titan-text-premier-v1:0 | 42K | $0.5 | $1.5 | — | |||
| OpenAI: o4 Mini Highopenai/o4-mini-high | 200K | $1.1 | $4.4 | — | |||
| global.openai.gpt-5.6-solbedrock_converse/global.openai.gpt-5.6-sol | 1M | $5 | $30 | — | |||
| @cf/nvidia/nemotron-3-120b-a12bcloudflare/@cf/nvidia/nemotron-3-120b-a12b | 256K | $0.5 | $1.5 | — | |||
| us.openai.gpt-5.6-terrabedrock_converse/us.openai.gpt-5.6-terra | 1M | $2.2 | $13.2 | — | |||
| deepseek-v4-flashdashscope/deepseek-v4-flash | 1M | $0.2 | $0.4 | — | |||
| deepseek-v4-flash-0731dashscope/deepseek-v4-flash-0731 | 1M | $0.2 | $0.4 | — | |||
| global.openai.gpt-5.6-terrabedrock_converse/global.openai.gpt-5.6-terra | 1M | $2 | $12 | — | |||
| us-gov-west-1/anthropic.claude-3-7-sonnet-20250219-v1:0bedrock/us-gov-west-1/anthropic.claude-3-7-sonnet-20250219-v1:0 | 200K | $3.6 | $18 | — | |||
| deepseek-v4-prodashscope/deepseek-v4-pro | 1M | $2.4 | $4.8 | — | |||
| glm-5.1dashscope/glm-5.1 | 202.745K | $1.4 | $4.4 | — | |||
| us.openai.gpt-5.6-lunabedrock_converse/us.openai.gpt-5.6-luna | 1M | $0.22 | $1.32 | — | |||
| global.openai.gpt-5.6-lunabedrock_converse/global.openai.gpt-5.6-luna | 1M | $0.2 | $1.2 | — | |||
| global.anthropic.claude-fable-5bedrock_converse/global.anthropic.claude-fable-5 | 1M | $10 | $50 | — | |||
| @cf/aisingapore/gemma-sea-lion-v4-27b-itcloudflare/@cf/aisingapore/gemma-sea-lion-v4-27b-it | 128K | $0.351 | $0.555 | — | |||
| us.anthropic.claude-fable-5bedrock_converse/us.anthropic.claude-fable-5 | 1M | $11 | $55 | — | |||
| eu.anthropic.claude-fable-5bedrock_converse/eu.anthropic.claude-fable-5 | 1M | $11 | $55 | — | |||
| anthropic.claude-opus-5bedrock_converse/anthropic.claude-opus-5 | 1M | $5 | $25 | — | |||
| global.anthropic.claude-opus-5bedrock_converse/global.anthropic.claude-opus-5 | 1M | $5 | $25 | — | |||
| us.anthropic.claude-opus-5bedrock_converse/us.anthropic.claude-opus-5 | 1M | $5.5 | $27.5 | — | |||
| glm-5.2dashscope/glm-5.2 | 1.04858M | $1.4 | $4.4 | — | |||
| kimi-k2.7-codedashscope/kimi-k2.7-code | 229.376K | $0.95 | $4 | — | |||
| us-gov-west-1/anthropic.claude-3-5-sonnet-20240620-v1:0bedrock/us-gov-west-1/anthropic.claude-3-5-sonnet-20240620-v1:0 | 200K | $3.6 | $18 | — | |||
| Qwen: Qwen3.5-9B (batch)qwen/qwen3.5-9b:batch | 262.144K | $0.17 | $0.25 | — | |||
| qwen3.8-maxdashscope/qwen3.8-max | 991.808K | $2 | $6 | — | |||
| nvidia/NVIDIA-Nemotron-3.5-Lightningdeepinfra/nvidia/nvidia-nemotron-3.5-lightning | 262.144K | $0.05 | $0.2 | — | |||
| gemini-3.7-flashvertex_ai/gemini-3.7-flash | 1.04858M | $0.75 | $3.75 | — | |||
| gemini-3.7-flashgemini/gemini-3.7-flash | 1.04858M | $0.75 | $3.75 | — | |||
| us-gov-west-1/anthropic.claude-3-haiku-20240307-v1:0bedrock/us-gov-west-1/anthropic.claude-3-haiku-20240307-v1:0 | 200K | $0.3 | $1.5 | — | |||
| eu.anthropic.claude-opus-5bedrock_converse/eu.anthropic.claude-opus-5 | 1M | $5.5 | $27.5 | — | |||
| au.anthropic.claude-opus-5bedrock_converse/au.anthropic.claude-opus-5 | 1M | $5.5 | $27.5 | — | |||
| OpenAI: GPT-5.6 Sol Pro (batch)openai/gpt-5.6-sol-pro:batch | 1.05M | $1 | $5 | — | |||
| Mistral: Ministral 3 8B 2512mistralai/ministral-8b-2512 | 262.144K | $0.15 | $0.15 | — | |||
| @cf/qwen/qwen3-30b-a3b-fp8cloudflare/@cf/qwen/qwen3-30b-a3b-fp8 | 32.768K | $0.051 | $0.335 | — | |||
| jp.anthropic.claude-opus-5bedrock_converse/jp.anthropic.claude-opus-5 | 1M | $5.5 | $27.5 | — | |||
| anthropic.claude-opus-4-8bedrock_converse/anthropic.claude-opus-4-8 | 1M | $5 | $25 | — | |||
| gemini-3.7-flashvertex_ai-language-models/gemini-3.7-flash | 1.04858M | $0.75 | $3.75 | — | |||
| claude-3-opus-20240229anthropic/claude-3-opus-20240229 | 200K | $15 | $75 | — | |||
| qwen/qwen3.6-27bgroq/qwen/qwen3.6-27b | 131.072K | $0.6 | $3 | — | |||
| us-gov-west-1/anthropic.claude-sonnet-4-5-20250929-v1:0bedrock/us-gov-west-1/anthropic.claude-sonnet-4-5-20250929-v1:0 | 200K | $3.6 | $18 | — | |||
| nvidia/nemotron-3.5-lightningopenrouter/nvidia/nemotron-3.5-lightning | 262.144K | $0.05 | $0.2 | — | |||
| Qwen: Qwen Plus 0728qwen/qwen-plus-2025-07-28 | 1M | $0.26 | $0.78 | — | |||
| Z.ai: GLM 4.5Vz-ai/glm-4.5v | 65.536K | $0.6 | $1.8 | — | |||
| global.anthropic.claude-opus-4-8bedrock_converse/global.anthropic.claude-opus-4-8 | 1M | $5 | $25 | — | |||
| Meta: Llama 3.3 70B Instruct (free)meta-llama/llama-3.3-70b-instruct:free | 65.536K | Free | Free | — | |||
| us.anthropic.claude-opus-4-8bedrock_converse/us.anthropic.claude-opus-4-8 | 1M | $5.5 | $27.5 | — | |||
| eu.anthropic.claude-opus-4-8bedrock_converse/eu.anthropic.claude-opus-4-8 | 1M | $5.5 | $27.5 | — | |||
| au.anthropic.claude-opus-4-8bedrock_converse/au.anthropic.claude-opus-4-8 | 1M | $5.5 | $27.5 | — | |||
| Thinking Machines: Inkling (free)thinkingmachines/inkling:free | 1.04858M | Free | Free | — | |||
| Qwen: Qwen2.5 VL 72B Instructqwen/qwen2.5-vl-72b-instruct | 128K | $0.8 | $1 | — |