Opus 4.7 is the next generation of Anthropic's Opus family, built for long-running, asynchronous agents. Building on the coding and agentic strengths of Opus 4.6, it delivers stronger performance on...
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
ERNIE-4.5-VL-424B-A47B is a multimodal Mixture-of-Experts (MoE) model from Baidu’s ERNIE 4.5 series, featuring 424B total parameters with 47B active per token. It is trained jointly on text and image data...
Llama 4 Maverick 17B Instruct (128E) is a high-capacity multimodal language model from Meta, built on a mixture-of-experts (MoE) architecture with 128 experts and 17 billion active parameters per forward...
No provider description is available for this model yet.
GPT-5.6 Sol Pro is the same underlying model as [GPT-5.6 Sol](https://openrouter.ai/openai/gpt-5.6-sol), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. Learn more in OpenAI's docs: https://developers.openai.com/api/docs/guides/reasoning#reasoning-mode
No provider description is available for this model yet.
No provider description is available for this model yet.
Qwen3.6 27B is a dense 27-billion-parameter language model from the Qwen Team at Alibaba, released in April 2026. It features hybrid multimodal capabilities — accepting text, image, and video inputs...
No provider description is available for this model yet.
Qwen3-VL-8B-Instruct is a multimodal vision-language model from the Qwen3-VL series, built for high-fidelity understanding and reasoning across text, images, and video. It features improved multimodal fusion with Interleaved-MRoPE for long-horizon...
GPT-5.6 Luna Pro is the same underlying model as [GPT-5.6 Luna](https://openrouter.ai/openai/gpt-5.6-luna), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. Learn more in OpenAI's docs: https://developers.openai.com/api/docs/guides/reasoning#reasoning-mode
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
GPT-5 Chat is designed for advanced, natural, multimodal, and context-aware conversations for enterprise applications.
No provider description is available for this model yet.
No provider description is available for this model yet.
Qwen3-VL-8B-Thinking is the reasoning-optimized variant of the Qwen3-VL-8B multimodal model, designed for advanced visual and textual reasoning across complex scenes, documents, and temporal sequences. It integrates enhanced multimodal alignment and...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Gemini 3.5 Flash Lite is a high-efficiency model from Google with upgraded agentic capabilities. It is suited for subagents that execute focused tasks within complex, multi-agent workflows.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Claude Haiku 4.5 is Anthropic’s fastest and most efficient model, delivering near-frontier intelligence at a fraction of the cost and latency of larger Claude models. Matching Claude Sonnet 4’s performance...
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| Anthropic: Claude Opus 4.7anthropic/claude-opus-4.7 | 1M | $5 | $25 | — | |||
| global/gpt-4o-2024-11-20azure/global/gpt-4o-2024-11-20 | 128K | $2.5 | $10 | — | |||
| apac.anthropic.claude-3-5-sonnet-20240620-v1:0bedrock/apac.anthropic.claude-3-5-sonnet-20240620-v1:0 | 200K | $3 | $15 | — | |||
| nvidia/Nemotron-3-Content-Safetynvidia/Nemotron-3-Content-Safety | Not documented | — | — | — | |||
| global/gpt-5.1azure/global/gpt-5.1 | 272K | $1.25 | $10 | — | |||
| global/gpt-5.1-chatazure/global/gpt-5.1-chat | 128K | $1.25 | $10 | — | |||
| nvidia/Nemotron-Labs-Diffusion-VLM-8Bnvidia/Nemotron-Labs-Diffusion-VLM-8B | Not documented | — | — | — | |||
| us-east-1/moonshotai.kimi-k2.5bedrock/us-east-1/moonshotai.kimi-k2.5 | 262.144K | $0.6 | $3 | — | |||
| Baidu: ERNIE 4.5 VL 424B A47B baidu/ernie-4.5-vl-424b-a47b | 123K | $0.42 | $1.25 | — | |||
| Meta: Llama 4 Maverickmeta-llama/llama-4-maverick | 128K | $0.2 | $0.696 | — | |||
| databricks-claude-opus-4-8databricks/databricks-claude-opus-4-8 | 1M | $5 | $25 | — | |||
| OpenAI: GPT-5.6 Sol Pro (batch)openai/gpt-5.6-sol-pro:batch | 1.05M | $1 | $5 | — | |||
| gpt-4-turbo-2024-04-09azure/gpt-4-turbo-2024-04-09 | 128K | $10 | $30 | — | |||
| databricks-claude-opus-5databricks/databricks-claude-opus-5 | 1M | $5 | $25 | — | |||
| Qwen: Qwen3.6 27Bqwen/qwen3.6-27b | 262.144K | $0.3 | $2 | — | |||
| gpt-4.1azure/gpt-4.1 | 1.04758M | $2 | $8 | — | |||
| Qwen: Qwen3 VL 8B Instructqwen/qwen3-vl-8b-instruct | 131.072K | $0.117 | $0.455 | — | |||
| OpenAI: GPT-5.6 Luna Proopenai/gpt-5.6-luna-pro | 1.05M | $0.2 | $1.2 | — | |||
| gpt-4.1-miniazure/gpt-4.1-mini | 1.04758M | $0.4 | $1.6 | — | |||
| databricks-claude-sonnet-5databricks/databricks-claude-sonnet-5 | 1M | $3 | $15 | — | |||
| gpt-4.1-mini-2025-04-14azure/gpt-4.1-mini-2025-04-14 | 1.04758M | $0.4 | $1.6 | — | |||
| gpt-4.1-nanoazure/gpt-4.1-nano | 1.04758M | $0.1 | $0.4 | — | |||
| gpt-4.1-nano-2025-04-14azure/gpt-4.1-nano-2025-04-14 | 1.04758M | $0.1 | $0.4 | — | |||
| gpt-4.5-previewazure/gpt-4.5-preview | 128K | $75 | $150 | — | |||
| gpt-4oazure/gpt-4o | 128K | $2.5 | $10 | — | |||
| gpt-4o-2024-05-13azure/gpt-4o-2024-05-13 | 128K | $5 | $15 | — | |||
| MiniMaxAI/MiniMax-M3together_ai/minimaxai/minimax-m3 | 524.288K | $0.3 | $1.2 | — | |||
| gpt-4o-2024-08-06azure/gpt-4o-2024-08-06 | 128K | $2.5 | $10 | — | |||
| apac.amazon.nova-pro-v1:0bedrock_converse/apac.amazon.nova-pro-v1:0 | 300K | $0.84 | $3.36 | — | |||
| Qwen/Qwen3.5-9Btogether_ai/qwen/qwen3.5-9b | 262.144K | $0.17 | $0.25 | — | |||
| apac.amazon.nova-lite-v1:0bedrock_converse/apac.amazon.nova-lite-v1:0 | 300K | $0.063 | $0.252 | — | |||
| gpt-4o-miniazure/gpt-4o-mini | 128K | $0.165 | $0.66 | — | |||
| us-gov-east-1/xai.grok-4.6bedrock_mantle/us-gov-east-1/xai.grok-4.6 | 500K | $2.64 | $7.92 | — | |||
| gpt-4o-mini-2024-07-18azure/gpt-4o-mini-2024-07-18 | 128K | $0.165 | $0.66 | — | |||
| OpenAI: GPT-5 Chatopenai/gpt-5-chat | 128K | $1.25 | $10 | — | |||
| gpt-5.1-chat-2025-11-13azure/gpt-5.1-chat-2025-11-13 | 128K | $1.25 | $10 | — | |||
| us-gov-west-1/google.gemma-4-31bbedrock_mantle/us-gov-west-1/google.gemma-4-31b | 256K | $0.168 | $0.48 | — | |||
| Qwen: Qwen3 VL 8B Thinkingqwen/qwen3-vl-8b-thinking | 131.072K | $0.18 | $2.1 | — | |||
| us-gov-west-1/google.gemma-4-26b-a4bbedrock_mantle/us-gov-west-1/google.gemma-4-26b-a4b | 256K | $0.156 | $0.48 | — | |||
| gpt-5-chatazure/gpt-5-chat | 128K | $1.25 | $10 | — | |||
| gpt-5-chat-latestazure/gpt-5-chat-latest | 128K | $1.25 | $10 | — | |||
| google/gemma-4-31B-ittogether_ai/google/gemma-4-31b-it | 262.144K | $0.39 | $0.97 | — | |||
| google/medgemma-1.5-4b-itgoogle/medgemma-1.5-4b-it | Not documented | — | — | — | |||
| gpt-5-miniazure/gpt-5-mini | 272K | $0.25 | $2 | — | |||
| Google: Gemini 3.5 Flash Lite (batch)google/gemini-3.5-flash-lite:batch | 1.04858M | $0.15 | $1.25 | — | |||
| gpt-5-mini-2025-08-07azure/gpt-5-mini-2025-08-07 | 272K | $0.25 | $2 | — | |||
| us-gov-west-1/google.gemma-4-e2bbedrock_mantle/us-gov-west-1/google.gemma-4-e2b | 128K | $0.048 | $0.096 | — | |||
| moonshotai/Kimi-K2.7-Codetogether_ai/moonshotai/kimi-k2.7-code | 262.144K | $0.95 | $4 | — | |||
| gpt-5-nano-2025-08-07azure/gpt-5-nano-2025-08-07 | 272K | $0.05 | $0.4 | — | |||
| Anthropic: Claude Haiku 4.5anthropic/claude-haiku-4.5 | 200K | $1 | $5 | — |