No provider description is available for this model yet.
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
GPT-5.4 Pro is OpenAI's most advanced model, building on GPT-5.4's unified architecture with enhanced reasoning capabilities for complex, high-stakes tasks. It features a 1M+ token context window (922K input, 128K...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Qwen3.8 2.4T A95B is an open-weight sparse mixture-of-experts model from Qwen and the open-weight variant of [Qwen3.8 Max](/qwen/qwen3.8-max), with 95 billion active parameters out of 2.4 trillion total. It is...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
GPT-5.6 Terra Pro is the same underlying model as [GPT-5.6 Terra](https://openrouter.ai/openai/gpt-5.6-terra), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. Learn more in OpenAI's docs: https://developers.openai.com/api/docs/guides/reasoning#reasoning-mode
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
GPT-5.6 Sol Pro is the same underlying model as [GPT-5.6 Sol](https://openrouter.ai/openai/gpt-5.6-sol), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. Learn more in OpenAI's docs: https://developers.openai.com/api/docs/guides/reasoning#reasoning-mode
No provider description is available for this model yet.
No provider description is available for this model yet.
Sonnet 4.6 is Anthropic's most capable Sonnet-class model yet, with frontier performance across coding, agents, and professional work. It excels at iterative development, complex codebase navigation, end-to-end project management with...
No provider description is available for this model yet.
Gemini 2.5 Pro is Google’s state-of-the-art AI model designed for advanced reasoning, coding, mathematics, and scientific tasks. It employs “thinking” capabilities, enabling it to reason through responses with enhanced accuracy...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Gemini 3.5 Flash Lite is a high-efficiency model from Google with upgraded agentic capabilities. It is suited for subagents that execute focused tasks within complex, multi-agent workflows.
No provider description is available for this model yet.
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| apac.anthropic.claude-sonnet-4-20250514-v1:0bedrock_converse/apac.anthropic.claude-sonnet-4-20250514-v1:0 | 1M | $3 | $15 | — | |||
| claude-opus-4-6azure_ai/claude-opus-4-6 | 1M | $5 | $25 | — | |||
| claude-opus-4-7azure_ai/claude-opus-4-7 | 1M | $5 | $25 | — | |||
| claude-fable-5azure_ai/claude-fable-5 | 1M | $10 | $50 | — | |||
| claude-opus-5azure_ai/claude-opus-5 | 1M | $5 | $25 | — | |||
| OpenAI: GPT-5.4 Pro (batch)openai/gpt-5.4-pro:batch | 1.05M | $15 | $90 | — | |||
| global.anthropic.claude-fable-5bedrock_converse/global.anthropic.claude-fable-5 | 1M | $10 | $50 | — | |||
| claude-sonnet-5azure_ai/claude-sonnet-5 | 1M | $2 | $10 | — | |||
| claude-sonnet-4-6azure_ai/claude-sonnet-4-6 | 1M | $3 | $15 | — | |||
| Qwen: Qwen3.8 2.4T A95B (batch)qwen/qwen3.8-2.4t-a95b:batch | 1.01M | $2 | $6 | — | |||
| deepseek-v4-flash-0731fireworks_ai/deepseek-v4-flash-0731 | 1.04858M | $0.14 | $0.28 | — | |||
| gpt-5.5azure_ai/gpt-5.5 | 1.05M | $5 | $30 | — | |||
| ap-southeast-2/minimax.minimax-m2.5bedrock/ap-southeast-2/minimax.minimax-m2.5 | 1M | $0.309 | $1.236 | — | |||
| gemini-3.8-flashvertex_ai-language-models/gemini-3.8-flash | 1.04858M | $0.75 | $3.75 | — | |||
| global.openai.gpt-5.6-lunabedrock_converse/global.openai.gpt-5.6-luna | 1M | $0.2 | $1.2 | — | |||
| gpt-5.5-2026-04-23azure_ai/gpt-5.5-2026-04-23 | 1.05M | $5 | $30 | — | |||
| us-gov.anthropic.claude-opus-4-8bedrock_converse/us-gov.anthropic.claude-opus-4-8 | 1M | $6 | $30 | — | |||
| gpt-5.4azure_ai/gpt-5.4 | 1.05M | $2.5 | $15 | — | |||
| accounts/fireworks/models/glm-5p3-flashfireworks_ai/accounts/fireworks/models/glm-5p3-flash | 1.04858M | $0.15 | $0.5 | — | |||
| accounts/fireworks/models/inklingfireworks_ai/accounts/fireworks/models/inkling | 1.04858M | $1 | $4.05 | — | |||
| us.openai.gpt-5.6-lunabedrock_converse/us.openai.gpt-5.6-luna | 1M | $0.22 | $1.32 | — | |||
| us-gov-west-1/anthropic.claude-sonnet-5bedrock/us-gov-west-1/anthropic.claude-sonnet-5 | 1M | $2.4 | $12 | — | |||
| FW-Inklingazure_ai/fw-inkling | 1.04858M | $1 | $4.05 | — | |||
| gpt-5.4-2026-03-05azure_ai/gpt-5.4-2026-03-05 | 1.05M | $2.5 | $15 | — | |||
| us-gov-east-1/anthropic.claude-sonnet-5bedrock/us-gov-east-1/anthropic.claude-sonnet-5 | 1M | $2.4 | $12 | — | |||
| us-gov-east-1/anthropic.claude-opus-4-8bedrock/us-gov-east-1/anthropic.claude-opus-4-8 | 1M | $6 | $30 | — | |||
| OpenAI: GPT-5.6 Terra Pro (batch)openai/gpt-5.6-terra-pro:batch | 1.05M | $1 | $6 | — | |||
| deepseek-v4-prodashscope/deepseek-v4-pro | 1M | $2.4 | $4.8 | — | |||
| FW-GLM-5.2-Fastazure_ai/fw-glm-5.2-fast | 1.04858M | $2.1 | $6.6 | — | |||
| databricks-claude-opus-4-7databricks/databricks-claude-opus-4-7 | 1M | $5 | $25 | — | |||
| databricks-claude-opus-4-8databricks/databricks-claude-opus-4-8 | 1M | $5 | $25 | — | |||
| OpenAI: GPT-5.6 Sol Pro (batch)openai/gpt-5.6-sol-pro:batch | 1.05M | $1 | $5 | — | |||
| databricks-claude-opus-5databricks/databricks-claude-opus-5 | 1M | $5 | $25 | — | |||
| gpt-4.1azure/gpt-4.1 | 1.04758M | $2 | $8 | — | |||
| Anthropic: Claude Sonnet 4.6anthropic/claude-sonnet-4.6 | 1M | $3 | $15 | — | |||
| gpt-4.1-2025-04-14azure/gpt-4.1-2025-04-14 | 1.04758M | $2 | $8 | — | |||
| Google: Gemini 2.5 Pro Preview 06-05google/gemini-2.5-pro-preview | 1.04858M | $1.25 | $10 | — | |||
| databricks-claude-sonnet-5databricks/databricks-claude-sonnet-5 | 1M | $3 | $15 | — | |||
| gpt-4.1-mini-2025-04-14azure/gpt-4.1-mini-2025-04-14 | 1.04758M | $0.4 | $1.6 | — | |||
| gpt-4.1-nanoazure/gpt-4.1-nano | 1.04758M | $0.1 | $0.4 | — | |||
| gpt-4.1-nano-2025-04-14azure/gpt-4.1-nano-2025-04-14 | 1.04758M | $0.1 | $0.4 | — | |||
| Qwen/Qwen3.6-Plustogether_ai/qwen/qwen3.6-plus | 1M | $0.5 | $3 | — | |||
| global.openai.gpt-5.6-terrabedrock_converse/global.openai.gpt-5.6-terra | 1M | $2 | $12 | — | |||
| Qwen/Qwen3.7-Plustogether_ai/qwen/qwen3.7-plus | 1M | $0.32 | $1.28 | — | |||
| deepseek-v4-flash-0731dashscope/deepseek-v4-flash-0731 | 1M | $0.2 | $0.4 | — | |||
| deepseek-ai/DeepSeek-V4-Flash-0731together_ai/deepseek-ai/deepseek-v4-flash-0731 | 1.04858M | $0.14 | $0.28 | — | |||
| deepseek-ai/DeepSeek-V4-Pro-0813together_ai/deepseek-ai/deepseek-v4-pro-0813 | 1.04858M | $1.32 | $3.96 | — | |||
| meta-llama/Llama-Guard-4-12Btogether_ai/meta-llama/llama-guard-4-12b | 1.04858M | $0.2 | $0.2 | — | |||
| Google: Gemini 3.5 Flash Lite (batch)google/gemini-3.5-flash-lite:batch | 1.04858M | $0.15 | $1.25 | — | |||
| deepseek-v4-flashdashscope/deepseek-v4-flash | 1M | $0.2 | $0.4 | — |