GPT-5.4 is OpenAI’s latest frontier model, unifying the Codex and GPT lines into a single system. It features a 1M+ token context window (922K input, 128K output) with support for...
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
No provider description is available for this model yet.
No provider description is available for this model yet.
GPT-5.6 Luna Pro is the same underlying model as [GPT-5.6 Luna](https://openrouter.ai/openai/gpt-5.6-luna), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. Learn more in OpenAI's docs: https://developers.openai.com/api/docs/guides/reasoning#reasoning-mode
No provider description is available for this model yet.
No provider description is available for this model yet.
GPT-4.1 Mini is a mid-sized model delivering performance competitive with GPT-4o at substantially lower latency and cost. It retains a 1 million token context window and scores 45.1% on hard...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Ministral 8B is an 8B parameter model featuring a unique interleaved sliding-window attention pattern for faster, memory-efficient inference. Designed for edge use cases, it supports up to 128k context length...
No provider description is available for this model yet.
Fugu Max is the cost-performance model in Sakana AI's Fugu family. Rather than a single monolithic model, Fugu is a learned multi-agent orchestration system: a language model trained to route...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
GLM-4.5V is a vision-language foundation model for multimodal agent applications. Built on a Mixture-of-Experts (MoE) architecture with 106B parameters and 12B activated parameters, it achieves state-of-the-art results in video understanding,...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
GLM-5.3 is a large-scale reasoning model from Z.ai, built for complex software engineering and long-horizon agent tasks. It supports text input and output with a 1M-token context window, and improves...
GPT-4o mini is OpenAI's newest model after [GPT-4 Omni](/models/openai/gpt-4o), supporting both text and image inputs with text outputs. As their most advanced small model, it is many multiples more affordable...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
GLM-5 is Z.ai’s flagship open-source foundation model engineered for complex systems design and long-horizon agent workflows. Built for expert developers, it delivers production-grade performance on large-scale programming tasks, rivaling leading...
Claude Sonnet 4.5 is Anthropic’s most advanced Sonnet model to date, optimized for real-world agents and coding workflows. It delivers state-of-the-art performance on coding benchmarks such as SWE-bench Verified, with...
No provider description is available for this model yet.
GPT-5.6 Sol is the flagship model in OpenAI's GPT-5.6 series. It is suited for complex reasoning, coding, and agentic workflows, and is particularly strong at command-line and multi-step coding tasks...
GPT-5 Pro is OpenAI’s most advanced model, offering major improvements in reasoning, code quality, and user experience. It is optimized for complex tasks that require step-by-step reasoning, instruction following, and...
The latest GPT-4 Turbo model with vision capabilities. Vision requests can now use JSON mode and function calling. Training data: up to December 2023.
Sonnet 4.6 is Anthropic's most capable Sonnet-class model yet, with frontier performance across coding, agents, and professional work. It excels at iterative development, complex codebase navigation, end-to-end project management with...
GPT-5.6 Terra Pro is the same underlying model as [GPT-5.6 Terra](https://openrouter.ai/openai/gpt-5.6-terra), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. Learn more in OpenAI's docs: https://developers.openai.com/api/docs/guides/reasoning#reasoning-mode
Mistral Large 3 2512 is Mistral’s most capable model to date, featuring a sparse mixture-of-experts architecture with 41B active parameters (675B total), and released under the Apache 2.0 license.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| OpenAI: GPT-5.4 (batch)openai/gpt-5.4:batch | 1.05M | $1.25 | $7.5 | — | |||
| upstage/Solar-Open-100Bupstage/Solar-Open-100B | Not documented | — | — | — | |||
| Qwen/Qwen3.8-2.4T-A95BQwen/Qwen3.8-2.4T-A95B | Not documented | — | — | — | |||
| OpenAI: GPT-5.6 Luna Pro (batch)openai/gpt-5.6-luna-pro:batch | 1.05M | $0.1 | $0.6 | — | |||
| accounts/fireworks/models/codegemma-7bfireworks_ai/accounts/fireworks/models/codegemma-7b | 8.192K | $0.2 | $0.2 | — | |||
| accounts/fireworks/models/codegemma-2bfireworks_ai/accounts/fireworks/models/codegemma-2b | 8.192K | $0.1 | $0.1 | — | |||
| OpenAI: GPT-4.1 Mini (batch)openai/gpt-4.1-mini:batch | 1.04758M | $0.2 | $0.8 | — | |||
| accounts/fireworks/models/llama4-maverick-instruct-basicfireworks_ai/accounts/fireworks/models/llama4-maverick-instruct-basic | 131.072K | $0.22 | $0.88 | — | |||
| nvidia/NVIDIA-Nemotron-Labs-Teacher-STEMnvidia/NVIDIA-Nemotron-Labs-Teacher-STEM | Not documented | — | — | — | |||
| nvidia/NVIDIA-Nemotron-Labs-Teacher-Instruction-Followingnvidia/NVIDIA-Nemotron-Labs-Teacher-Instruction-Following | Not documented | — | — | — | |||
| accounts/fireworks/models/code-qwen-1p5-7bfireworks_ai/accounts/fireworks/models/code-qwen-1p5-7b | 65.536K | $0.2 | $0.2 | — | |||
| nvidia/NVIDIA-Nemotron-Labs-Teacher-Competition-Codingnvidia/NVIDIA-Nemotron-Labs-Teacher-Competition-Coding | Not documented | — | — | — | |||
| openai/gpt-oss-20breplicate/openai/gpt-oss-20b | Not documented | $0.09 | $0.36 | — | |||
| nvidia/NVIDIA-Nemotron-Labs-Teacher-Chatnvidia/NVIDIA-Nemotron-Labs-Teacher-Chat | Not documented | — | — | — | |||
| Mistral: Ministral 8Bmistralai/ministral-8b | 128K | $0.11 | $0.11 | — | |||
| accounts/fireworks/models/code-llama-7b-pythonfireworks_ai/accounts/fireworks/models/code-llama-7b-python | 16.384K | $0.2 | $0.2 | — | |||
| Sakana: Fugu Maxsakana/fugu-max | 1M | $2 | $6 | — | |||
| openai/gpt-5.6-solopenrouter/openai/gpt-5.6-sol | 1.05M | $2 | $10 | — | |||
| accounts/fireworks/models/llama-v3p2-90b-vision-instructfireworks_ai/accounts/fireworks/models/llama-v3p2-90b-vision-instruct | 16.384K | $0.9 | $0.9 | — | |||
| ibm-granite/granite-4.2-8bwandb/ibm-granite/granite-4.2-8b | Not documented | $0.1 | $0.15 | — | |||
| Z.ai: GLM 4.5Vz-ai/glm-4.5v | 65.536K | $0.6 | $1.8 | — | |||
| stabilityai/japanese-stablelm-instruct-alpha-7bstabilityai/japanese-stablelm-instruct-alpha-7b | Not documented | — | — | — | |||
| apac.amazon.nova-lite-v1:0bedrock_converse/apac.amazon.nova-lite-v1:0 | 300K | $0.063 | $0.252 | — | |||
| accounts/fireworks/models/code-llama-7b-instructfireworks_ai/accounts/fireworks/models/code-llama-7b-instruct | 16.384K | $0.2 | $0.2 | — | |||
| accounts/fireworks/models/code-llama-7bfireworks_ai/accounts/fireworks/models/code-llama-7b | 16.384K | $0.2 | $0.2 | — | |||
| Z.ai: GLM 5.3 (batch)z-ai/glm-5.3:batch | 1.04858M | $0.7 | $2.2 | — | |||
| OpenAI: GPT-4o-mini (batch)openai/gpt-4o-mini:batch | 128K | $0.075 | $0.3 | — | |||
| accounts/fireworks/models/llama-v3p2-3b-instructfireworks_ai/accounts/fireworks/models/llama-v3p2-3b-instruct | 16.384K | $0.1 | $0.1 | — | |||
| nvidia/NVIDIA-Nemotron-Labs-Teacher-General-Reasoningnvidia/NVIDIA-Nemotron-Labs-Teacher-General-Reasoning | Not documented | — | — | — | |||
| accounts/fireworks/models/code-llama-70b-pythonfireworks_ai/accounts/fireworks/models/code-llama-70b-python | 4.096K | $0.9 | $0.9 | — | |||
| Z.ai: GLM 5z-ai/glm-5 | 198K | $0.6 | $1.92 | — | |||
| Anthropic: Claude Sonnet 4.5 (batch)anthropic/claude-sonnet-4.5:batch | 1M | $1.5 | $7.5 | — | |||
| accounts/fireworks/models/code-llama-70b-instructfireworks_ai/accounts/fireworks/models/code-llama-70b-instruct | 4.096K | $0.9 | $0.9 | — | |||
| OpenAI: GPT-5.6 Sol (batch)openai/gpt-5.6-sol:batch | 1.05M | $1 | $5 | — | |||
| OpenAI: GPT-5 Pro (batch)openai/gpt-5-pro:batch | 400K | $7.5 | $60 | — | |||
| OpenAI: GPT-4 Turbo (batch)openai/gpt-4-turbo:batch | 128K | $5 | $15 | — | |||
| Anthropic: Claude Sonnet 4.6 (batch)anthropic/claude-sonnet-4.6:batch | 1M | $1.5 | $7.5 | — | |||
| OpenAI: GPT-5.6 Terra Pro (batch)openai/gpt-5.6-terra-pro:batch | 1.05M | $1 | $6 | — | |||
| Mistral: Mistral Large 3 2512mistralai/mistral-large-2512 | 262.144K | $0.5 | $1.5 | — | |||
| databricks-glm-5-2databricks/databricks-glm-5-2 | 1M | $1.4 | $4.4 | — | |||
| databricks-kimi-k3databricks/databricks-kimi-k3 | 1M | $3 | $15 | — | |||
| accounts/fireworks/models/deepseek-v4-pro-0813fireworks_ai/accounts/fireworks/models/deepseek-v4-pro-0813 | 1.04858M | $1.32 | $3.96 | — | |||
| accounts/fireworks/models/llama-v3p2-1b-instructfireworks_ai/accounts/fireworks/models/llama-v3p2-1b-instruct | 16.384K | $0.1 | $0.1 | — | |||
| ministral-14b-latestmistral/ministral-14b-latest | 262.144K | $0.2 | $0.2 | — | |||
| ministral-3b-2512mistral/ministral-3b-2512 | 131.072K | $0.1 | $0.1 | — | |||
| ministral-3b-latestmistral/ministral-3b-latest | 131.072K | $0.1 | $0.1 | — | |||
| mistral-medium-3mistral/mistral-medium-3 | 262.144K | $1.5 | $7.5 | — | |||
| voxtral-small-2507mistral/voxtral-small-2507 | 32.768K | $0.1 | $0.4 | — | |||
| zai-org/GLM-5.3-Flashtogether_ai/zai-org/glm-5.3-flash | 1.04858M | $0.15 | $0.5 | — | |||
| stabilityai/StableBeluga-13Bstabilityai/StableBeluga-13B | Not documented | — | — | — |