Amazon Nova Premier is the most capable of Amazon’s multimodal models for complex reasoning tasks and for use as the best teacher for distilling custom models.
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
Olmo 3 32B Think is a large-scale, 32-billion-parameter model purpose-built for deep reasoning, complex logic chains and advanced instruction-following scenarios. Its capacity enables strong performance on demanding evaluation tasks and...
No provider description is available for this model yet.
No provider description is available for this model yet.
The smallest model in the Ministral 3 family, Ministral 3 3B is a powerful, efficient tiny language model with vision capabilities.
Transform your natural language requests into structured OpenRouter API request objects. Describe what you want to accomplish with AI models, and Body Builder will construct the appropriate API calls. Example:...
GLM-4.6V is a large multimodal model designed for high-fidelity visual understanding and long-context reasoning across images, documents, and mixed media. It supports up to 128K tokens, processes complex page layouts...
The relace-search model uses 4-12 `view_file` and `grep` tools in parallel to explore a codebase and return relevant files to the user request. In contrast to RAG, relace-search performs agentic...
No provider description is available for this model yet.
No provider description is available for this model yet.
Cogito v2.1 671B MoE represents one of the strongest open models globally, matching performance of frontier closed and open models. This model is trained using self play with reinforcement learning...
Nova 2 Lite is a fast, cost-effective reasoning model for everyday workloads that can process text, images, and videos to generate text. Nova 2 Lite demonstrates standout capabilities in processing...
No provider description is available for this model yet.
GPT-5.2 Chat (AKA Instant) is the fast, lightweight member of the 5.2 family, optimized for low-latency chat while retaining strong general intelligence. It uses adaptive reasoning to selectively “think” on...
GLM-4.7 is Z.ai’s latest flagship model, featuring upgrades in two key areas: enhanced programming capabilities and more stable multi-step reasoning/execution. It demonstrates significant improvements in executing complex agent tasks while...
No provider description is available for this model yet.
MiniMax-M2.1 is a lightweight, state-of-the-art large language model optimized for coding, agentic workflows, and modern application development. With only 10 billion activated parameters, it delivers a major jump in real-world...
The gpt-audio model is OpenAI's first generally available audio model. The new snapshot features an upgraded decoder for more natural sounding voices and maintains better voice consistency. Audio is priced...
No provider description is available for this model yet.
MiniMax M2-her is a dialogue-first large language model built for immersive roleplay, character-driven chat, and expressive multi-turn conversations. Designed to stay consistent in tone and personality, it supports rich message...
Solar Pro 3 is Upstage's powerful Mixture-of-Experts (MoE) language model. With 102B total parameters and 12B active parameters per forward pass, it delivers exceptional performance while maintaining computational efficiency. Optimized...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
The Qwen3.5 native vision-language series Plus models are built on a hybrid architecture that integrates linear attention mechanisms with sparse mixture-of-experts models, achieving higher inference efficiency. In a variety of...
Aion-2.0 is a variant of DeepSeek V3.2 optimized for immersive roleplaying and storytelling. It is particularly strong at introducing tension, crises, and conflict into stories, making narratives feel more engaging....
No provider description is available for this model yet.
The Qwen3.5 122B-A10B native vision-language model is built on a hybrid architecture that integrates a linear attention mechanism with a sparse mixture-of-experts model, achieving higher inference efficiency. In terms of...
The Qwen3.5 27B native vision-language Dense model incorporates a linear attention mechanism, delivering fast response times while balancing inference speed and performance. Its overall capabilities are comparable to those of...
No provider description is available for this model yet.
The Qwen3.5 Series 35B-A3B is a native vision-language model designed with a hybrid architecture that integrates linear attention mechanisms and a sparse mixture-of-experts model, achieving higher inference efficiency. Its overall...
Qwen3.5-9B is a multimodal foundation model from the Qwen3.5 family, designed to deliver strong reasoning, coding, and visual understanding in an efficient 9B-parameter architecture. It uses a unified vision-language design...
No provider description is available for this model yet.
NVIDIA Nemotron 3 Super is a 120B-parameter open hybrid MoE model, activating just 12B parameters for maximum compute efficiency and accuracy in complex multi-agent applications. Built on a hybrid Mamba-Transformer...
Reka Edge is an extremely efficient 7B multimodal vision-language model that accepts image/video+text inputs and generates text outputs. This model is optimized specifically to deliver industry-leading performance in image understanding,...
No provider description is available for this model yet.
No provider description is available for this model yet.
KAT-Coder-Pro V2 is the latest high-performance model in KwaiKAT’s KAT-Coder series, designed for complex enterprise-grade software engineering and SaaS integration. It builds on the agentic coding strengths of earlier versions,...
This model always redirects to the latest model in the Claude Opus family.
No provider description is available for this model yet.
[GPT-5.4](https://openrouter.ai/openai/gpt-5.4) Image 2 combines OpenAI's GPT-5.4 model with state-of-the-art image generation capabilities from GPT Image 2. It enables rich multimodal workflows, allowing users to seamlessly move between reasoning, coding, and...
Mistral Small 4 is the next major release in the Mistral Small family, unifying the capabilities of several flagship Mistral models into a single system. It combines strong reasoning from...
No provider description is available for this model yet.
The Pareto Router maintains a tiered shortlist of strong coding models, ranked by [Artificial Analysis](https://artificialanalysis.ai/) coding percentiles. Set min_coding_score between 0 and 1 on the [pareto-router plugin](https://openrouter.ai/docs/guides/routing/routers/pareto-router#the-min_coding_score-parameter) to control how...
Ling-2.6-flash is an instant (instruct) model from inclusionAI with 104B total parameters and 7.4B active parameters, designed for real-world agents that require fast responses, strong execution, and high token efficiency....
No provider description is available for this model yet.
Qwen3.6 27B is a dense 27-billion-parameter language model from the Qwen Team at Alibaba, released in April 2026. It features hybrid multimodal capabilities — accepting text, image, and video inputs...
Qwen3.6-Max-Preview is a proprietary frontier model from Alibaba Cloud built on a sparse mixture-of-experts architecture with approximately 1 trillion total parameters. It is optimized for agentic coding, tool use, and...
No provider description is available for this model yet.
No provider description is available for this model yet.
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| Amazon: Nova Premier 1.0amazon/nova-premier-v1 | 1M | $2.5 | $12.5 | — | |||
| AllenAI: Olmo 3 32B Thinkallenai/olmo-3-32b-think | 65.536K | $0.15 | $0.5 | — | |||
| minimax/minimax-m2-heropenrouter/minimax/minimax-m2-her | 65.536K | $0.3 | $1.2 | — | |||
| meta-llama/llama-4-maverick-17b-128e-instruct-fp8watsonx/meta-llama/llama-4-maverick-17b-128e-instruct-fp8 | 131.072K | $0.371 | $1.484 | — | |||
| Mistral: Ministral 3 3B 2512mistralai/ministral-3b-2512 | 131.072K | $0.1 | $0.1 | — | |||
| Body Builder (beta)openrouter/bodybuilder | 128K | — | — | — | |||
| Z.ai: GLM 4.6Vz-ai/glm-4.6v | 131.072K | $0.3 | $0.9 | — | |||
| Relace: Relace Searchrelace/relace-search | 256K | $1 | $3 | — | |||
| qwen/qwen3-coder-nextopenrouter/qwen/qwen3-coder-next | 262.144K | $0.12 | $0.8 | — | |||
| bigscience/mt0-xxlwatsonx/bigscience/mt0-xxl | 4.096K | $1.908 | $1.908 | — | |||
| Deep Cogito: Cogito v2.1 671Bdeepcogito/cogito-v2.1-671b | 128K | $1.25 | $1.25 | — | |||
| Amazon: Nova 2 Liteamazon/nova-2-lite-v1 | 1M | $0.3 | $2.5 | — | |||
| qwen/qwen3-max-thinkingopenrouter/qwen/qwen3-max-thinking | 262.144K | $0.78 | $3.9 | — | |||
| OpenAI: GPT-5.2 Chatopenai/gpt-5.2-chat | 128K | $1.75 | $14 | — | |||
| Z.ai: GLM 4.7z-ai/glm-4.7 | 202.752K | $0.4 | $1.75 | — | |||
| google/gemini-3.1-pro-preview-customtoolsopenrouter/google/gemini-3.1-pro-preview-customtools | 1.04858M | $2 | $12 | — | |||
| MiniMax: MiniMax M2.1minimax/minimax-m2.1 | 204.8K | $0.3 | $1.2 | — | |||
| OpenAI: GPT Audioopenai/gpt-audio | 128K | $2.5 | $10 | — | |||
| google/gemini-3.1-flash-image-previewopenrouter/google/gemini-3.1-flash-image-preview | 65.536K | $0.5 | $3 | — | |||
| MiniMax: MiniMax M2-herminimax/minimax-m2-her | 65.536K | $0.3 | $1.2 | — | |||
| Upstage: Solar Pro 3upstage/solar-pro-3 | 131.072K | $0.15 | $0.6 | — | |||
| openai/gpt-5.4-proopenrouter/openai/gpt-5.4-pro | 1.05M | $30 | $180 | — | |||
| zai-org/GLM-5.3-Flashnebius/zai-org/glm-5.3-flash | 1.024M | $0.15 | $0.5 | — | |||
| zai-org/GLM-5.2nebius/zai-org/glm-5.2 | 1.04858M | $1.4 | $4.4 | — | |||
| Qwen: Qwen3.5 Plus 2026-02-15qwen/qwen3.5-plus-02-15 | 1M | $0.26 | $1.56 | — | |||
| AionLabs: Aion-2.0aion-labs/aion-2.0 | 131.072K | $0.8 | $1.6 | — | |||
| qwen/qwen3.5-9bopenrouter/qwen/qwen3.5-9b | 262.144K | $0.1 | $0.15 | — | |||
| Qwen: Qwen3.5-122B-A10Bqwen/qwen3.5-122b-a10b | 262.144K | $0.26 | $2.08 | — | |||
| Qwen: Qwen3.5-27Bqwen/qwen3.5-27b | 262.144K | $0.195 | $1.56 | — | |||
| nvidia/nemotron-3-super-120b-a12b:freeopenrouter/nvidia/nemotron-3-super-120b-a12b:free | 262.144K | Free | Free | — | |||
| Qwen: Qwen3.5-35B-A3Bqwen/qwen3.5-35b-a3b | 256K | $0.312 | $1.25 | — | |||
| Qwen: Qwen3.5-9Bqwen/qwen3.5-9b | 262.144K | $0.1 | $0.15 | — | |||
| nvidia/nemotron-3-super-120b-a12bopenrouter/nvidia/nemotron-3-super-120b-a12b | 1M | $0.085 | $0.4 | — | |||
| NVIDIA: Nemotron 3 Super (free)nvidia/nemotron-3-super-120b-a12b:free | 262.144K | Free | Free | — | |||
| Reka Edgerekaai/reka-edge | 16.384K | $0.1 | $0.1 | — | |||
| z-ai/glm-5-turboopenrouter/z-ai/glm-5-turbo | 202.752K | $1.2 | $4 | — | |||
| zai-org/GLM-5.1nebius/zai-org/glm-5.1 | 202.752K | $1.4 | $4.4 | — | |||
| Kwaipilot: KAT-Coder-Pro V2kwaipilot/kat-coder-pro-v2 | 262.144K | $0.3 | $1.2 | — | |||
| Anthropic: Claude Opus Latest~anthropic/claude-opus-latest | 1M | $5 | $25 | — | |||
| mistralai/mistral-small-2603openrouter/mistralai/mistral-small-2603 | 262.144K | $0.15 | $0.6 | — | |||
| OpenAI: GPT-5.4 Image 2openai/gpt-5.4-image-2 | 272K | $8 | $15 | — | |||
| Mistral: Mistral Small 4mistralai/mistral-small-2603 | 262.144K | $0.15 | $0.6 | — | |||
| minimax/minimax-m2.7:freeopenrouter/minimax/minimax-m2.7:free | 196.608K | Free | Free | — | |||
| Pareto Code Routeropenrouter/pareto-code | 2M | — | — | — | |||
| inclusionAI: Ling-2.6-flashinclusionai/ling-2.6-flash | 262.144K | $0.01 | $0.03 | — | |||
| minimax/minimax-m2.7openrouter/minimax/minimax-m2.7 | 204.8K | $0.3 | $1.2 | — | |||
| Qwen: Qwen3.6 27Bqwen/qwen3.6-27b | 262.144K | $0.3 | $2 | — | |||
| Qwen: Qwen3.6 Max Previewqwen/qwen3.6-max-preview | 262.144K | $1.027 | $6.162 | — | |||
| z-ai/glm-5v-turboopenrouter/z-ai/glm-5v-turbo | 202.752K | $1.2 | $4 | — | |||
| Qwen/Qwen3.5-397B-A17Bnebius/qwen/qwen3.5-397b-a17b | 262.144K | $0.6 | $3.6 | — |