4,182 models

Amazon Nova Premier is the most capable of Amazon’s multimodal models for complex reasoning tasks and for use as the best teacher for distilling custom models.

amazon/nova-premier-v1 1M context $2.5/M input $12.5/M output

Olmo 3 32B Think is a large-scale, 32-billion-parameter model purpose-built for deep reasoning, complex logic chains and advanced instruction-following scenarios. Its capacity enables strong performance on demanding evaluation tasks and...

allenai/olmo-3-32b-think 65.536K context $0.15/M input $0.5/M output

No provider description is available for this model yet.

openrouter/minimax/minimax-m2-her 65.536K context $0.3/M input $1.2/M output

The smallest model in the Ministral 3 family, Ministral 3 3B is a powerful, efficient tiny language model with vision capabilities.

mistralai/ministral-3b-2512 131.072K context $0.1/M input $0.1/M output

Transform your natural language requests into structured OpenRouter API request objects. Describe what you want to accomplish with AI models, and Body Builder will construct the appropriate API calls. Example:...

openrouter/bodybuilder 128K context Input not listed Output not listed

GLM-4.6V is a large multimodal model designed for high-fidelity visual understanding and long-context reasoning across images, documents, and mixed media. It supports up to 128K tokens, processes complex page layouts...

z-ai/glm-4.6v 131.072K context $0.3/M input $0.9/M output

The relace-search model uses 4-12 `view_file` and `grep` tools in parallel to explore a codebase and return relevant files to the user request. In contrast to RAG, relace-search performs agentic...

relace/relace-search 256K context $1/M input $3/M output

No provider description is available for this model yet.

openrouter/qwen/qwen3-coder-next 262.144K context $0.12/M input $0.8/M output

No provider description is available for this model yet.

watsonx/bigscience/mt0-xxl 4.096K context $1.908/M input $1.908/M output

Cogito v2.1 671B MoE represents one of the strongest open models globally, matching performance of frontier closed and open models. This model is trained using self play with reinforcement learning...

deepcogito/cogito-v2.1-671b 128K context $1.25/M input $1.25/M output

Nova 2 Lite is a fast, cost-effective reasoning model for everyday workloads that can process text, images, and videos to generate text. Nova 2 Lite demonstrates standout capabilities in processing...

amazon/nova-2-lite-v1 1M context $0.3/M input $2.5/M output

No provider description is available for this model yet.

openrouter/qwen/qwen3-max-thinking 262.144K context $0.78/M input $3.9/M output

GPT-5.2 Chat (AKA Instant) is the fast, lightweight member of the 5.2 family, optimized for low-latency chat while retaining strong general intelligence. It uses adaptive reasoning to selectively “think” on...

openai/gpt-5.2-chat 128K context $1.75/M input $14/M output

GLM-4.7 is Z.ai’s latest flagship model, featuring upgrades in two key areas: enhanced programming capabilities and more stable multi-step reasoning/execution. It demonstrates significant improvements in executing complex agent tasks while...

z-ai/glm-4.7 202.752K context $0.4/M input $1.75/M output

MiniMax-M2.1 is a lightweight, state-of-the-art large language model optimized for coding, agentic workflows, and modern application development. With only 10 billion activated parameters, it delivers a major jump in real-world...

minimax/minimax-m2.1 204.8K context $0.3/M input $1.2/M output

The gpt-audio model is OpenAI's first generally available audio model. The new snapshot features an upgraded decoder for more natural sounding voices and maintains better voice consistency. Audio is priced...

openai/gpt-audio 128K context $2.5/M input $10/M output

MiniMax M2-her is a dialogue-first large language model built for immersive roleplay, character-driven chat, and expressive multi-turn conversations. Designed to stay consistent in tone and personality, it supports rich message...

minimax/minimax-m2-her 65.536K context $0.3/M input $1.2/M output

Solar Pro 3 is Upstage's powerful Mixture-of-Experts (MoE) language model. With 102B total parameters and 12B active parameters per forward pass, it delivers exceptional performance while maintaining computational efficiency. Optimized...

upstage/solar-pro-3 131.072K context $0.15/M input $0.6/M output

No provider description is available for this model yet.

openrouter/openai/gpt-5.4-pro 1.05M context $30/M input $180/M output

No provider description is available for this model yet.

nebius/zai-org/glm-5.3-flash 1.024M context $0.15/M input $0.5/M output

No provider description is available for this model yet.

nebius/zai-org/glm-5.2 1.04858M context $1.4/M input $4.4/M output

The Qwen3.5 native vision-language series Plus models are built on a hybrid architecture that integrates linear attention mechanisms with sparse mixture-of-experts models, achieving higher inference efficiency. In a variety of...

qwen/qwen3.5-plus-02-15 1M context $0.26/M input $1.56/M output

Aion-2.0 is a variant of DeepSeek V3.2 optimized for immersive roleplaying and storytelling. It is particularly strong at introducing tension, crises, and conflict into stories, making narratives feel more engaging....

aion-labs/aion-2.0 131.072K context $0.8/M input $1.6/M output

No provider description is available for this model yet.

openrouter/qwen/qwen3.5-9b 262.144K context $0.1/M input $0.15/M output

The Qwen3.5 122B-A10B native vision-language model is built on a hybrid architecture that integrates a linear attention mechanism with a sparse mixture-of-experts model, achieving higher inference efficiency. In terms of...

qwen/qwen3.5-122b-a10b 262.144K context $0.26/M input $2.08/M output

The Qwen3.5 27B native vision-language Dense model incorporates a linear attention mechanism, delivering fast response times while balancing inference speed and performance. Its overall capabilities are comparable to those of...

qwen/qwen3.5-27b 262.144K context $0.195/M input $1.56/M output

The Qwen3.5 Series 35B-A3B is a native vision-language model designed with a hybrid architecture that integrates linear attention mechanisms and a sparse mixture-of-experts model, achieving higher inference efficiency. Its overall...

qwen/qwen3.5-35b-a3b 256K context $0.312/M input $1.25/M output

Qwen3.5-9B is a multimodal foundation model from the Qwen3.5 family, designed to deliver strong reasoning, coding, and visual understanding in an efficient 9B-parameter architecture. It uses a unified vision-language design...

qwen/qwen3.5-9b 262.144K context $0.1/M input $0.15/M output

NVIDIA Nemotron 3 Super is a 120B-parameter open hybrid MoE model, activating just 12B parameters for maximum compute efficiency and accuracy in complex multi-agent applications. Built on a hybrid Mamba-Transformer...

nvidia/nemotron-3-super-120b-a12b:free 262.144K context Free input Free output

Reka Edge is an extremely efficient 7B multimodal vision-language model that accepts image/video+text inputs and generates text outputs. This model is optimized specifically to deliver industry-leading performance in image understanding,...

rekaai/reka-edge 16.384K context $0.1/M input $0.1/M output

No provider description is available for this model yet.

openrouter/z-ai/glm-5-turbo 202.752K context $1.2/M input $4/M output

No provider description is available for this model yet.

nebius/zai-org/glm-5.1 202.752K context $1.4/M input $4.4/M output

KAT-Coder-Pro V2 is the latest high-performance model in KwaiKAT’s KAT-Coder series, designed for complex enterprise-grade software engineering and SaaS integration. It builds on the agentic coding strengths of earlier versions,...

kwaipilot/kat-coder-pro-v2 262.144K context $0.3/M input $1.2/M output

[GPT-5.4](https://openrouter.ai/openai/gpt-5.4) Image 2 combines OpenAI's GPT-5.4 model with state-of-the-art image generation capabilities from GPT Image 2. It enables rich multimodal workflows, allowing users to seamlessly move between reasoning, coding, and...

openai/gpt-5.4-image-2 272K context $8/M input $15/M output

Mistral Small 4 is the next major release in the Mistral Small family, unifying the capabilities of several flagship Mistral models into a single system. It combines strong reasoning from...

mistralai/mistral-small-2603 262.144K context $0.15/M input $0.6/M output

The Pareto Router maintains a tiered shortlist of strong coding models, ranked by [Artificial Analysis](https://artificialanalysis.ai/) coding percentiles. Set min_coding_score between 0 and 1 on the [pareto-router plugin](https://openrouter.ai/docs/guides/routing/routers/pareto-router#the-min_coding_score-parameter) to control how...

openrouter/pareto-code 2M context Input not listed Output not listed

Ling-2.6-flash is an instant (instruct) model from inclusionAI with 104B total parameters and 7.4B active parameters, designed for real-world agents that require fast responses, strong execution, and high token efficiency....

inclusionai/ling-2.6-flash 262.144K context $0.01/M input $0.03/M output

No provider description is available for this model yet.

openrouter/minimax/minimax-m2.7 204.8K context $0.3/M input $1.2/M output

Qwen3.6 27B is a dense 27-billion-parameter language model from the Qwen Team at Alibaba, released in April 2026. It features hybrid multimodal capabilities — accepting text, image, and video inputs...

qwen/qwen3.6-27b 262.144K context $0.3/M input $2/M output

Qwen3.6-Max-Preview is a proprietary frontier model from Alibaba Cloud built on a sparse mixture-of-experts architecture with approximately 1 trillion total parameters. It is optimized for agentic coding, tool use, and...

qwen/qwen3.6-max-preview 262.144K context $1.027/M input $6.162/M output

No provider description is available for this model yet.

openrouter/z-ai/glm-5v-turbo 202.752K context $1.2/M input $4/M output

No provider description is available for this model yet.

nebius/qwen/qwen3.5-397b-a17b 262.144K context $0.6/M input $3.6/M output