No provider description is available for this model yet.
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
Llama Guard 4 is a Llama 4 Scout-derived multimodal pretrained model, fine-tuned for content safety classification. Similar to previous versions, it can be used to classify content in both LLM...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
This model always redirects to the latest model in the Kimi family.
This model always redirects to the latest model in the Gemini Flash family.
No provider description is available for this model yet.
Laguna XS 2.1 is the latest coding agent model in the 33B-A3B category from [Poolside](https://poolside.ai/) and a step forward from their Laguna XS.2 model (released in April 2026). It combines...
No provider description is available for this model yet.
NVIDIA Nemotron 3.5 Content Safety is a compact 4B-parameter multimodal guardrail model from NVIDIA, fine-tuned from Google Gemma-3-4B. It moderates both inputs to and responses from LLMs and VLMs, accepting...
No provider description is available for this model yet.
Mistral Medium 3.5 is a dense 128B instruction-following model from Mistral AI. It supports text and image inputs with text output, and is designed for agentic workflows, coding, and complex...
No provider description is available for this model yet.
Ring-2.6-1T is a 1T-parameter-scale thinking model with 63B active parameters, built for real-world agent workflows that require both strong capability and operational efficiency. It is optimized for coding agents, tool...
No provider description is available for this model yet.
Mistral Small 3.1 24B Instruct is an upgraded variant of Mistral Small 3 (2501), featuring 24 billion parameters with advanced multimodal capabilities. It provides state-of-the-art performance in text-based reasoning and...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
GPT-5 Image Mini combines OpenAI's advanced language capabilities, powered by [GPT-5 Mini](https://openrouter.ai/openai/gpt-5-mini), with GPT Image 1 Mini for efficient image generation. This natively multimodal model features superior instruction following, text...
The Qwen3.5 native vision-language Flash models are built on a hybrid architecture that integrates a linear attention mechanism with a sparse mixture-of-experts model, achieving higher inference efficiency. Compared to the...
The Qwen3.5 series 397B-A17B native vision-language model is built on a hybrid architecture that integrates a linear attention mechanism with a sparse mixture-of-experts model, achieving higher inference efficiency. It delivers...
MiniMax-M2.5 is a SOTA large language model designed for real-world productivity. Trained in a diverse range of complex real-world digital working environments, M2.5 builds upon the coding expertise of M2.1...
Opus 4.6 is Anthropic’s strongest model for coding and long-running professional tasks. It is built for agents that operate across entire workflows rather than single prompts, making it especially effective...
No provider description is available for this model yet.
The simplest way to get free inference. openrouter/free is a router that selects free models at random from the models available on OpenRouter. The router smartly filters for models that...
No provider description is available for this model yet.
The largest model in the Ministral 3 family, Ministral 3 14B offers frontier capabilities and performance comparable to its larger Mistral Small 3.2 24B counterpart. A powerful and efficient language...
No provider description is available for this model yet.
Qwen3 Coder Flash is Alibaba's fast and cost efficient version of their proprietary Qwen3 Coder Plus. It is a powerful coding agent model specializing in autonomous programming via tool calling...
No provider description is available for this model yet.
Voxtral Small is an enhancement of Mistral Small 3, incorporating state-of-the-art audio input capabilities while retaining best-in-class text performance. It excels at speech transcription, translation and audio understanding. Input audio...
Muse Spark 1.3 Contributor is the cost-efficient contributor tier of Meta’s multimodal reasoning model for experimentation, learning, and early-stage agentic, multi-agent, and coding workflows. It is designed to track information...
Granite-4.0-H-Micro is a 3B parameter from the Granite 4 family of models. These models are the latest in a series of models released by IBM. They are fine-tuned for long...
No provider description is available for this model yet.
Claude Haiku 4.5 is Anthropic’s fastest and most efficient model, delivering near-frontier intelligence at a fraction of the cost and latency of larger Claude models. Matching Claude Sonnet 4’s performance...
No provider description is available for this model yet.
Qwen3-Next-80B-A3B-Instruct is an instruction-tuned chat model in the Qwen3-Next series optimized for fast, stable responses without “thinking” traces. It targets complex tasks across reasoning, code generation, knowledge QA, and multilingual...
No provider description is available for this model yet.
Qwen3-235B-A22B-Instruct-2507 is a multilingual, instruction-tuned mixture-of-experts language model based on the Qwen3-235B architecture, with 22B active parameters per forward pass. It is optimized for general-purpose text generation, including instruction following,...
No provider description is available for this model yet.
No provider description is available for this model yet.
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| microsoft/Dayhoff-170M-GRS-SS-110000microsoft/Dayhoff-170M-GRS-SS-110000 | Not documented | — | — | — | |||
| Meta: Llama Guard 4 12Bmeta-llama/llama-guard-4-12b | 163.84K | $0.18 | $0.18 | — | |||
| ministral-14b-2512mistral/ministral-14b-2512 | 262.144K | $0.2 | $0.2 | — | |||
| nvidia/NVIDIA-Nemotron-Labs-Teacher-General-Reasoningnvidia/NVIDIA-Nemotron-Labs-Teacher-General-Reasoning | Not documented | — | — | — | |||
| openai/gpt-5.6-solopenrouter/openai/gpt-5.6-sol | 1.05M | $2 | $10 | — | |||
| x-ai/grok-4.5openrouter/x-ai/grok-4.5 | 500K | $2 | $6 | — | |||
| x-ai/grok-4.3openrouter/x-ai/grok-4.3 | 1M | $1.25 | $2.5 | — | |||
| qwen-plus-2025-01-25qwencloud/qwen-plus-2025-01-25 | 129.024K | $0.4 | $1.2 | — | |||
| x-ai/grok-4.20-multi-agentopenrouter/x-ai/grok-4.20-multi-agent | 1M | $1.25 | $2.5 | — | |||
| x-ai/grok-4.20openrouter/x-ai/grok-4.20 | 1M | $1.25 | $2.5 | — | |||
| qwen-plusqwencloud/qwen-plus | 129.024K | $0.4 | $1.2 | — | |||
| meta-llama/Llama-3.1-70B-Instructwandb/meta-llama/llama-3.1-70b-instruct | 128K | $0.8 | $0.8 | — | |||
| MoonshotAI: Kimi Latest~moonshotai/kimi-latest | 1.04858M | $2.1 | $10.95 | — | |||
| Google: Gemini Flash Latest~google/gemini-flash-latest | 1.04858M | $0.75 | $3.75 | — | |||
| openai/o4-miniopenrouter/openai/o4-mini | 200K | $1.1 | $4.4 | — | |||
| Poolside: Laguna XS 2.1 (free)poolside/laguna-xs-2.1:free | 262.144K | Free | Free | — | |||
| openai/o3openrouter/openai/o3 | 200K | $2 | $8 | — | |||
| NVIDIA: Nemotron 3.5 Content Safety (free)nvidia/nemotron-3.5-content-safety:free | 128K | Free | Free | — | |||
| qwen-maxqwencloud/qwen-max | 30.72K | $1.6 | $6.4 | — | |||
| Mistral: Mistral Medium 3.5mistralai/mistral-medium-3-5 | 262.144K | $1.5 | $7.5 | — | |||
| openai/gpt-5.6-terraopenrouter/openai/gpt-5.6-terra | 922K | $2 | $12 | — | |||
| inclusionAI: Ring-2.6-1Tinclusionai/ring-2.6-1t | 262.144K | $0.075 | $0.625 | — | |||
| openai/gpt-5.6-lunaopenrouter/openai/gpt-5.6-luna | 922K | $0.2 | $1.2 | — | |||
| Mistral: Mistral Small 3.1 24Bmistralai/mistral-small-3.1-24b-instruct | 128K | $0.351 | $0.555 | — | |||
| qwen-flash-2025-07-28qwencloud/qwen-flash-2025-07-28 | 997.952K | — | — | — | |||
| JetBrains/Mellum2-12B-A2.5B-Instructwandb/jetbrains/mellum2-12b-a2.5b-instruct | 131.072K | $0.05 | $0.1 | — | |||
| microsoft/Dayhoff-170M-GRS-SS-122000microsoft/Dayhoff-170M-GRS-SS-122000 | Not documented | — | — | — | |||
| OpenAI: GPT-5 Image Miniopenai/gpt-5-image-mini | 400K | $2.5 | $2 | — | |||
| Qwen: Qwen3.5-Flashqwen/qwen3.5-flash-02-23 | 1M | $0.065 | $0.26 | — | |||
| Qwen: Qwen3.5 397B A17Bqwen/qwen3.5-397b-a17b | 262.144K | $0.55 | $3.5 | — | |||
| MiniMax: MiniMax M2.5minimax/minimax-m2.5 | 200K | $0.27 | $1.08 | — | |||
| Anthropic: Claude Opus 4.6anthropic/claude-opus-4.6 | 1M | $5 | $25 | — | |||
| openai/gpt-5.5openrouter/openai/gpt-5.5 | 1.05M | $5 | $30 | — | |||
| Free Models Routeropenrouter/free | 200K | Free | Free | — | |||
| openai/gpt-5.4-nanoopenrouter/openai/gpt-5.4-nano | 272K | $0.2 | $1.25 | — | |||
| Mistral: Ministral 3 14B 2512mistralai/ministral-14b-2512 | 262.144K | $0.2 | $0.2 | — | |||
| openai/gpt-5.4-miniopenrouter/openai/gpt-5.4-mini | 272K | $0.75 | $4.5 | — | |||
| Qwen: Qwen3 Coder Flashqwen/qwen3-coder-flash | 1M | $0.195 | $0.975 | — | |||
| openai/gpt-5.4openrouter/openai/gpt-5.4 | 1.05M | $2.5 | $15 | — | |||
| Mistral: Voxtral Small 24B 2507mistralai/voxtral-small-24b-2507 | 32.768K | $0.1 | $0.3 | — | |||
| Meta: Muse Spark 1.3 Contributormeta/muse-spark-1.3-contributor | 1.04858M | $0.1 | $0.2 | — | |||
| IBM: Granite 4.0 Microibm-granite/granite-4.0-h-micro | 131K | $0.017 | $0.112 | — | |||
| ibm-granite/granite-4.1-8bwandb/ibm-granite/granite-4.1-8b | 131.072K | $0.05 | $0.1 | — | |||
| Anthropic: Claude Haiku 4.5anthropic/claude-haiku-4.5 | 200K | $1 | $5 | — | |||
| openai/gpt-5.3-codexopenrouter/openai/gpt-5.3-codex | 272K | $1.75 | $14 | — | |||
| Qwen: Qwen3 Next 80B A3B Instructqwen/qwen3-next-80b-a3b-instruct | 262.144K | $0.09 | $1.1 | — | |||
| openai/gpt-5.1openrouter/openai/gpt-5.1 | 272K | $1.25 | $10 | — | |||
| Qwen: Qwen3 235B A22B Instruct 2507qwen/qwen3-235b-a22b-2507 | 262.144K | $0.087 | $0.35 | — | |||
| qwen-flashqwencloud/qwen-flash | 997.952K | — | — | — | |||
| openai/gpt-4o-miniopenrouter/openai/gpt-4o-mini | 128K | $0.15 | $0.6 | — |