NVIDIA Nemotron 3.5 Content Safety is a compact 4B-parameter multimodal guardrail model from NVIDIA, fine-tuned from Google Gemma-3-4B. It moderates both inputs to and responses from LLMs and VLMs, accepting...
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
The Qwen3.5 series 397B-A17B native vision-language model is built on a hybrid architecture that integrates a linear attention mechanism with a sparse mixture-of-experts model, achieving higher inference efficiency. It delivers...
Opus 4.6 is Anthropic’s strongest model for coding and long-running professional tasks. It is built for agents that operate across entire workflows rather than single prompts, making it especially effective...
The simplest way to get free inference. openrouter/free is a router that selects free models at random from the models available on OpenRouter. The router smartly filters for models that...
No provider description is available for this model yet.
No provider description is available for this model yet.
This model always redirects to the latest model in the OpenAI GPT family.
This model always redirects to the latest model in the Claude Sonnet family.
This model always redirects to the latest model in the Gemini Flash family.
This model always redirects to the latest model in the Kimi family.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
This model always redirects to the latest model in the Gemini Pro family.
This model always redirects to the latest model in the GPT Mini family.
No provider description is available for this model yet.
No provider description is available for this model yet.
This model always redirects to the latest model in the Claude Haiku family.
No provider description is available for this model yet.
NVIDIA Nemotron™ 3 Nano Omni is a 30B-A3B open multimodal model designed to function as a perception and context sub-agent in enterprise agent systems. It accepts text, image, video, and...
Llama Guard 4 is a Llama 4 Scout-derived multimodal pretrained model, fine-tuned for content safety classification. Similar to previous versions, it can be used to classify content in both LLM...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
UI-TARS-1.5 is a multimodal vision-language agent optimized for GUI-based environments, including desktop interfaces, web browsers, mobile systems, and games. Built by ByteDance, it builds upon the UI-TARS framework with reinforcement...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Claude 3 Haiku is Anthropic's fastest and most compact model for near-instant responsiveness. Quick and accurate targeted performance. See the launch announcement and benchmark results [here](https://www.anthropic.com/news/claude-3-haiku) #multimodal
GPT-4o mini is OpenAI's newest model after [GPT-4 Omni](/models/openai/gpt-4o), supporting both text and image inputs with text outputs. As their most advanced small model, it is many multiples more affordable...
No provider description is available for this model yet.
Qwen3.5 Plus (April 2026) is a large-scale multimodal language model from Alibaba. It accepts text, image, and video input and produces text output, with a 1M token context window. This...
Amazon Nova Lite 1.0 is a very low-cost multimodal model from Amazon that focused on fast processing of image, video, and text inputs to generate text output. Amazon Nova Lite...
No provider description is available for this model yet.
No provider description is available for this model yet.
MiniMax-01 is a combines MiniMax-Text-01 for text generation and MiniMax-VL-01 for image understanding. It has 456 billion parameters, with 45.9 billion parameters activated per inference, and can handle a context...
Llama 4 Maverick 17B Instruct (128E) is a high-capacity multimodal language model from Meta, built on a mixture-of-experts (MoE) architecture with 128 experts and 17 billion active parameters per forward...
Gemini 2.5 Pro is Google’s state-of-the-art AI model designed for advanced reasoning, coding, mathematics, and scientific tasks. It employs “thinking” capabilities, enabling it to reason through responses with enhanced accuracy...
Qwen3-VL-30B-A3B-Instruct is a multimodal model that unifies strong text generation with visual understanding for images and videos. Its Instruct variant optimizes instruction-following for general multimodal tasks. It excels in perception...
No provider description is available for this model yet.
No provider description is available for this model yet.
Qwen3-VL-30B-A3B-Thinking is a multimodal model that unifies strong text generation with visual understanding for images and videos. Its Thinking variant enhances reasoning in STEM, math, and complex tasks. It excels...
[GPT-5](https://openrouter.ai/openai/gpt-5) Image combines OpenAI's GPT-5 model with state-of-the-art image generation capabilities. It offers major improvements in reasoning, code quality, and user experience while incorporating GPT Image 1's superior instruction following,...
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| NVIDIA: Nemotron 3.5 Content Safety (free)nvidia/nemotron-3.5-content-safety:free | 128K | Free | Free | — | |||
| Qwen: Qwen3.5 397B A17Bqwen/qwen3.5-397b-a17b | 262.144K | $0.55 | $3.5 | — | |||
| Anthropic: Claude Opus 4.6anthropic/claude-opus-4.6 | 1M | $5 | $25 | — | |||
| Free Models Routeropenrouter/free | 200K | Free | Free | — | |||
| qwen/qwen3-vl-235b-a22b-instructopenrouter/qwen/qwen3-vl-235b-a22b-instruct | 262.144K | $0.21 | $1.9 | — | |||
| openai/o4-miniopenrouter/openai/o4-mini | 200K | $1.1 | $4.4 | — | |||
| OpenAI GPT Latest~openai/gpt-latest | 1.05M | $2 | $10 | — | |||
| Anthropic: Claude Sonnet Latest~anthropic/claude-sonnet-latest | 1M | $2 | $10 | — | |||
| Google: Gemini Flash Latest~google/gemini-flash-latest | 1.04858M | $0.75 | $3.75 | — | |||
| MoonshotAI: Kimi Latest~moonshotai/kimi-latest | 1.04858M | $2.1 | $10.95 | — | |||
| qwen/qwen3-vl-235b-a22b-thinkingopenrouter/qwen/qwen3-vl-235b-a22b-thinking | 131.072K | $0.4 | $4 | — | |||
| openai/o3openrouter/openai/o3 | 200K | $2 | $8 | — | |||
| moonshotai/Kimi-K2.7-Codenebius/moonshotai/kimi-k2.7-code | 262.144K | $0.95 | $4 | — | |||
| databricks-gemini-3-pro-imagedatabricks/databricks-gemini-3-pro-image | 65.536K | — | — | — | |||
| openai/gpt-5-proopenrouter/openai/gpt-5-pro | 400K | $15 | $120 | — | |||
| Google: Gemini Pro Latest~google/gemini-pro-latest | 1.04858M | $2 | $12 | — | |||
| OpenAI: GPT Mini Latest~openai/gpt-mini-latest | 400K | $0.75 | $4.5 | — | |||
| qwen/qwen3-vl-30b-a3b-instructopenrouter/qwen/qwen3-vl-30b-a3b-instruct | 262.144K | $0.15 | $0.6 | — | |||
| openai/gpt-5.6-terraopenrouter/openai/gpt-5.6-terra | 922K | $2 | $12 | — | |||
| Anthropic: Claude Haiku Latest~anthropic/claude-haiku-latest | 200K | $1 | $5 | — | |||
| qwen/qwen3-vl-30b-a3b-thinkingopenrouter/qwen/qwen3-vl-30b-a3b-thinking | 262.144K | $0.2 | $2.4 | — | |||
| NVIDIA: Nemotron 3 Nano Omni (free)nvidia/nemotron-3-nano-omni-30b-a3b-reasoning:free | 256K | Free | Free | — | |||
| Meta: Llama Guard 4 12Bmeta-llama/llama-guard-4-12b | 163.84K | $0.18 | $0.18 | — | |||
| google/gemini-2.5-flash-imageopenrouter/google/gemini-2.5-flash-image | 32.768K | $0.3 | $2.5 | — | |||
| openai/gpt-5.6-lunaopenrouter/openai/gpt-5.6-luna | 922K | $0.2 | $1.2 | — | |||
| moonshotai/Kimi-K2.6nebius/moonshotai/kimi-k2.6 | 262.144K | $0.95 | $4 | — | |||
| qwen/qwen3-vl-8b-instructopenrouter/qwen/qwen3-vl-8b-instruct | 262.144K | $0.117 | $0.455 | — | |||
| qwen/qwen3-vl-8b-thinkingopenrouter/qwen/qwen3-vl-8b-thinking | 131.072K | $0.18 | $2.1 | — | |||
| openai/gpt-5.5openrouter/openai/gpt-5.5 | 1.05M | $5 | $30 | — | |||
| ByteDance: UI-TARS 7B bytedance/ui-tars-1.5-7b | 128K | $0.1 | $0.2 | — | |||
| qwen/qwen3-vl-32b-instructopenrouter/qwen/qwen3-vl-32b-instruct | 131.072K | $0.104 | $0.416 | — | |||
| openai/gpt-5.1-codex-miniopenrouter/openai/gpt-5.1-codex-mini | 400K | $0.25 | $2 | — | |||
| openai/gpt-5.4-nanoopenrouter/openai/gpt-5.4-nano | 272K | $0.2 | $1.25 | — | |||
| MiniMaxAI/MiniMax-M3nebius/minimaxai/minimax-m3 | 1.04858M | $0.3 | $1.2 | — | |||
| databricks-gemini-3-1-flash-imagedatabricks/databricks-gemini-3-1-flash-image | 131.072K | — | — | — | |||
| Anthropic: Claude 3 Haikuanthropic/claude-3-haiku | 200K | $0.25 | $1.25 | — | |||
| OpenAI: GPT-4o-mini (2024-07-18)openai/gpt-4o-mini-2024-07-18 | 128K | $0.15 | $0.6 | — | |||
| openai/gpt-5.1-codexopenrouter/openai/gpt-5.1-codex | 400K | $1.25 | $10 | — | |||
| Qwen: Qwen3.5 Plus 2026-04-20qwen/qwen3.5-plus-20260420 | 1M | $0.3 | $1.8 | — | |||
| Amazon: Nova Lite 1.0amazon/nova-lite-v1 | 300K | $0.06 | $0.24 | — | |||
| google/gemini-3-pro-image-previewopenrouter/google/gemini-3-pro-image-preview | 65.536K | $2 | $12 | — | |||
| openai/gpt-5.4-miniopenrouter/openai/gpt-5.4-mini | 272K | $0.75 | $4.5 | — | |||
| MiniMax: MiniMax-01minimax/minimax-01 | 1.00019M | $0.2 | $1.1 | — | |||
| Meta: Llama 4 Maverickmeta-llama/llama-4-maverick | 128K | $0.2 | $0.696 | — | |||
| Google: Gemini 2.5 Pro Preview 06-05google/gemini-2.5-pro-preview | 1.04858M | $1.25 | $10 | — | |||
| Qwen: Qwen3 VL 30B A3B Instructqwen/qwen3-vl-30b-a3b-instruct | 262.144K | $0.15 | $0.6 | — | |||
| z-ai/glm-4.6vopenrouter/z-ai/glm-4.6v | 131.072K | $0.3 | $0.9 | — | |||
| openai/gpt-5.4openrouter/openai/gpt-5.4 | 1.05M | $2.5 | $15 | — | |||
| Qwen: Qwen3 VL 30B A3B Thinkingqwen/qwen3-vl-30b-a3b-thinking | 131.072K | $0.2 | $2.4 | — | |||
| OpenAI: GPT-5 Imageopenai/gpt-5-image | 400K | $10 | $10 | — |