1,564 models

NVIDIA Nemotron 3.5 Content Safety is a compact 4B-parameter multimodal guardrail model from NVIDIA, fine-tuned from Google Gemma-3-4B. It moderates both inputs to and responses from LLMs and VLMs, accepting...

nvidia/nemotron-3.5-content-safety:free 128K context Free input Free output

The Qwen3.5 series 397B-A17B native vision-language model is built on a hybrid architecture that integrates a linear attention mechanism with a sparse mixture-of-experts model, achieving higher inference efficiency. It delivers...

qwen/qwen3.5-397b-a17b 262.144K context $0.55/M input $3.5/M output

Opus 4.6 is Anthropic’s strongest model for coding and long-running professional tasks. It is built for agents that operate across entire workflows rather than single prompts, making it especially effective...

anthropic/claude-opus-4.6 1M context $5/M input $25/M output

The simplest way to get free inference. openrouter/free is a router that selects free models at random from the models available on OpenRouter. The router smartly filters for models that...

openrouter/free 200K context Free input Free output

No provider description is available for this model yet.

openrouter/openai/o4-mini 200K context $1.1/M input $4.4/M output

This model always redirects to the latest model in the OpenAI GPT family.

~openai/gpt-latest 1.05M context $2/M input $10/M output

This model always redirects to the latest model in the Gemini Flash family.

~google/gemini-flash-latest 1.04858M context $0.75/M input $3.75/M output

This model always redirects to the latest model in the Kimi family.

~moonshotai/kimi-latest 1.04858M context $2.1/M input $10.95/M output

No provider description is available for this model yet.

openrouter/openai/o3 200K context $2/M input $8/M output

No provider description is available for this model yet.

nebius/moonshotai/kimi-k2.7-code 262.144K context $0.95/M input $4/M output

No provider description is available for this model yet.

openrouter/openai/gpt-5-pro 400K context $15/M input $120/M output

This model always redirects to the latest model in the Gemini Pro family.

~google/gemini-pro-latest 1.04858M context $2/M input $12/M output

This model always redirects to the latest model in the GPT Mini family.

~openai/gpt-mini-latest 400K context $0.75/M input $4.5/M output

No provider description is available for this model yet.

openrouter/openai/gpt-5.6-terra 922K context $2/M input $12/M output

NVIDIA Nemotron™ 3 Nano Omni is a 30B-A3B open multimodal model designed to function as a perception and context sub-agent in enterprise agent systems. It accepts text, image, video, and...

nvidia/nemotron-3-nano-omni-30b-a3b-reasoning:free 256K context Free input Free output

Llama Guard 4 is a Llama 4 Scout-derived multimodal pretrained model, fine-tuned for content safety classification. Similar to previous versions, it can be used to classify content in both LLM...

meta-llama/llama-guard-4-12b 163.84K context $0.18/M input $0.18/M output

No provider description is available for this model yet.

openrouter/openai/gpt-5.6-luna 922K context $0.2/M input $1.2/M output

No provider description is available for this model yet.

nebius/moonshotai/kimi-k2.6 262.144K context $0.95/M input $4/M output

No provider description is available for this model yet.

openrouter/qwen/qwen3-vl-8b-instruct 262.144K context $0.117/M input $0.455/M output

No provider description is available for this model yet.

openrouter/openai/gpt-5.5 1.05M context $5/M input $30/M output

UI-TARS-1.5 is a multimodal vision-language agent optimized for GUI-based environments, including desktop interfaces, web browsers, mobile systems, and games. Built by ByteDance, it builds upon the UI-TARS framework with reinforcement...

bytedance/ui-tars-1.5-7b 128K context $0.1/M input $0.2/M output

No provider description is available for this model yet.

openrouter/openai/gpt-5.4-nano 272K context $0.2/M input $1.25/M output

No provider description is available for this model yet.

nebius/minimaxai/minimax-m3 1.04858M context $0.3/M input $1.2/M output

Claude 3 Haiku is Anthropic's fastest and most compact model for near-instant responsiveness. Quick and accurate targeted performance. See the launch announcement and benchmark results [here](https://www.anthropic.com/news/claude-3-haiku) #multimodal

anthropic/claude-3-haiku 200K context $0.25/M input $1.25/M output

GPT-4o mini is OpenAI's newest model after [GPT-4 Omni](/models/openai/gpt-4o), supporting both text and image inputs with text outputs. As their most advanced small model, it is many multiples more affordable...

openai/gpt-4o-mini-2024-07-18 128K context $0.15/M input $0.6/M output

No provider description is available for this model yet.

openrouter/openai/gpt-5.1-codex 400K context $1.25/M input $10/M output

Qwen3.5 Plus (April 2026) is a large-scale multimodal language model from Alibaba. It accepts text, image, and video input and produces text output, with a 1M token context window. This...

qwen/qwen3.5-plus-20260420 1M context $0.3/M input $1.8/M output

Amazon Nova Lite 1.0 is a very low-cost multimodal model from Amazon that focused on fast processing of image, video, and text inputs to generate text output. Amazon Nova Lite...

amazon/nova-lite-v1 300K context $0.06/M input $0.24/M output

No provider description is available for this model yet.

openrouter/openai/gpt-5.4-mini 272K context $0.75/M input $4.5/M output

MiniMax-01 is a combines MiniMax-Text-01 for text generation and MiniMax-VL-01 for image understanding. It has 456 billion parameters, with 45.9 billion parameters activated per inference, and can handle a context...

minimax/minimax-01 1.00019M context $0.2/M input $1.1/M output

Llama 4 Maverick 17B Instruct (128E) is a high-capacity multimodal language model from Meta, built on a mixture-of-experts (MoE) architecture with 128 experts and 17 billion active parameters per forward...

meta-llama/llama-4-maverick 128K context $0.2/M input $0.696/M output

Gemini 2.5 Pro is Google’s state-of-the-art AI model designed for advanced reasoning, coding, mathematics, and scientific tasks. It employs “thinking” capabilities, enabling it to reason through responses with enhanced accuracy...

google/gemini-2.5-pro-preview 1.04858M context $1.25/M input $10/M output

Qwen3-VL-30B-A3B-Instruct is a multimodal model that unifies strong text generation with visual understanding for images and videos. Its Instruct variant optimizes instruction-following for general multimodal tasks. It excels in perception...

qwen/qwen3-vl-30b-a3b-instruct 262.144K context $0.15/M input $0.6/M output

No provider description is available for this model yet.

openrouter/z-ai/glm-4.6v 131.072K context $0.3/M input $0.9/M output

No provider description is available for this model yet.

openrouter/openai/gpt-5.4 1.05M context $2.5/M input $15/M output

Qwen3-VL-30B-A3B-Thinking is a multimodal model that unifies strong text generation with visual understanding for images and videos. Its Thinking variant enhances reasoning in STEM, math, and complex tasks. It excels...

qwen/qwen3-vl-30b-a3b-thinking 131.072K context $0.2/M input $2.4/M output

[GPT-5](https://openrouter.ai/openai/gpt-5) Image combines OpenAI's GPT-5 model with state-of-the-art image generation capabilities. It offers major improvements in reasoning, code quality, and user experience while incorporating GPT Image 1's superior instruction following,...

openai/gpt-5-image 400K context $10/M input $10/M output