3,223 models

DeepSeek V4 Flash Vision Exp is an experimental vision-enabled version of [DeepSeek V4 Flash 0731](https://openrouter.ai/deepseek/deepseek-v4-flash-0731) from DeepSeek, adding image understanding while matching the base model on text capabilities including agents,...

deepseek/deepseek-v4-flash-vision-exp:batch 1.04858M context $0.11/M input $0.33/M output

Amazon Nova Pro 1.0 is a capable multimodal model from Amazon focused on providing a combination of accuracy, speed, and cost for a wide range of tasks. As of December...

amazon/nova-pro-v1 300K context $0.8/M input $3.2/M output

Skyfall 36B v2 is an enhanced iteration of Mistral Small 2501, specifically fine-tuned for improved creativity, nuanced writing, role-playing, and coherent storytelling.

thedrummer/skyfall-36b-v2 32.768K context $0.55/M input $0.8/M output

No provider description is available for this model yet.

azure_ai/gpt-5.5-2026-04-23 1.05M context $5/M input $30/M output

Qwen3.6 27B is a dense 27-billion-parameter language model from the Qwen Team at Alibaba, released in April 2026. It features hybrid multimodal capabilities — accepting text, image, and video inputs...

qwen/qwen3.6-27b 262.144K context $0.3/M input $2/M output

Qwen2.5 72B is the latest series of Qwen large language models. Qwen2.5 brings the following improvements upon Qwen2: - Significantly more knowledge and has greatly improved capabilities in coding and...

qwen/qwen-2.5-72b-instruct 32.768K context $0.36/M input $0.4/M output

No provider description is available for this model yet.

azure_ai/gpt-5.4 1.05M context $2.5/M input $15/M output

No provider description is available for this model yet.

xai/grok-build-latest 500K context $2/M input $6/M output

No provider description is available for this model yet.

azure/command-r-plus 128K context $3/M input $15/M output

No provider description is available for this model yet.

scaleway/glm-5.2 256K context $1.8/M input $5.5/M output

No provider description is available for this model yet.

scaleway/deepseek-v4-flash-0731 256K context $0.4/M input $0.8/M output

No provider description is available for this model yet.

azure_ai/fw-glm-5.2-fast 1.04858M context $2.1/M input $6.6/M output

The Auto Router automatically selects the best model for your prompt, powered by the wisdom of the market. It routes you based on what the OpenRouter community collectively spends on...

openrouter/auto 2M context Input not listed Output not listed

Llama 3.2 3B is a 3-billion-parameter multilingual large language model, optimized for advanced natural language processing tasks like dialogue generation, reasoning, and summarization. Designed with the latest transformer architecture, it...

meta-llama/llama-3.2-3b-instruct:free 131.072K context Free input Free output

No provider description is available for this model yet.

azure_ai/gpt-5.4-2026-03-05 1.05M context $2.5/M input $15/M output

Gemma 4 31B Instruct is Google DeepMind's 30.7B dense multimodal model supporting text and image input with text output. Features a 256K token context window, configurable thinking/reasoning mode, native function...

google/gemma-4-31b-it:batch 262.144K context $0.39/M input $0.97/M output

Seed 2.1 Turbo is a multimodal model from ByteDance Seed for coding and long-horizon agent workflows. It is suited for end-to-end software delivery, multi-step task execution, and understanding visual and...

bytedance-seed/seed-2-1-turbo 262.144K context $0.5/M input $2.5/M output

No provider description is available for this model yet.

azure_ai/gpt-5.4-mini 272K context $0.75/M input $4.5/M output

No provider description is available for this model yet.

azure_ai/fw-glm-5.2 1.04858M context $1.54/M input $4.84/M output

Inkling Small is an open-weight multimodal mixture-of-experts model from Thinking Machines Lab, with 12B active parameters out of 276B total. It is positioned as the smaller, more efficient member of...

thinkingmachines/inkling-small:free 1.04858M context Free input Free output

No provider description is available for this model yet.

azure_ai/gpt-5.4-mini-2026-03-17 272K context $0.75/M input $4.5/M output

Reka Flash 3 is a general-purpose, instruction-tuned large language model with 21 billion parameters, developed by Reka. It excels at general chat, coding tasks, instruction-following, and function calling. Featuring a...

rekaai/reka-flash-3 65.536K context $0.1/M input $0.2/M output

No provider description is available for this model yet.

azure_ai/gpt-5.4-nano-2026-03-17 272K context $0.2/M input $1.25/M output

Gemini 2.5 Pro is Google’s state-of-the-art AI model designed for advanced reasoning, coding, mathematics, and scientific tasks. It employs “thinking” capabilities, enabling it to reason through responses with enhanced accuracy...

google/gemini-2.5-pro-preview 1.04858M context $1.25/M input $10/M output