DeepSeek V4 Flash Vision Exp is an experimental vision-enabled version of [DeepSeek V4 Flash 0731](https://openrouter.ai/deepseek/deepseek-v4-flash-0731) from DeepSeek, adding image understanding while matching the base model on text capabilities including agents,...
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
Amazon Nova Pro 1.0 is a capable multimodal model from Amazon focused on providing a combination of accuracy, speed, and cost for a wide range of tasks. As of December...
Skyfall 36B v2 is an enhanced iteration of Mistral Small 2501, specifically fine-tuned for improved creativity, nuanced writing, role-playing, and coherent storytelling.
No provider description is available for this model yet.
No provider description is available for this model yet.
Qwen3.6 27B is a dense 27-billion-parameter language model from the Qwen Team at Alibaba, released in April 2026. It features hybrid multimodal capabilities — accepting text, image, and video inputs...
Qwen2.5 72B is the latest series of Qwen large language models. Qwen2.5 brings the following improvements upon Qwen2: - Significantly more knowledge and has greatly improved capabilities in coding and...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
The Auto Router automatically selects the best model for your prompt, powered by the wisdom of the market. It routes you based on what the OpenRouter community collectively spends on...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Llama 3.2 3B is a 3-billion-parameter multilingual large language model, optimized for advanced natural language processing tasks like dialogue generation, reasoning, and summarization. Designed with the latest transformer architecture, it...
No provider description is available for this model yet.
No provider description is available for this model yet.
Gemma 4 31B Instruct is Google DeepMind's 30.7B dense multimodal model supporting text and image input with text output. Features a 256K token context window, configurable thinking/reasoning mode, native function...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Seed 2.1 Turbo is a multimodal model from ByteDance Seed for coding and long-horizon agent workflows. It is suited for end-to-end software delivery, multi-step task execution, and understanding visual and...
No provider description is available for this model yet.
No provider description is available for this model yet.
Inkling Small is an open-weight multimodal mixture-of-experts model from Thinking Machines Lab, with 12B active parameters out of 276B total. It is positioned as the smaller, more efficient member of...
No provider description is available for this model yet.
No provider description is available for this model yet.
Reka Flash 3 is a general-purpose, instruction-tuned large language model with 21 billion parameters, developed by Reka. It excels at general chat, coding tasks, instruction-following, and function calling. Featuring a...
No provider description is available for this model yet.
No provider description is available for this model yet.
Gemini 2.5 Pro is Google’s state-of-the-art AI model designed for advanced reasoning, coding, mathematics, and scientific tasks. It employs “thinking” capabilities, enabling it to reason through responses with enhanced accuracy...
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| DeepSeek: DeepSeek V4 Flash Vision Exp (batch)deepseek/deepseek-v4-flash-vision-exp:batch | 1.04858M | $0.11 | $0.33 | — | |||
| Amazon: Nova Pro 1.0amazon/nova-pro-v1 | 300K | $0.8 | $3.2 | — | |||
| TheDrummer: Skyfall 36B V2thedrummer/skyfall-36b-v2 | 32.768K | $0.55 | $0.8 | — | |||
| us-gov.anthropic.claude-sonnet-5bedrock_converse/us-gov.anthropic.claude-sonnet-5 | 1M | $2.4 | $12 | — | |||
| gpt-5.5-2026-04-23azure_ai/gpt-5.5-2026-04-23 | 1.05M | $5 | $30 | — | |||
| Qwen: Qwen3.6 27Bqwen/qwen3.6-27b | 262.144K | $0.3 | $2 | — | |||
| Qwen2.5 72B Instructqwen/qwen-2.5-72b-instruct | 32.768K | $0.36 | $0.4 | — | |||
| us-west-2/1-month-commitment/anthropic.claude-instant-v1bedrock/us-west-2/1-month-commitment/anthropic.claude-instant-v1 | 100K | — | — | — | |||
| gpt-5.4azure_ai/gpt-5.4 | 1.05M | $2.5 | $15 | — | |||
| grok-build-latestxai/grok-build-latest | 500K | $2 | $6 | — | |||
| command-r-plusazure/command-r-plus | 128K | $3 | $15 | — | |||
| accounts/fireworks/models/inklingfireworks_ai/accounts/fireworks/models/inkling | 1.04858M | $1 | $4.05 | — | |||
| glm-5.2scaleway/glm-5.2 | 256K | $1.8 | $5.5 | — | |||
| us-west-2/1-month-commitment/anthropic.claude-v1bedrock/us-west-2/1-month-commitment/anthropic.claude-v1 | 100K | — | — | — | |||
| au.anthropic.claude-sonnet-4-5-20250929-v1:0bedrock_converse/au.anthropic.claude-sonnet-4-5-20250929-v1:0 | 200K | $3.3 | $16.5 | — | |||
| deepseek-v4-flash-0731scaleway/deepseek-v4-flash-0731 | 256K | $0.4 | $0.8 | — | |||
| FW-GLM-5.2-Fastazure_ai/fw-glm-5.2-fast | 1.04858M | $2.1 | $6.6 | — | |||
| us-west-2/6-month-commitment/anthropic.claude-instant-v1bedrock/us-west-2/6-month-commitment/anthropic.claude-instant-v1 | 100K | — | — | — | |||
| openai/gpt-5.6-luna-proopenrouter/openai/gpt-5.6-luna-pro | 1.05M | $0.2 | $1.2 | — | |||
| us-west-2/6-month-commitment/anthropic.claude-v1bedrock/us-west-2/6-month-commitment/anthropic.claude-v1 | 100K | — | — | — | |||
| us-west-2/6-month-commitment/anthropic.claude-v2:1bedrock/us-west-2/6-month-commitment/anthropic.claude-v2:1 | 100K | — | — | — | |||
| us-gov-west-1/nvidia.nemotron-nano-3-30bbedrock/us-gov-west-1/nvidia.nemotron-nano-3-30b | 262.144K | $0.072 | $0.288 | — | |||
| us-gov.nvidia.nemotron-nano-12b-v2bedrock_converse/us-gov.nvidia.nemotron-nano-12b-v2 | 128K | $0.24 | $0.72 | — | |||
| us-gov-west-1/nvidia.nemotron-super-3-120bbedrock/us-gov-west-1/nvidia.nemotron-super-3-120b | 256K | $0.18 | $0.78 | — | |||
| us-west-2/anthropic.claude-instant-v1bedrock/us-west-2/anthropic.claude-instant-v1 | 100K | $0.8 | $2.4 | — | |||
| us-gov-west-1/openai.gpt-oss-20b-1:0bedrock/us-gov-west-1/openai.gpt-oss-20b-1:0 | 128K | $0.084 | $0.36 | — | |||
| Auto Routeropenrouter/auto | 2M | — | — | — | |||
| us-gov-west-1/anthropic.claude-sonnet-5bedrock/us-gov-west-1/anthropic.claude-sonnet-5 | 1M | $2.4 | $12 | — | |||
| us-gov-west-1/anthropic.claude-opus-4-8bedrock/us-gov-west-1/anthropic.claude-opus-4-8 | 1M | $6 | $30 | — | |||
| us-west-2/anthropic.claude-v1bedrock/us-west-2/anthropic.claude-v1 | 100K | $8 | $24 | — | |||
| apac.anthropic.claude-sonnet-4-20250514-v1:0bedrock_converse/apac.anthropic.claude-sonnet-4-20250514-v1:0 | 1M | $3 | $15 | — | |||
| apac.anthropic.claude-3-sonnet-20240229-v1:0bedrock/apac.anthropic.claude-3-sonnet-20240229-v1:0 | 200K | $3 | $15 | — | |||
| us-gov-east-1/nvidia.nemotron-nano-3-30bbedrock/us-gov-east-1/nvidia.nemotron-nano-3-30b | 262.144K | $0.072 | $0.288 | — | |||
| Meta: Llama 3.2 3B Instruct (free)meta-llama/llama-3.2-3b-instruct:free | 131.072K | Free | Free | — | |||
| us-gov.nvidia.nemotron-nano-3-30bbedrock_converse/us-gov.nvidia.nemotron-nano-3-30b | 262.144K | $0.072 | $0.288 | — | |||
| gpt-5.4-2026-03-05azure_ai/gpt-5.4-2026-03-05 | 1.05M | $2.5 | $15 | — | |||
| Google: Gemma 4 31B (batch)google/gemma-4-31b-it:batch | 262.144K | $0.39 | $0.97 | — | |||
| us-west-2/mistral.mistral-large-2402-v1:0bedrock/us-west-2/mistral.mistral-large-2402-v1:0 | 32K | $8 | $24 | — | |||
| apac.anthropic.claude-haiku-4-5-20251001-v1:0bedrock_converse/apac.anthropic.claude-haiku-4-5-20251001-v1:0 | 200K | $1.1 | $5.5 | — | |||
| us-west-2/mistral.mixtral-8x7b-instruct-v0:1bedrock/us-west-2/mistral.mixtral-8x7b-instruct-v0:1 | 32K | $0.45 | $0.7 | — | |||
| ByteDance Seed: Seed 2.1 Turbobytedance-seed/seed-2-1-turbo | 262.144K | $0.5 | $2.5 | — | |||
| gpt-5.4-miniazure_ai/gpt-5.4-mini | 272K | $0.75 | $4.5 | — | |||
| FW-GLM-5.2azure_ai/fw-glm-5.2 | 1.04858M | $1.54 | $4.84 | — | |||
| Thinking Machines: Inkling Small (free)thinkingmachines/inkling-small:free | 1.04858M | Free | Free | — | |||
| gpt-5.4-mini-2026-03-17azure_ai/gpt-5.4-mini-2026-03-17 | 272K | $0.75 | $4.5 | — | |||
| us-gov-east-1/openai.gpt-oss-120b-1:0bedrock/us-gov-east-1/openai.gpt-oss-120b-1:0 | 128K | $0.18 | $0.72 | — | |||
| Reka Flash 3rekaai/reka-flash-3 | 65.536K | $0.1 | $0.2 | — | |||
| us-gov-east-1/anthropic.claude-sonnet-5bedrock/us-gov-east-1/anthropic.claude-sonnet-5 | 1M | $2.4 | $12 | — | |||
| gpt-5.4-nano-2026-03-17azure_ai/gpt-5.4-nano-2026-03-17 | 272K | $0.2 | $1.25 | — | |||
| Google: Gemini 2.5 Pro Preview 06-05google/gemini-2.5-pro-preview | 1.04858M | $1.25 | $10 | — |