No provider description is available for this model yet.
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Hermes 3 is a generalist language model with many improvements over Hermes 2, including advanced agentic capabilities, much better roleplaying, reasoning, multi-turn conversation, long context coherence, and improvements across the...
No provider description is available for this model yet.
GPT-5.5 Pro is OpenAI’s high-capability model optimized for deep reasoning and accuracy on complex, high-stakes workloads. It features a 1M+ token context window (922K input, 128K output) with support for...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
This model offers four times the context length of gpt-3.5-turbo, allowing it to support approximately 20 pages of text in a single request at a higher cost. Training data: up...
Gemini 2.5 Pro is Google’s state-of-the-art AI model designed for advanced reasoning, coding, mathematics, and scientific tasks. It employs “thinking” capabilities, enabling it to reason through responses with enhanced accuracy...
No provider description is available for this model yet.
No provider description is available for this model yet.
Amazon Nova Pro 1.0 is a capable multimodal model from Amazon focused on providing a combination of accuracy, speed, and cost for a wide range of tasks. As of December...
No provider description is available for this model yet.
No provider description is available for this model yet.
The Auto Router automatically selects the best model for your prompt, powered by the wisdom of the market. It routes you based on what the OpenRouter community collectively spends on...
Morph's fastest apply model for code edits. ~10,500 tokens/sec with 96% accuracy for rapid code transformations. The model requires the prompt to be in the following format: <instruction>{instruction}</instruction> <code>{initial_code}</code> <update>{edit_snippet}</update>...
Qwen3.6-Max-Preview is a proprietary frontier model from Alibaba Cloud built on a sparse mixture-of-experts architecture with approximately 1 trillion total parameters. It is optimized for agentic coding, tool use, and...
No provider description is available for this model yet.
No provider description is available for this model yet.
Reka Flash 3 is a general-purpose, instruction-tuned large language model with 21 billion parameters, developed by Reka. It excels at general chat, coding tasks, instruction-following, and function calling. Featuring a...
No provider description is available for this model yet.
No provider description is available for this model yet.
Command A is an open-weights 111B parameter model with a 256k context window focused on delivering great performance across agentic, multilingual, and coding use cases. Compared to other leading proprietary...
Morph's high-accuracy apply model for complex code edits. ~4,500 tokens/sec with 98% accuracy for precise code transformations. The model requires the prompt to be in the following format: <instruction>{instruction}</instruction> <code>{initial_code}</code>...
No provider description is available for this model yet.
No provider description is available for this model yet.
DeepSeek V3, a 685B-parameter, mixture-of-experts model, is the latest iteration of the flagship chat model family from the DeepSeek team. It succeeds the [DeepSeek V3](/deepseek/deepseek-chat-v3) model and performs really well...
No provider description is available for this model yet.
No provider description is available for this model yet.
Llama 4 Scout 17B Instruct (16E) is a mixture-of-experts (MoE) language model developed by Meta, activating 17 billion parameters out of a total of 109B. It supports native multimodal input...
Hunyuan-A13B is a 13B active parameter Mixture-of-Experts (MoE) language model developed by Tencent, with a total parameter count of 80B and support for reasoning via Chain-of-Thought. It offers competitive benchmark...
Qwen3.6-35B-A3B is an open-weight multimodal model from Alibaba Cloud with 35 billion total parameters and 3 billion active parameters per token. It uses a hybrid sparse mixture-of-experts architecture combining Gated...
GPT-5.6 Sol Pro is the same underlying model as [GPT-5.6 Sol](https://openrouter.ai/openai/gpt-5.6-sol), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. Learn more in OpenAI's docs: https://developers.openai.com/api/docs/guides/reasoning#reasoning-mode
GPT-5.6 Luna Pro is the same underlying model as [GPT-5.6 Luna](https://openrouter.ai/openai/gpt-5.6-luna), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. Learn more in OpenAI's docs: https://developers.openai.com/api/docs/guides/reasoning#reasoning-mode
Qwen3-30B-A3B-Thinking-2507 is a 30B parameter Mixture-of-Experts reasoning model optimized for complex tasks requiring extended multi-step thinking. The model is designed specifically for “thinking mode,” where internal reasoning traces are separated...
No provider description is available for this model yet.
No provider description is available for this model yet.
Qwen3-32B is a dense 32.8B parameter causal language model from the Qwen3 series, optimized for both complex reasoning and efficient dialogue. It supports seamless switching between a "thinking" mode for...
No provider description is available for this model yet.
No provider description is available for this model yet.
Qwen3-8B is a dense 8.2B parameter causal language model from the Qwen3 series, designed for both reasoning-heavy tasks and efficient dialogue. It supports seamless switching between "thinking" mode for math,...
Venice Uncensored Dolphin Mistral 24B Venice Edition is a fine-tuned variant of Mistral-Small-24B-Instruct-2501, developed by dphn.ai in collaboration with Venice.ai. This model is designed as an “uncensored” instruct-tuned LLM, preserving...
No provider description is available for this model yet.
No provider description is available for this model yet.
Qwen3, the latest generation in the Qwen large language model series, features both dense and mixture-of-experts (MoE) architectures to excel in reasoning, multilingual support, and advanced agent tasks. Its unique...
No provider description is available for this model yet.
No provider description is available for this model yet.
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| FW-DeepSeek-V3.2azure_ai/fw-deepseek-v3.2 | 163.84K | $0.62 | $1.85 | — | |||
| chatdolphinnlp_cloud/chatdolphin | 16.384K | $0.5 | $0.5 | — | |||
| us.anthropic.claude-opus-4-6-v1bedrock_converse/us.anthropic.claude-opus-4-6-v1 | 1M | $5.5 | $27.5 | — | |||
| eu.anthropic.claude-opus-4-6-v1bedrock_converse/eu.anthropic.claude-opus-4-6-v1 | 1M | $5.5 | $27.5 | — | |||
| Nous: Hermes 3 405B Instruct (free)nousresearch/hermes-3-llama-3.1-405b:free | 131.072K | Free | Free | — | |||
| deepseek-ai/DeepSeek-V3-0324gmi/deepseek-ai/deepseek-v3-0324 | 163.84K | $0.28 | $0.88 | — | |||
| OpenAI: GPT-5.5 Pro (batch)openai/gpt-5.5-pro:batch | 1.05M | $15 | $90 | — | |||
| google/gemini-3-pro-previewgmi/google/gemini-3-pro-preview | 1.04858M | $2 | $12 | — | |||
| google/gemini-3-flash-previewgmi/google/gemini-3-flash-preview | 1.04858M | $0.5 | $3 | — | |||
| au.anthropic.claude-opus-4-6-v1bedrock_converse/au.anthropic.claude-opus-4-6-v1 | 1M | $5.5 | $27.5 | — | |||
| OpenAI: GPT-3.5 Turbo 16kopenai/gpt-3.5-turbo-16k | 16.385K | $3 | $4 | — | |||
| Google: Gemini 2.5 Pro Preview 06-05google/gemini-2.5-pro-preview | 1.04858M | $1.25 | $10 | — | |||
| us-gov.anthropic.claude-3-haiku-20240307-v1:0bedrock_converse/us-gov.anthropic.claude-3-haiku-20240307-v1:0 | 200K | $0.3 | $1.5 | — | |||
| us-east-1/mistral.mistral-large-2402-v1:0bedrock/us-east-1/mistral.mistral-large-2402-v1:0 | 32K | $8 | $24 | — | |||
| Amazon: Nova Pro 1.0amazon/nova-pro-v1 | 300K | $0.8 | $3.2 | — | |||
| us-east-1/mistral.mistral-7b-instruct-v0:2bedrock/us-east-1/mistral.mistral-7b-instruct-v0:2 | 32K | $0.15 | $0.2 | — | |||
| llama3:8bollama/llama3:8b | 8.192K | — | — | — | |||
| Auto Routeropenrouter/auto | 2M | — | — | — | |||
| Morph: Morph V3 Fastmorph/morph-v3-fast | 81.92K | $0.8 | $1.2 | — | |||
| Qwen: Qwen3.6 Max Previewqwen/qwen3.6-max-preview | 262.144K | $1.027 | $6.162 | — | |||
| llama3:70bollama/llama3:70b | 8.192K | — | — | — | |||
| us-east-1/meta.llama3-8b-instruct-v1:0bedrock/us-east-1/meta.llama3-8b-instruct-v1:0 | 8.192K | $0.3 | $0.6 | — | |||
| Reka Flash 3rekaai/reka-flash-3 | 65.536K | $0.1 | $0.2 | — | |||
| us-east-1/meta.llama3-70b-instruct-v1:0bedrock/us-east-1/meta.llama3-70b-instruct-v1:0 | 8.192K | $2.65 | $3.5 | — | |||
| us-east-1/anthropic.claude-v2:1bedrock/us-east-1/anthropic.claude-v2:1 | 100K | $8 | $24 | — | |||
| Cohere: Command Acohere/command-a | 256K | $2.5 | $10 | — | |||
| Morph: Morph V3 Largemorph/morph-v3-large | 262.144K | $0.9 | $1.9 | — | |||
| us-east-1/anthropic.claude-v1bedrock/us-east-1/anthropic.claude-v1 | 100K | $8 | $24 | — | |||
| global.xai.grok-4.6bedrock_converse/global.xai.grok-4.6 | 500K | $2 | $6 | — | |||
| DeepSeek: DeepSeek V3 0324deepseek/deepseek-chat-v3-0324 | 163.84K | $0.25 | $1 | — | |||
| us.xai.grok-4.6bedrock_converse/us.xai.grok-4.6 | 500K | $2.2 | $6.6 | — | |||
| us-east-1/anthropic.claude-instant-v1bedrock/us-east-1/anthropic.claude-instant-v1 | 100K | $0.8 | $2.4 | — | |||
| Meta: Llama 4 Scoutmeta-llama/llama-4-scout | 327.68K | $0.1 | $0.3 | — | |||
| Tencent: Hunyuan A13B Instructtencent/hunyuan-a13b-instruct | 131.072K | $0.14 | $0.57 | — | |||
| Qwen: Qwen3.6 35B A3Bqwen/qwen3.6-35b-a3b | 262.144K | $0.1 | $0.9 | — | |||
| OpenAI: GPT-5.6 Sol Proopenai/gpt-5.6-sol-pro | 1.05M | $2 | $10 | — | |||
| OpenAI: GPT-5.6 Luna Proopenai/gpt-5.6-luna-pro | 1.05M | $0.2 | $1.2 | — | |||
| Qwen: Qwen3 30B A3B Thinking 2507qwen/qwen3-30b-a3b-thinking-2507 | 81.92K | $0.2 | $2.4 | — | |||
| us-gov-west-1/amazon.nova-micro-v1:0bedrock/us-gov-west-1/amazon.nova-micro-v1:0 | 128K | $0.042 | $0.168 | — | |||
| us-gov-west-1/amazon.nova-lite-v1:0bedrock/us-gov-west-1/amazon.nova-lite-v1:0 | 300K | $0.072 | $0.288 | — | |||
| Qwen: Qwen3 32Bqwen/qwen3-32b | 40.96K | $0.08 | $0.28 | — | |||
| llama3.1ollama/llama3.1 | 8.192K | — | — | — | |||
| us-east-1/6-month-commitment/anthropic.claude-v2:1bedrock/us-east-1/6-month-commitment/anthropic.claude-v2:1 | 100K | — | — | — | |||
| Qwen: Qwen3 8Bqwen/qwen3-8b | 131.072K | $0.117 | $0.455 | — | |||
| Venice: Uncensoredcognitivecomputations/dolphin-mistral-24b-venice-edition | 128K | $0.2 | $0.9 | — | |||
| xai.grok-4.6bedrock_mantle/xai.grok-4.6 | 500K | $2.2 | $6.6 | — | |||
| deepseek/deepseek-v4-pro-0813openrouter/deepseek/deepseek-v4-pro-0813 | 1.04858M | $1.32 | $3.96 | — | |||
| Qwen: Qwen3 30B A3Bqwen/qwen3-30b-a3b | 40.96K | $0.12 | $0.5 | — | |||
| us-east-1/6-month-commitment/anthropic.claude-v1bedrock/us-east-1/6-month-commitment/anthropic.claude-v1 | 100K | — | — | — | |||
| deepseek/deepseek-v4-proopenrouter/deepseek/deepseek-v4-pro | 1.04858M | $1.32 | $3.96 | — |