No provider description is available for this model yet.
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
No provider description is available for this model yet.
This is a series of models designed to replicate the prose quality of the Claude 3 models, specifically Sonnet(https://openrouter.ai/anthropic/claude-3.5-sonnet) and Opus(https://openrouter.ai/anthropic/claude-3-opus). The model is fine-tuned on top of [Qwen2.5 72B](https://openrouter.ai/qwen/qwen-2.5-72b-instruct).
Mistral's cutting-edge language model for coding released end of July 2025. Codestral specializes in low-latency, high-frequency tasks such as fill-in-the-middle (FIM), code correction and test generation. [Blog Post](https://mistral.ai/news/codestral-25-08)
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Mistral's cutting-edge language model for coding released end of July 2025. Codestral specializes in low-latency, high-frequency tasks such as fill-in-the-middle (FIM), code correction and test generation. [Blog Post](https://mistral.ai/news/codestral-25-08)
No provider description is available for this model yet.
UnslopNemo v4.1 is the latest addition from the creator of Rocinante, designed for adventure writing and role-play scenarios.
Gemma 3n E4B-it is optimized for efficient execution on mobile and low-resource devices, such as phones, laptops, and tablets. It supports multimodal inputs—including text, visual data, and audio—enabling diverse tasks...
Qwen3.6-Max-Preview is a proprietary frontier model from Alibaba Cloud built on a sparse mixture-of-experts architecture with approximately 1 trillion total parameters. It is optimized for agentic coding, tool use, and...
No provider description is available for this model yet.
No provider description is available for this model yet.
Qwen2.5-Coder is the latest series of Code-Specific Qwen large language models (formerly known as CodeQwen). Qwen2.5-Coder brings the following improvements upon CodeQwen1.5: - Significantly improvements in **code generation**, **code reasoning**...
No provider description is available for this model yet.
No provider description is available for this model yet.
This is Mistral AI's flagship model, Mistral Large 2 (version mistral-large-2407). It's a proprietary weights-available model and excels at reasoning, code, JSON, chat, and more. Read the launch announcement [here](https://mistral.ai/news/mistral-large-2407/)....
Morph's high-accuracy apply model for complex code edits. ~4,500 tokens/sec with 98% accuracy for precise code transformations. The model requires the prompt to be in the following format: <instruction>{instruction}</instruction> <code>{initial_code}</code>...
No provider description is available for this model yet.
No provider description is available for this model yet.
Amazon Nova Pro 1.0 is a capable multimodal model from Amazon focused on providing a combination of accuracy, speed, and cost for a wide range of tasks. As of December...
Venice Uncensored Dolphin Mistral 24B Venice Edition is a fine-tuned variant of Mistral-Small-24B-Instruct-2501, developed by dphn.ai in collaboration with Venice.ai. This model is designed as an “uncensored” instruct-tuned LLM, preserving...
No provider description is available for this model yet.
The Auto Router automatically selects the best model for your prompt, powered by the wisdom of the market. It routes you based on what the OpenRouter community collectively spends on...
Hunyuan-A13B is a 13B active parameter Mixture-of-Experts (MoE) language model developed by Tencent, with a total parameter count of 80B and support for reasoning via Chain-of-Thought. It offers competitive benchmark...
Qwen3.6-35B-A3B is an open-weight multimodal model from Alibaba Cloud with 35 billion total parameters and 3 billion active parameters per token. It uses a hybrid sparse mixture-of-experts architecture combining Gated...
GPT-5.6 Sol Pro is the same underlying model as [GPT-5.6 Sol](https://openrouter.ai/openai/gpt-5.6-sol), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. Learn more in OpenAI's docs: https://developers.openai.com/api/docs/guides/reasoning#reasoning-mode
GPT-5.6 Luna Pro is the same underlying model as [GPT-5.6 Luna](https://openrouter.ai/openai/gpt-5.6-luna), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. Learn more in OpenAI's docs: https://developers.openai.com/api/docs/guides/reasoning#reasoning-mode
Kimi K2 0905 is the September update of [Kimi K2 0711](moonshotai/kimi-k2). It is a large-scale Mixture-of-Experts (MoE) language model developed by Moonshot AI, featuring 1 trillion total parameters with 32...
No provider description is available for this model yet.
No provider description is available for this model yet.
Reka Flash 3 is a general-purpose, instruction-tuned large language model with 21 billion parameters, developed by Reka. It excels at general chat, coding tasks, instruction-following, and function calling. Featuring a...
No provider description is available for this model yet.
No provider description is available for this model yet.
Command A is an open-weights 111B parameter model with a 256k context window focused on delivering great performance across agentic, multilingual, and coding use cases. Compared to other leading proprietary...
Venice Uncensored Dolphin Mistral 24B Venice Edition is a fine-tuned variant of Mistral-Small-24B-Instruct-2501, developed by dphn.ai in collaboration with Venice.ai. This model is designed as an “uncensored” instruct-tuned LLM, preserving...
No provider description is available for this model yet.
DeepSeek V3, a 685B-parameter, mixture-of-experts model, is the latest iteration of the flagship chat model family from the DeepSeek team. It succeeds the [DeepSeek V3](/deepseek/deepseek-chat-v3) model and performs really well...
No provider description is available for this model yet.
Llama 4 Scout 17B Instruct (16E) is a mixture-of-experts (MoE) language model developed by Meta, activating 17 billion parameters out of a total of 109B. It supports native multimodal input...
Morph's fastest apply model for code edits. ~10,500 tokens/sec with 96% accuracy for rapid code transformations. The model requires the prompt to be in the following format: <instruction>{instruction}</instruction> <code>{initial_code}</code> <update>{edit_snippet}</update>...
Ling-2.6-1T is an instant (instruct) model from inclusionAI and the company’s trillion-parameter flagship, designed for real-world agents that require fast execution and high efficiency at scale. It uses a “fast...
No provider description is available for this model yet.
No provider description is available for this model yet.
Qwen3-32B is a dense 32.8B parameter causal language model from the Qwen3 series, optimized for both complex reasoning and efficient dialogue. It supports seamless switching between a "thinking" mode for...
No provider description is available for this model yet.
No provider description is available for this model yet.
Qwen3-8B is a dense 8.2B parameter causal language model from the Qwen3 series, designed for both reasoning-heavy tasks and efficient dialogue. It supports seamless switching between "thinking" mode for math,...
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| gpt-audio-miniazure/gpt-audio-mini | 128K | $0.6 | $2.4 | — | |||
| us-gov-west-1/amazon.nova-pro-v1:0bedrock/us-gov-west-1/amazon.nova-pro-v1:0 | 300K | $0.96 | $3.84 | — | |||
| Magnum v4 72Banthracite-org/magnum-v4-72b | 32.768K | $2.5 | $5 | — | |||
| Mistral: Codestral 2508mistralai/codestral-2508 | 256K | $0.3 | $0.9 | — | |||
| muse-glimmer-30bfireworks_ai/muse-glimmer-30b | 131.072K | $0.35 | $1.5 | — | |||
| anthropic.claude-fable-5bedrock_converse/anthropic.claude-fable-5 | 1M | $10 | $50 | — | |||
| qwen3p8-maxfireworks_ai/qwen3p8-max | 262.144K | $2 | $6 | — | |||
| us-gov-west-1/amazon.titan-text-express-v1bedrock/us-gov-west-1/amazon.titan-text-express-v1 | 42K | $1.3 | $1.7 | — | |||
| Mistral: Codestral 2508 (batch)mistralai/codestral-2508:batch | 256K | $0.15 | $0.45 | — | |||
| us.openai.gpt-5.6-solbedrock_converse/us.openai.gpt-5.6-sol | 1M | $5.5 | $33 | — | |||
| TheDrummer: UnslopNemo 12Bthedrummer/unslopnemo-12b | 1.024M | $0.4 | $0.4 | — | |||
| Google: Gemma 3n 4Bgoogle/gemma-3n-e4b-it | 32.768K | $0.06 | $0.12 | — | |||
| Qwen: Qwen3.6 Max Previewqwen/qwen3.6-max-preview | 262.144K | $1.027 | $6.162 | — | |||
| kimi-k3-usfireworks_ai/kimi-k3-us | 1.04858M | $3.3 | $16.5 | — | |||
| FW-Kimi-K2.7-Codeazure_ai/fw-kimi-k2.7-code | 262.144K | $1.05 | $4.4 | — | |||
| Qwen2.5 Coder 32B Instructqwen/qwen-2.5-coder-32b-instruct | 32.768K | $0.66 | $1 | — | |||
| FW-Kimi-K2.6azure_ai/fw-kimi-k2.6 | 262.144K | $1.045 | $4.4 | — | |||
| kimi-k3-fastfireworks_ai/kimi-k3-fast | 1.04858M | $4.5 | $22.5 | — | |||
| Mistral Large 2407mistralai/mistral-large-2407 | 131.072K | $2 | $6 | — | |||
| Morph: Morph V3 Largemorph/morph-v3-large | 262.144K | $0.9 | $1.9 | — | |||
| us-gov.xai.grok-4.6bedrock_converse/us-gov.xai.grok-4.6 | 500K | $2.64 | $7.92 | — | |||
| us-gov.openai.gpt-oss-120b-1:0bedrock_converse/us-gov.openai.gpt-oss-120b-1:0 | 128K | $0.18 | $0.72 | — | |||
| Amazon: Nova Pro 1.0amazon/nova-pro-v1 | 300K | $0.8 | $3.2 | — | |||
| Venice: Uncensored (free)cognitivecomputations/dolphin-mistral-24b-venice-edition:free | 32.768K | Free | Free | — | |||
| us-gov.openai.gpt-oss-20b-1:0bedrock_converse/us-gov.openai.gpt-oss-20b-1:0 | 128K | $0.084 | $0.36 | — | |||
| Auto Routeropenrouter/auto | 2M | — | — | — | |||
| Tencent: Hunyuan A13B Instructtencent/hunyuan-a13b-instruct | 131.072K | $0.14 | $0.57 | — | |||
| Qwen: Qwen3.6 35B A3Bqwen/qwen3.6-35b-a3b | 262.144K | $0.1 | $0.9 | — | |||
| OpenAI: GPT-5.6 Sol Proopenai/gpt-5.6-sol-pro | 1.05M | $2 | $10 | — | |||
| OpenAI: GPT-5.6 Luna Proopenai/gpt-5.6-luna-pro | 1.05M | $0.2 | $1.2 | — | |||
| MoonshotAI: Kimi K2 0905moonshotai/kimi-k2-0905 | 262.144K | $0.6 | $2.5 | — | |||
| us-gov.nvidia.nemotron-super-3-120bbedrock_converse/us-gov.nvidia.nemotron-super-3-120b | 256K | $0.18 | $0.78 | — | |||
| kimi-k3fireworks_ai/kimi-k3 | 1.04858M | $3 | $15 | — | |||
| Reka Flash 3rekaai/reka-flash-3 | 65.536K | $0.1 | $0.2 | — | |||
| glm-5p2-fast-usfireworks_ai/glm-5p2-fast-us | 1.04858M | $2.1 | $6.6 | — | |||
| FW-Kimi-K2.5azure_ai/fw-kimi-k2.5 | 262.144K | $0.66 | $3.3 | — | |||
| Cohere: Command Acohere/command-a | 256K | $2.5 | $10 | — | |||
| Venice: Uncensoredcognitivecomputations/dolphin-mistral-24b-venice-edition | 128K | $0.2 | $0.9 | — | |||
| us-gov.nvidia.nemotron-nano-9b-v2bedrock_converse/us-gov.nvidia.nemotron-nano-9b-v2 | 128K | $0.072 | $0.276 | — | |||
| DeepSeek: DeepSeek V3 0324deepseek/deepseek-chat-v3-0324 | 163.84K | $0.25 | $1 | — | |||
| glm-5p2-fastfireworks_ai/glm-5p2-fast | 1.04858M | $2.1 | $6.6 | — | |||
| Meta: Llama 4 Scoutmeta-llama/llama-4-scout | 327.68K | $0.1 | $0.3 | — | |||
| Morph: Morph V3 Fastmorph/morph-v3-fast | 81.92K | $0.8 | $1.2 | — | |||
| inclusionAI: Ling-2.6-1Tinclusionai/ling-2.6-1t | 262.144K | $0.075 | $0.625 | — | |||
| deepseek-v4-flash-0731fireworks_ai/deepseek-v4-flash-0731 | 1.04858M | $0.14 | $0.28 | — | |||
| FW-Inklingazure_ai/fw-inkling | 1.04858M | $1 | $4.05 | — | |||
| Qwen: Qwen3 32Bqwen/qwen3-32b | 40.96K | $0.08 | $0.28 | — | |||
| FW-GLM-5.2-Fastazure_ai/fw-glm-5.2-fast | 1.04858M | $2.1 | $6.6 | — | |||
| us-gov.nvidia.nemotron-nano-12b-v2bedrock_converse/us-gov.nvidia.nemotron-nano-12b-v2 | 128K | $0.24 | $0.72 | — | |||
| Qwen: Qwen3 8Bqwen/qwen3-8b | 131.072K | $0.117 | $0.455 | — |