No provider description is available for this model yet.
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
No provider description is available for this model yet.
No provider description is available for this model yet.
The Meta Llama 3.3 multilingual large language model (LLM) is a pretrained and instruction tuned generative model in 70B (text in/text out). The Llama 3.3 instruction tuned text only model...
No provider description is available for this model yet.
No provider description is available for this model yet.
Euryale L3.3 70B is a model focused on creative roleplay from [Sao10k](https://ko-fi.com/sao10k). It is the successor of [Euryale L3 70B v2.2](/models/sao10k/l3-euryale-70b).
[GPT-5.4](https://openrouter.ai/openai/gpt-5.4) Image 2 combines OpenAI's GPT-5.4 model with state-of-the-art image generation capabilities from GPT Image 2. It enables rich multimodal workflows, allowing users to seamlessly move between reasoning, coding, and...
Fast-mode variant of [Opus 4.8](/anthropic/claude-opus-4.8) - identical capabilities with higher output speed at 2x pricing relative to regular Opus 4.8. Learn more in Anthropic's docs: https://platform.claude.com/docs/en/build-with-claude/fast-mode Note: As of September...
No provider description is available for this model yet.
No provider description is available for this model yet.
WizardLM-2 8x22B is Microsoft AI's most advanced Wizard model. It demonstrates highly competitive performance compared to leading proprietary models, and it consistently outperforms all existing state-of-the-art opensource models. It is...
No provider description is available for this model yet.
No provider description is available for this model yet.
Mistral's official instruct fine-tuned version of [Mixtral 8x22B](/models/mistralai/mixtral-8x22b). It uses 39B active parameters out of 141B, offering unparalleled cost efficiency for its size. Its strengths include: - strong math, coding,...
[Microsoft Research](/microsoft) Phi-4 is designed to perform well in complex reasoning tasks and can operate efficiently in situations with limited memory or where quick responses are needed. At 14 billion...
No provider description is available for this model yet.
No provider description is available for this model yet.
Qwen2.5 7B is the latest series of Qwen large language models. Qwen2.5 brings the following improvements upon Qwen2: - Significantly more knowledge and has greatly improved capabilities in coding and...
No provider description is available for this model yet.
No provider description is available for this model yet.
This is a series of models designed to replicate the prose quality of the Claude 3 models, specifically Sonnet(https://openrouter.ai/anthropic/claude-3.5-sonnet) and Opus(https://openrouter.ai/anthropic/claude-3-opus). The model is fine-tuned on top of [Qwen2.5 72B](https://openrouter.ai/qwen/qwen-2.5-72b-instruct).
MiniMax-01 is a combines MiniMax-Text-01 for text generation and MiniMax-VL-01 for image understanding. It has 456 billion parameters, with 45.9 billion parameters activated per inference, and can handle a context...
Mistral Small 4 is the next major release in the Mistral Small family, unifying the capabilities of several flagship Mistral models into a single system. It combines strong reasoning from...
No provider description is available for this model yet.
No provider description is available for this model yet.
UnslopNemo v4.1 is the latest addition from the creator of Rocinante, designed for adventure writing and role-play scenarios.
No provider description is available for this model yet.
No provider description is available for this model yet.
Qwen2.5-Coder is the latest series of Code-Specific Qwen large language models (formerly known as CodeQwen). Qwen2.5-Coder brings the following improvements upon CodeQwen1.5: - Significantly improvements in **code generation**, **code reasoning**...
DeepSeek R1 Distill Llama 70B is a distilled large language model based on [Llama-3.3-70B-Instruct](/meta-llama/llama-3.3-70b-instruct), using outputs from [DeepSeek R1](/deepseek/deepseek-r1). The model combines advanced distillation techniques to achieve high performance across...
No provider description is available for this model yet.
No provider description is available for this model yet.
This is Mistral AI's flagship model, Mistral Large 2 (version mistral-large-2407). It's a proprietary weights-available model and excels at reasoning, code, JSON, chat, and more. Read the launch announcement [here](https://mistral.ai/news/mistral-large-2407/)....
No provider description is available for this model yet.
No provider description is available for this model yet.
This model offers four times the context length of gpt-3.5-turbo, allowing it to support approximately 20 pages of text in a single request at a higher cost. Training data: up...
OpenAI o3-mini-high is the same model as [o3-mini](/openai/o3-mini) with reasoning_effort set to high. o3-mini is a cost-efficient language model optimized for STEM reasoning tasks, particularly excelling in science, mathematics, and...
The Pareto Router maintains a tiered shortlist of strong coding models, ranked by [Artificial Analysis](https://artificialanalysis.ai/) coding percentiles. Set min_coding_score between 0 and 1 on the [pareto-router plugin](https://openrouter.ai/docs/guides/routing/routers/pareto-router#the-min_coding_score-parameter) to control how...
Qwen3-VL-8B-Instruct is a multimodal vision-language model from the Qwen3-VL series, built for high-fidelity understanding and reasoning across text, images, and video. It features improved multimodal fusion with Interleaved-MRoPE for long-horizon...
Mistral Small 3 is a 24B-parameter language model optimized for low-latency performance across common AI tasks. Released under the Apache 2.0 license, it features both pre-trained and instruction-tuned versions designed...
No provider description is available for this model yet.
No provider description is available for this model yet.
Amazon Nova Pro 1.0 is a capable multimodal model from Amazon focused on providing a combination of accuracy, speed, and cost for a wide range of tasks. As of December...
No provider description is available for this model yet.
No provider description is available for this model yet.
The Auto Router automatically selects the best model for your prompt, powered by the wisdom of the market. It routes you based on what the OpenRouter community collectively spends on...
Mistral Saba is a 24B-parameter language model specifically designed for the Middle East and South Asia, delivering accurate and contextually relevant responses while maintaining efficient performance. Trained on curated regional...
No provider description is available for this model yet.
No provider description is available for this model yet.
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| kimi-k3fireworks_ai/kimi-k3 | 1.04858M | $3 | $15 | — | |||
| glm-5p2-fast-usfireworks_ai/glm-5p2-fast-us | 1.04858M | $2.1 | $6.6 | — | |||
| FW-Kimi-K2.5azure_ai/fw-kimi-k2.5 | 262.144K | $0.66 | $3.3 | — | |||
| Meta: Llama 3.3 70B Instructmeta-llama/llama-3.3-70b-instruct | 131.072K | $0.1 | $0.32 | — | |||
| us-gov.nvidia.nemotron-nano-9b-v2bedrock_converse/us-gov.nvidia.nemotron-nano-9b-v2 | 128K | $0.072 | $0.276 | — | |||
| glm-5p2-fastfireworks_ai/glm-5p2-fast | 1.04858M | $2.1 | $6.6 | — | |||
| Sao10K: Llama 3.3 Euryale 70Bsao10k/l3.3-euryale-70b | 131.072K | $0.65 | $0.75 | — | |||
| OpenAI: GPT-5.4 Image 2openai/gpt-5.4-image-2 | 272K | $8 | $15 | — | |||
| Anthropic: Claude Opus 4.8 (Fast)anthropic/claude-opus-4.8-fast | 1M | $10 | $50 | — | |||
| deepseek-v4-flash-0731fireworks_ai/deepseek-v4-flash-0731 | 1.04858M | $0.14 | $0.28 | — | |||
| FW-Inklingazure_ai/fw-inkling | 1.04858M | $1 | $4.05 | — | |||
| WizardLM-2 8x22Bmicrosoft/wizardlm-2-8x22b | 65.535K | $0.62 | $0.62 | — | |||
| FW-GLM-5.2-Fastazure_ai/fw-glm-5.2-fast | 1.04858M | $2.1 | $6.6 | — | |||
| us-gov.nvidia.nemotron-nano-12b-v2bedrock_converse/us-gov.nvidia.nemotron-nano-12b-v2 | 128K | $0.24 | $0.72 | — | |||
| Mistral: Mixtral 8x22B Instructmistralai/mixtral-8x22b-instruct | 65.536K | $2 | $6 | — | |||
| Microsoft: Phi 4microsoft/phi-4 | 16.384K | $0.07 | $0.14 | — | |||
| us-gov.nvidia.nemotron-nano-3-30bbedrock_converse/us-gov.nvidia.nemotron-nano-3-30b | 262.144K | $0.072 | $0.288 | — | |||
| FW-GLM-5.2azure_ai/fw-glm-5.2 | 1.04858M | $1.54 | $4.84 | — | |||
| Qwen: Qwen2.5 7B Instructqwen/qwen-2.5-7b-instruct | 32.768K | $0.1 | $0.2 | — | |||
| FW-GLM-5.1azure_ai/fw-glm-5.1 | 202.8K | $1.54 | $4.84 | — | |||
| accounts/fireworks/models/kimi-k3fireworks_ai/accounts/fireworks/models/kimi-k3 | 1.04858M | $3 | $15 | — | |||
| Magnum v4 72Banthracite-org/magnum-v4-72b | 32.768K | $2.5 | $5 | — | |||
| MiniMax: MiniMax-01minimax/minimax-01 | 1.00019M | $0.2 | $1.1 | — | |||
| Mistral: Mistral Small 4mistralai/mistral-small-2603 | 262.144K | $0.15 | $0.6 | — | |||
| accounts/fireworks/models/deepseek-v4-flash-0731fireworks_ai/accounts/fireworks/models/deepseek-v4-flash-0731 | 1.04858M | $0.14 | $0.28 | — | |||
| us-gov.anthropic.claude-fable-5-1bedrock_converse/us-gov.anthropic.claude-fable-5-1 | 1M | $12 | $60 | — | |||
| TheDrummer: UnslopNemo 12Bthedrummer/unslopnemo-12b | 1.024M | $0.4 | $0.4 | — | |||
| us-gov.anthropic.claude-opus-5bedrock_converse/us-gov.anthropic.claude-opus-5 | 1M | $6 | $30 | — | |||
| mistralollama/mistral | 8.192K | — | — | — | |||
| Qwen2.5 Coder 32B Instructqwen/qwen-2.5-coder-32b-instruct | 32.768K | $0.66 | $1 | — | |||
| DeepSeek: R1 Distill Llama 70Bdeepseek/deepseek-r1-distill-llama-70b | 8.192K | $0.8 | $0.8 | — | |||
| us-east-1/deepseek.v3.2bedrock/us-east-1/deepseek.v3.2 | 163.84K | $0.62 | $1.85 | — | |||
| us-east-1/mistral.mixtral-8x7b-instruct-v0:1bedrock/us-east-1/mistral.mixtral-8x7b-instruct-v0:1 | 32K | $0.45 | $0.7 | — | |||
| Mistral Large 2407mistralai/mistral-large-2407 | 131.072K | $2 | $6 | — | |||
| FW-GLM-5azure_ai/fw-glm-5 | 200K | $1.1 | $3.52 | — | |||
| FW-DeepSeek-V4-Proazure_ai/fw-deepseek-v4-pro | 1M | $1.925 | $3.828 | — | |||
| OpenAI: GPT-3.5 Turbo 16kopenai/gpt-3.5-turbo-16k | 16.385K | $3 | $4 | — | |||
| OpenAI: o3 Mini Highopenai/o3-mini-high | 200K | $1.1 | $4.4 | — | |||
| Pareto Code Routeropenrouter/pareto-code | 2M | — | — | — | |||
| Qwen: Qwen3 VL 8B Instructqwen/qwen3-vl-8b-instruct | 131.072K | $0.117 | $0.455 | — | |||
| Mistral: Mistral Small 3mistralai/mistral-small-24b-instruct-2501 | 32.768K | $0.05 | $0.08 | — | |||
| FW-DeepSeek-V3.2azure_ai/fw-deepseek-v3.2 | 163.84K | $0.62 | $1.85 | — | |||
| us-gov.anthropic.claude-3-haiku-20240307-v1:0bedrock_converse/us-gov.anthropic.claude-3-haiku-20240307-v1:0 | 200K | $0.3 | $1.5 | — | |||
| Amazon: Nova Pro 1.0amazon/nova-pro-v1 | 300K | $0.8 | $3.2 | — | |||
| us-east-1/mistral.mistral-large-2402-v1:0bedrock/us-east-1/mistral.mistral-large-2402-v1:0 | 32K | $8 | $24 | — | |||
| us-east-1/mistral.mistral-7b-instruct-v0:2bedrock/us-east-1/mistral.mistral-7b-instruct-v0:2 | 32K | $0.15 | $0.2 | — | |||
| Auto Routeropenrouter/auto | 2M | — | — | — | |||
| Mistral: Sabamistralai/mistral-saba | 32.768K | $0.2 | $0.6 | — | |||
| llama3:8bollama/llama3:8b | 8.192K | — | — | — | |||
| llama3:70bollama/llama3:70b | 8.192K | — | — | — |