No provider description is available for this model yet.
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
No provider description is available for this model yet.
No provider description is available for this model yet.
This model always redirects to the latest model in the Gemini Pro family.
No provider description is available for this model yet.
Ling 3.0 Flash Sante is a health and medicine-focused mixture-of-experts model from InclusionAI, built on Ling 3.0 Flash with 5.1B active parameters out of 124B total. It is designed for...
No provider description is available for this model yet.
This model always redirects to the latest model in the Gemini Flash family.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
NVIDIA Nemotron 3.5 Content Safety is a compact 4B-parameter multimodal guardrail model from NVIDIA, fine-tuned from Google Gemma-3-4B. It moderates both inputs to and responses from LLMs and VLMs, accepting...
Nex-N2-Pro is an agentic mixture-of-experts model from Nex AGI, with 17B active parameters out of 397B total. Built on the Qwen3.5 architecture, it accepts text and image input and produces...
Mistral Medium 3.5 is a dense 128B instruction-following model from Mistral AI. It supports text and image inputs with text output, and is designed for agentic workflows, coding, and complex...
No provider description is available for this model yet.
Granite 4.1 8B is a dense, decoder-only 8-billion-parameter language model from IBM, part of the Granite 4.1 family. It supports a 131K-token context window and is designed for enterprise tasks...
No provider description is available for this model yet.
Mistral Small 3.1 24B Instruct is an upgraded variant of Mistral Small 3 (2501), featuring 24 billion parameters with advanced multimodal capabilities. It provides state-of-the-art performance in text-based reasoning and...
No provider description is available for this model yet.
Gemma 4 31B Instruct is Google DeepMind's 30.7B dense multimodal model supporting text and image input with text output. Features a 256K token context window, configurable thinking/reasoning mode, native function...
No provider description is available for this model yet.
GLM-5 Turbo is a new model from Z.ai designed for fast inference and strong performance in agent-driven environments such as OpenClaw scenarios. It is deeply optimized for real-world agent workflows...
No provider description is available for this model yet.
The Qwen3.5 series 397B-A17B native vision-language model is built on a hybrid architecture that integrates a linear attention mechanism with a sparse mixture-of-experts model, achieving higher inference efficiency. It delivers...
No provider description is available for this model yet.
Qwen3-Coder-Next is an open-weight causal language model optimized for coding agents and local development workflows. It uses a sparse MoE design with 80B total parameters and only 3B activated per...
No provider description is available for this model yet.
Opus 4.6 is Anthropic’s strongest model for coding and long-running professional tasks. It is built for agents that operate across entire workflows rather than single prompts, making it especially effective...
NVIDIA Nemotron 3 Nano 30B A3B is a small language MoE model with highest compute efficiency and accuracy for developers to build specialized agentic AI systems. The model is fully...
The largest model in the Ministral 3 family, Ministral 3 14B offers frontier capabilities and performance comparable to its larger Mistral Small 3.2 24B counterpart. A powerful and efficient language...
No provider description is available for this model yet.
The preview GPT-4 model with improved instruction following, JSON mode, reproducible outputs, parallel function calling, and more. Training data: up to Dec 2023. **Note:** heavily rate limited by OpenAI while...
No provider description is available for this model yet.
Voxtral Small is an enhancement of Mistral Small 3, incorporating state-of-the-art audio input capabilities while retaining best-in-class text performance. It excels at speech transcription, translation and audio understanding. Input audio...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Qwen3-Next-80B-A3B-Instruct is an instruction-tuned chat model in the Qwen3-Next series optimized for fast, stable responses without “thinking” traces. It targets complex tasks across reasoning, code generation, knowledge QA, and multilingual...
Claude Opus 4.8 is Anthropic's most capable generally available model in the Opus family. It supports text, image, and file inputs with text output, with reasoning support and a 1M-token...
Qwen3-VL-8B-Thinking is the reasoning-optimized variant of the Qwen3-VL-8B multimodal model, designed for advanced visual and textual reasoning across complex scenes, documents, and temporal sequences. It integrates enhanced multimodal alignment and...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
May 28th update to the [original DeepSeek R1](/deepseek/deepseek-r1) Performance on par with [OpenAI o1](/openai/o1), but open-sourced and with fully open reasoning tokens. It's 671B parameters in size, with 37B active...
No provider description is available for this model yet.
Qwen3-8B is a dense 8.2B parameter causal language model from the Qwen3 series, designed for both reasoning-heavy tasks and efficient dialogue. It supports seamless switching between "thinking" mode for math,...
No provider description is available for this model yet.
Llama 4 Scout 17B Instruct (16E) is a mixture-of-experts (MoE) language model developed by Meta, activating 17 billion parameters out of a total of 109B. It supports native multimodal input...
DeepSeek V3, a 685B-parameter, mixture-of-experts model, is the latest iteration of the flagship chat model family from the DeepSeek team. It succeeds the [DeepSeek V3](/deepseek/deepseek-chat-v3) model and performs really well...
No provider description is available for this model yet.
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| qwen-turbo-2025-04-28qwencloud/qwen-turbo-2025-04-28 | 1M | $0.05 | $0.2 | — | |||
| openai/gpt-5.3-codexopenrouter/openai/gpt-5.3-codex | 272K | $1.75 | $14 | — | |||
| openai/gpt-5.1openrouter/openai/gpt-5.1 | 272K | $1.25 | $10 | — | |||
| Google: Gemini Pro Latest~google/gemini-pro-latest | 1.04858M | $2 | $12 | — | |||
| qwen-turbo-2024-11-01qwencloud/qwen-turbo-2024-11-01 | 1M | $0.05 | $0.2 | — | |||
| inclusionAI: Ling 3.0 Flash Sante (free)inclusionai/ling-3.0-flash-sante:free | 262.144K | Free | Free | — | |||
| openai/gpt-4o-miniopenrouter/openai/gpt-4o-mini | 128K | $0.15 | $0.6 | — | |||
| Google: Gemini Flash Latest~google/gemini-flash-latest | 1.04858M | $0.75 | $3.75 | — | |||
| google/gemini-3.8-flashopenrouter/google/gemini-3.8-flash | 1.04858M | $0.75 | $3.75 | — | |||
| qwen-turboqwencloud/qwen-turbo | 129.024K | $0.05 | $0.2 | — | |||
| google/gemini-3.7-flashopenrouter/google/gemini-3.7-flash | 1.04858M | $0.75 | $3.75 | — | |||
| NVIDIA: Nemotron 3.5 Content Safety (free)nvidia/nemotron-3.5-content-safety:free | 128K | Free | Free | — | |||
| Nex AGI: Nex-N2-Pronex-agi/nex-n2-pro | 262.144K | $0.25 | $1 | — | |||
| Mistral: Mistral Medium 3.5mistralai/mistral-medium-3-5 | 262.144K | $1.5 | $7.5 | — | |||
| google/gemini-3.6-flashopenrouter/google/gemini-3.6-flash | 1.04858M | $0.75 | $3.75 | — | |||
| IBM: Granite 4.1 8Bibm-granite/granite-4.1-8b | 131.072K | $0.05 | $0.1 | — | |||
| qwen-plus-latestqwencloud/qwen-plus-latest | 997.952K | — | — | — | |||
| Mistral: Mistral Small 3.1 24Bmistralai/mistral-small-3.1-24b-instruct | 128K | $0.351 | $0.555 | — | |||
| Qwen/Qwen3-30B-A3B-Instruct-2507wandb/qwen/qwen3-30b-a3b-instruct-2507 | 262.144K | $0.1 | $0.3 | — | |||
| Google: Gemma 4 31B (free)google/gemma-4-31b-it:free | 262.144K | Free | Free | — | |||
| google/gemini-3.5-flash-liteopenrouter/google/gemini-3.5-flash-lite | 1.04858M | $0.3 | $2.5 | — | |||
| Z.ai: GLM 5 Turboz-ai/glm-5-turbo | 202.752K | $1.2 | $4 | — | |||
| google/gemini-3.5-flashopenrouter/google/gemini-3.5-flash | 1.04858M | $1.5 | $9 | — | |||
| Qwen: Qwen3.5 397B A17Bqwen/qwen3.5-397b-a17b | 262.144K | $0.55 | $3.5 | — | |||
| qwen-plus-2025-09-11qwencloud/qwen-plus-2025-09-11 | 997.952K | — | — | — | |||
| Qwen: Qwen3 Coder Nextqwen/qwen3-coder-next | 262.144K | $0.12 | $0.8 | — | |||
| google/gemini-2.5-flash-liteopenrouter/google/gemini-2.5-flash-lite | 1.04858M | $0.1 | $0.4 | — | |||
| Anthropic: Claude Opus 4.6anthropic/claude-opus-4.6 | 1M | $5 | $25 | — | |||
| NVIDIA: Nemotron 3 Nano 30B A3B (free)nvidia/nemotron-3-nano-30b-a3b:free | 256K | Free | Free | — | |||
| Mistral: Ministral 3 14B 2512mistralai/ministral-14b-2512 | 262.144K | $0.2 | $0.2 | — | |||
| anthropic/claude-sonnet-5openrouter/anthropic/claude-sonnet-5 | 1M | $2 | $10 | — | |||
| OpenAI: GPT-4 Turbo Previewopenai/gpt-4-turbo-preview | 128K | $10 | $30 | — | |||
| qwen-plus-2025-07-28qwencloud/qwen-plus-2025-07-28 | 997.952K | — | — | — | |||
| Mistral: Voxtral Small 24B 2507mistralai/voxtral-small-24b-2507 | 32.768K | $0.1 | $0.3 | — | |||
| Qwen/Qwen3.5-35B-A3Bwandb/qwen/qwen3.5-35b-a3b | 262.144K | $0.25 | $1.25 | — | |||
| xiaomimimo/mimo-v2.5-pronovita/xiaomimimo/mimo-v2.5-pro | 1.04858M | $0.522 | $1.044 | — | |||
| ministral-14b-latestmistral/ministral-14b-latest | 262.144K | $0.2 | $0.2 | — | |||
| Qwen: Qwen3 Next 80B A3B Instructqwen/qwen3-next-80b-a3b-instruct | 262.144K | $0.09 | $1.1 | — | |||
| Anthropic: Claude Opus 4.8 (batch)anthropic/claude-opus-4.8:batch | 1M | $2.5 | $12.5 | — | |||
| Qwen: Qwen3 VL 8B Thinkingqwen/qwen3-vl-8b-thinking | 131.072K | $0.18 | $2.1 | — | |||
| anthropic/claude-opus-4.8openrouter/anthropic/claude-opus-4.8 | 1M | $5 | $25 | — | |||
| qwen-plus-2025-07-14qwencloud/qwen-plus-2025-07-14 | 129.024K | $0.4 | $1.2 | — | |||
| anthropic/claude-fable-5.1openrouter/anthropic/claude-fable-5.1 | 1M | $10 | $50 | — | |||
| DeepSeek: R1 0528deepseek/deepseek-r1-0528 | 163.84K | $0.5 | $2.15 | — | |||
| qwen-plus-2025-04-28qwencloud/qwen-plus-2025-04-28 | 129.024K | $0.4 | $1.2 | — | |||
| Qwen: Qwen3 8Bqwen/qwen3-8b | 131.072K | $0.117 | $0.455 | — | |||
| Qwen/Qwen3.6-27Bwandb/qwen/qwen3.6-27b | 262.144K | $0.6 | $3.6 | — | |||
| Meta: Llama 4 Scoutmeta-llama/llama-4-scout | 327.68K | $0.1 | $0.3 | — | |||
| DeepSeek: DeepSeek V3 0324deepseek/deepseek-chat-v3-0324 | 163.84K | $0.25 | $1 | — | |||
| qwen-plus-2025-01-25qwencloud/qwen-plus-2025-01-25 | 129.024K | $0.4 | $1.2 | — |