No provider description is available for this model yet.
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Qwen3-235B-A22B-Instruct-2507 is a multilingual, instruction-tuned mixture-of-experts language model based on the Qwen3-235B architecture, with 22B active parameters per forward pass. It is optimized for general-purpose text generation, including instruction following,...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
NVIDIA Nemotron Nano 2 VL is a 12-billion-parameter open multimodal reasoning model designed for video understanding and document intelligence. It introduces a hybrid Transformer-Mamba architecture, combining transformer-level accuracy with Mamba’s...
No provider description is available for this model yet.
Voxtral Small is an enhancement of Mistral Small 3, incorporating state-of-the-art audio input capabilities while retaining best-in-class text performance. It excels at speech transcription, translation and audio understanding. Input audio...
No provider description is available for this model yet.
No provider description is available for this model yet.
Claude Opus 4.5 is Anthropic’s frontier reasoning model optimized for complex software engineering, agentic workflows, and long-horizon computer use. It offers strong multimodal capabilities, competitive performance across real-world coding and...
The preview GPT-4 model with improved instruction following, JSON mode, reproducible outputs, parallel function calling, and more. Training data: up to Dec 2023. **Note:** heavily rate limited by OpenAI while...
Granite-4.0-H-Micro is a 3B parameter from the Granite 4 family of models. These models are the latest in a series of models released by IBM. They are fine-tuned for long...
Qwen3-Next-80B-A3B-Instruct is an instruction-tuned chat model in the Qwen3-Next series optimized for fast, stable responses without “thinking” traces. It targets complex tasks across reasoning, code generation, knowledge QA, and multilingual...
Claude Haiku 4.5 is Anthropic’s fastest and most efficient model, delivering near-frontier intelligence at a fraction of the cost and latency of larger Claude models. Matching Claude Sonnet 4’s performance...
Qwen3-VL-8B-Thinking is the reasoning-optimized variant of the Qwen3-VL-8B multimodal model, designed for advanced visual and textual reasoning across complex scenes, documents, and temporal sequences. It integrates enhanced multimodal alignment and...
Qwen3-Coder-30B-A3B-Instruct is a 30.5B parameter Mixture-of-Experts (MoE) model with 128 experts (8 active per forward pass), designed for advanced code generation, repository-scale understanding, and agentic tool use. Built on the...
No provider description is available for this model yet.
ERNIE-4.5-VL-424B-A47B is a multimodal Mixture-of-Experts (MoE) model from Baidu’s ERNIE 4.5 series, featuring 424B total parameters with 47B active per token. It is trained jointly on text and image data...
May 28th update to the [original DeepSeek R1](/deepseek/deepseek-r1) Performance on par with [OpenAI o1](/openai/o1), but open-sourced and with fully open reasoning tokens. It's 671B parameters in size, with 37B active...
Qwen3 Coder Flash is Alibaba's fast and cost efficient version of their proprietary Qwen3 Coder Plus. It is a powerful coding agent model specializing in autonomous programming via tool calling...
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| mistralai/mistral-small-24b-instruct-2501openrouter/mistralai/mistral-small-24b-instruct-2501 | 32.768K | $0.05 | $0.08 | — | |||
| qwen/qwen-plusopenrouter/qwen/qwen-plus | 1M | $0.26 | $0.78 | — | |||
| z-ai/glm-5.3openrouter/z-ai/glm-5.3 | 1.31072M | $1.4 | $4.4 | — | |||
| qwen/qwen2.5-vl-72b-instructopenrouter/qwen/qwen2.5-vl-72b-instruct | 128K | $0.8 | $1 | — | |||
| mistralai/mistral-sabaopenrouter/mistralai/mistral-saba | 32.768K | $0.2 | $0.6 | — | |||
| deepseek/deepseek-v4-flash-vision-expopenrouter/deepseek/deepseek-v4-flash-vision-exp | 1.04858M | $0.22 | $0.66 | — | |||
| nvidia/Nemotron-3-Nano-Omninebius/nvidia/nemotron-3-nano-omni | 262.144K | $0.06 | $0.24 | — | |||
| google/gemma-3-27b-itopenrouter/google/gemma-3-27b-it | 131.072K | $0.08 | $0.45 | — | |||
| z-ai/glm-5.3-flashopenrouter/z-ai/glm-5.3-flash | 1.31072M | $0.075 | $0.25 | — | |||
| google/gemma-3-12b-itopenrouter/google/gemma-3-12b-it | 131.072K | $0.05 | $0.15 | — | |||
| google/gemma-3-4b-itopenrouter/google/gemma-3-4b-it | 131.072K | $0.05 | $0.1 | — | |||
| qwen/qwen3.8-flashopenrouter/qwen/qwen3.8-flash | 1M | $0.15 | $0.47 | — | |||
| nvidia/NVIDIA-Nemotron-3-Nano-30B-A3Bnebius/nvidia/nvidia-nemotron-3-nano-30b-a3b | 262.144K | $0.06 | $0.24 | — | |||
| databricks-gemini-3-8-flashdatabricks/databricks-gemini-3-8-flash | 1.04858M | — | — | — | |||
| openai/o1-proopenrouter/openai/o1-pro | 200K | $150 | $600 | — | |||
| meta-llama/llama-4-scoutopenrouter/meta-llama/llama-4-scout | 1.31072M | $0.1 | $0.3 | — | |||
| openai/gpt-6-astraopenrouter/openai/gpt-6-astra | 1.05M | $10 | $50 | — | |||
| meta-llama/llama-4-maverickopenrouter/meta-llama/llama-4-maverick | 1.04858M | $0.2 | $0.696 | — | |||
| openai/o4-mini-highopenrouter/openai/o4-mini-high | 200K | $1.1 | $4.4 | — | |||
| qwen/qwen3.7-plusopenrouter/qwen/qwen3.7-plus | 1M | $0.32 | $1.28 | — | |||
| nvidia/Llama-3_1-Nemotron-Ultra-253B-v1nebius/nvidia/llama-3_1-nemotron-ultra-253b-v1 | 131.072K | $0.6 | $1.8 | — | |||
| qwen/qwen3-235b-a22bopenrouter/qwen/qwen3-235b-a22b | 131.072K | $0.455 | $1.82 | — | |||
| qwen/qwen3-32bopenrouter/qwen/qwen3-32b | 131.072K | $0.08 | $0.28 | — | |||
| minimax/minimax-m3openrouter/minimax/minimax-m3 | 1.04858M | $0.3 | $1.2 | — | |||
| qwen/qwen3-14bopenrouter/qwen/qwen3-14b | 131.072K | $0.12 | $0.24 | — | |||
| Qwen: Qwen3 235B A22B Instruct 2507qwen/qwen3-235b-a22b-2507 | 262.144K | $0.087 | $0.35 | — | |||
| zai-org/GLM-5.3baseten/zai-org/glm-5.3 | 1.04858M | $1.4 | $4.4 | — | |||
| nvidia/Cosmos3-Super-Reasonernebius/nvidia/cosmos3-super-reasoner | 262.144K | $0.1 | $0.3 | — | |||
| databricks-gemini-3-pro-imagedatabricks/databricks-gemini-3-pro-image | 65.536K | — | — | — | |||
| global.openai.gpt-6-astrabedrock_converse/global.openai.gpt-6-astra | 1.05M | $10 | $50 | — | |||
| cohere-command-aazure_ai/cohere-command-a | 131.072K | $2.5 | $10 | — | |||
| qwen/qwen3-8bopenrouter/qwen/qwen3-8b | 131.072K | $0.117 | $0.455 | — | |||
| qwen/qwen3-30b-a3bopenrouter/qwen/qwen3-30b-a3b | 131.072K | $0.12 | $0.5 | — | |||
| x-ai/grok-build-0.1openrouter/x-ai/grok-build-0.1 | 256K | $1 | $2 | — | |||
| NVIDIA: Nemotron Nano 12B 2 VL (free)nvidia/nemotron-nano-12b-v2-vl:free | 128K | Free | Free | — | |||
| meta-llama/llama-guard-4-12bopenrouter/meta-llama/llama-guard-4-12b | 163.84K | $0.18 | $0.18 | — | |||
| Mistral: Voxtral Small 24B 2507mistralai/voxtral-small-24b-2507 | 32.768K | $0.1 | $0.3 | — | |||
| x-ai/grok-4.6openrouter/x-ai/grok-4.6 | 500K | $2 | $6 | — | |||
| NousResearch/Hermes-4-70Bnebius/nousresearch/hermes-4-70b | 131.072K | $0.13 | $0.4 | — | |||
| Anthropic: Claude Opus 4.5anthropic/claude-opus-4.5 | 200K | $5 | $25 | — | |||
| OpenAI: GPT-4 Turbo Previewopenai/gpt-4-turbo-preview | 128K | $10 | $30 | — | |||
| IBM: Granite 4.0 Microibm-granite/granite-4.0-h-micro | 131K | $0.017 | $0.112 | — | |||
| Qwen: Qwen3 Next 80B A3B Instructqwen/qwen3-next-80b-a3b-instruct | 262.144K | $0.09 | $1.1 | — | |||
| Anthropic: Claude Haiku 4.5anthropic/claude-haiku-4.5 | 200K | $1 | $5 | — | |||
| Qwen: Qwen3 VL 8B Thinkingqwen/qwen3-vl-8b-thinking | 131.072K | $0.18 | $2.1 | — | |||
| Qwen: Qwen3 Coder 30B A3B Instructqwen/qwen3-coder-30b-a3b-instruct | 262.144K | $0.07 | $0.28 | — | |||
| google/gemini-2.5-pro-preview-05-06openrouter/google/gemini-2.5-pro-preview-05-06 | 1.04858M | $1.25 | $10 | — | |||
| Baidu: ERNIE 4.5 VL 424B A47B baidu/ernie-4.5-vl-424b-a47b | 123K | $0.42 | $1.25 | — | |||
| DeepSeek: R1 0528deepseek/deepseek-r1-0528 | 163.84K | $0.5 | $2.15 | — | |||
| Qwen: Qwen3 Coder Flashqwen/qwen3-coder-flash | 1M | $0.195 | $0.975 | — |