3,199 models

No provider description is available for this model yet.

openrouter/qwen/qwen-plus 1M context $0.26/M input $0.78/M output

No provider description is available for this model yet.

openrouter/z-ai/glm-5.3 1.31072M context $1.4/M input $4.4/M output

No provider description is available for this model yet.

openrouter/mistralai/mistral-saba 32.768K context $0.2/M input $0.6/M output

No provider description is available for this model yet.

nebius/nvidia/nemotron-3-nano-omni 262.144K context $0.06/M input $0.24/M output

No provider description is available for this model yet.

openrouter/google/gemma-3-27b-it 131.072K context $0.08/M input $0.45/M output

No provider description is available for this model yet.

openrouter/z-ai/glm-5.3-flash 1.31072M context $0.075/M input $0.25/M output

No provider description is available for this model yet.

openrouter/google/gemma-3-12b-it 131.072K context $0.05/M input $0.15/M output

No provider description is available for this model yet.

openrouter/google/gemma-3-4b-it 131.072K context $0.05/M input $0.1/M output

No provider description is available for this model yet.

openrouter/qwen/qwen3.8-flash 1M context $0.15/M input $0.47/M output

No provider description is available for this model yet.

openrouter/openai/o1-pro 200K context $150/M input $600/M output

No provider description is available for this model yet.

openrouter/openai/gpt-6-astra 1.05M context $10/M input $50/M output

No provider description is available for this model yet.

openrouter/openai/o4-mini-high 200K context $1.1/M input $4.4/M output

No provider description is available for this model yet.

openrouter/qwen/qwen3.7-plus 1M context $0.32/M input $1.28/M output

No provider description is available for this model yet.

openrouter/qwen/qwen3-235b-a22b 131.072K context $0.455/M input $1.82/M output

No provider description is available for this model yet.

openrouter/qwen/qwen3-32b 131.072K context $0.08/M input $0.28/M output

No provider description is available for this model yet.

openrouter/minimax/minimax-m3 1.04858M context $0.3/M input $1.2/M output

No provider description is available for this model yet.

openrouter/qwen/qwen3-14b 131.072K context $0.12/M input $0.24/M output

Qwen3-235B-A22B-Instruct-2507 is a multilingual, instruction-tuned mixture-of-experts language model based on the Qwen3-235B architecture, with 22B active parameters per forward pass. It is optimized for general-purpose text generation, including instruction following,...

qwen/qwen3-235b-a22b-2507 262.144K context $0.087/M input $0.35/M output

No provider description is available for this model yet.

baseten/zai-org/glm-5.3 1.04858M context $1.4/M input $4.4/M output

No provider description is available for this model yet.

azure_ai/cohere-command-a 131.072K context $2.5/M input $10/M output

No provider description is available for this model yet.

openrouter/qwen/qwen3-8b 131.072K context $0.117/M input $0.455/M output

No provider description is available for this model yet.

openrouter/qwen/qwen3-30b-a3b 131.072K context $0.12/M input $0.5/M output

No provider description is available for this model yet.

openrouter/x-ai/grok-build-0.1 256K context $1/M input $2/M output

NVIDIA Nemotron Nano 2 VL is a 12-billion-parameter open multimodal reasoning model designed for video understanding and document intelligence. It introduces a hybrid Transformer-Mamba architecture, combining transformer-level accuracy with Mamba’s...

nvidia/nemotron-nano-12b-v2-vl:free 128K context Free input Free output

Voxtral Small is an enhancement of Mistral Small 3, incorporating state-of-the-art audio input capabilities while retaining best-in-class text performance. It excels at speech transcription, translation and audio understanding. Input audio...

mistralai/voxtral-small-24b-2507 32.768K context $0.1/M input $0.3/M output

No provider description is available for this model yet.

openrouter/x-ai/grok-4.6 500K context $2/M input $6/M output

No provider description is available for this model yet.

nebius/nousresearch/hermes-4-70b 131.072K context $0.13/M input $0.4/M output

Claude Opus 4.5 is Anthropic’s frontier reasoning model optimized for complex software engineering, agentic workflows, and long-horizon computer use. It offers strong multimodal capabilities, competitive performance across real-world coding and...

anthropic/claude-opus-4.5 200K context $5/M input $25/M output

The preview GPT-4 model with improved instruction following, JSON mode, reproducible outputs, parallel function calling, and more. Training data: up to Dec 2023. **Note:** heavily rate limited by OpenAI while...

openai/gpt-4-turbo-preview 128K context $10/M input $30/M output

Granite-4.0-H-Micro is a 3B parameter from the Granite 4 family of models. These models are the latest in a series of models released by IBM. They are fine-tuned for long...

ibm-granite/granite-4.0-h-micro 131K context $0.017/M input $0.112/M output

Qwen3-Next-80B-A3B-Instruct is an instruction-tuned chat model in the Qwen3-Next series optimized for fast, stable responses without “thinking” traces. It targets complex tasks across reasoning, code generation, knowledge QA, and multilingual...

qwen/qwen3-next-80b-a3b-instruct 262.144K context $0.09/M input $1.1/M output

Claude Haiku 4.5 is Anthropic’s fastest and most efficient model, delivering near-frontier intelligence at a fraction of the cost and latency of larger Claude models. Matching Claude Sonnet 4’s performance...

anthropic/claude-haiku-4.5 200K context $1/M input $5/M output

Qwen3-VL-8B-Thinking is the reasoning-optimized variant of the Qwen3-VL-8B multimodal model, designed for advanced visual and textual reasoning across complex scenes, documents, and temporal sequences. It integrates enhanced multimodal alignment and...

qwen/qwen3-vl-8b-thinking 131.072K context $0.18/M input $2.1/M output

Qwen3-Coder-30B-A3B-Instruct is a 30.5B parameter Mixture-of-Experts (MoE) model with 128 experts (8 active per forward pass), designed for advanced code generation, repository-scale understanding, and agentic tool use. Built on the...

qwen/qwen3-coder-30b-a3b-instruct 262.144K context $0.07/M input $0.28/M output

ERNIE-4.5-VL-424B-A47B is a multimodal Mixture-of-Experts (MoE) model from Baidu’s ERNIE 4.5 series, featuring 424B total parameters with 47B active per token. It is trained jointly on text and image data...

baidu/ernie-4.5-vl-424b-a47b 123K context $0.42/M input $1.25/M output

May 28th update to the [original DeepSeek R1](/deepseek/deepseek-r1) Performance on par with [OpenAI o1](/openai/o1), but open-sourced and with fully open reasoning tokens. It's 671B parameters in size, with 37B active...

deepseek/deepseek-r1-0528 163.84K context $0.5/M input $2.15/M output

Qwen3 Coder Flash is Alibaba's fast and cost efficient version of their proprietary Qwen3 Coder Plus. It is a powerful coding agent model specializing in autonomous programming via tool calling...

qwen/qwen3-coder-flash 1M context $0.195/M input $0.975/M output