1,445 models

No provider description is available for this model yet.

azure_ai/gpt-5.4-mini-2026-03-17 272K context $0.75/M input $4.5/M output

No provider description is available for this model yet.

azure_ai/gpt-5.4-nano 272K context $0.2/M input $1.25/M output

No provider description is available for this model yet.

azure_ai/gpt-5.4-nano-2026-03-17 272K context $0.2/M input $1.25/M output

Inkling Small is an open-weight multimodal mixture-of-experts model from Thinking Machines Lab, with 12B active parameters out of 276B total. It is positioned as the smaller, more efficient member of...

thinkingmachines/inkling-small:free 1.04858M context Free input Free output

No provider description is available for this model yet.

azure/eu/gpt-5-2025-08-07 272K context $1.375/M input $11/M output

No provider description is available for this model yet.

azure/eu/gpt-5-mini-2025-08-07 272K context $0.275/M input $2.2/M output

No provider description is available for this model yet.

azure/eu/gpt-5.1 272K context $1.38/M input $11/M output

No provider description is available for this model yet.

azure/eu/gpt-5.1-chat 128K context $1.38/M input $11/M output

No provider description is available for this model yet.

azure/eu/gpt-5-nano-2025-08-07 272K context $0.055/M input $0.44/M output

Qwen3.5 Plus (April 2026) is a large-scale multimodal language model from Alibaba. It accepts text, image, and video input and produces text output, with a 1M token context window. This...

qwen/qwen3.5-plus-20260420 1M context $0.3/M input $1.8/M output

No provider description is available for this model yet.

azure/us-gov/gpt-5.1 272K context $1.719/M input $13.75/M output

GPT-5 Chat is designed for advanced, natural, multimodal, and context-aware conversations for enterprise applications.

openai/gpt-5-chat 128K context $1.25/M input $10/M output

Qwen 3.6 Plus builds on a hybrid architecture that combines efficient linear attention with sparse mixture-of-experts routing, enabling strong scalability and high-performance inference. Compared to the 3.5 series, it delivers...

qwen/qwen3.6-plus 1M context $0.325/M input $1.95/M output

No provider description is available for this model yet.

azure/global/gpt-4o-2024-08-06 128K context $2.5/M input $10/M output

Amazon Nova Pro 1.0 is a capable multimodal model from Amazon focused on providing a combination of accuracy, speed, and cost for a wide range of tasks. As of December...

amazon/nova-pro-v1 300K context $0.8/M input $3.2/M output

No provider description is available for this model yet.

azure/global/gpt-5.1 272K context $1.25/M input $10/M output

Amazon Nova Lite 1.0 is a very low-cost multimodal model from Amazon that focused on fast processing of image, video, and text inputs to generate text output. Amazon Nova Lite...

amazon/nova-lite-v1 300K context $0.06/M input $0.24/M output

Mistral Small 4 is the next major release in the Mistral Small family, unifying the capabilities of several flagship Mistral models into a single system. It combines strong reasoning from...

mistralai/mistral-small-2603 262.144K context $0.15/M input $0.6/M output

Claude Opus 4 is benchmarked as the world’s best coding model, at time of release, bringing sustained performance on complex, long-running tasks and agent workflows. It sets new benchmarks in...

anthropic/claude-opus-4 200K context $15/M input $75/M output

GPT-5.6 Sol Pro is the same underlying model as [GPT-5.6 Sol](https://openrouter.ai/openai/gpt-5.6-sol), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. Learn more in OpenAI's docs: https://developers.openai.com/api/docs/guides/reasoning#reasoning-mode

openai/gpt-5.6-sol-pro:batch 1.05M context $1/M input $5/M output

No provider description is available for this model yet.

azure/gpt-4-turbo-2024-04-09 128K context $10/M input $30/M output

The Auto Router automatically selects the best model for your prompt, powered by the wisdom of the market. It routes you based on what the OpenRouter community collectively spends on...

openrouter/auto 2M context Input not listed Output not listed

No provider description is available for this model yet.

azure/gpt-4-turbo-vision-preview 128K context $10/M input $30/M output

No provider description is available for this model yet.

azure/gpt-4.1 1.04758M context $2/M input $8/M output

No provider description is available for this model yet.

azure/gpt-4.1-mini 1.04758M context $0.4/M input $1.6/M output

No provider description is available for this model yet.

azure/gpt-4.1-mini-2025-04-14 1.04758M context $0.4/M input $1.6/M output

Llama 4 Scout 17B Instruct (16E) is a mixture-of-experts (MoE) language model developed by Meta, activating 17 billion parameters out of a total of 109B. It supports native multimodal input...

meta-llama/llama-4-scout 327.68K context $0.1/M input $0.3/M output