4,182 models

DeepSeek V3, a 685B-parameter, mixture-of-experts model, is the latest iteration of the flagship chat model family from the DeepSeek team. It succeeds the [DeepSeek V3](/deepseek/deepseek-chat-v3) model and performs really well...

deepseek/deepseek-chat-v3-0324 163.84K context $0.25/M input $1/M output

OpenAI o3-mini-high is the same model as [o3-mini](/openai/o3-mini) with reasoning_effort set to high. o3-mini is a cost-efficient language model optimized for STEM reasoning tasks, particularly excelling in science, mathematics, and...

openai/o3-mini-high 200K context $1.1/M input $4.4/M output

No provider description is available for this model yet.

bedrock_converse/us.xai.grok-4.6 500K context $2.2/M input $6.6/M output

Llama 4 Scout 17B Instruct (16E) is a mixture-of-experts (MoE) language model developed by Meta, activating 17 billion parameters out of a total of 109B. It supports native multimodal input...

meta-llama/llama-4-scout 327.68K context $0.1/M input $0.3/M output

Qwen3-32B is a dense 32.8B parameter causal language model from the Qwen3 series, optimized for both complex reasoning and efficient dialogue. It supports seamless switching between a "thinking" mode for...

qwen/qwen3-32b 40.96K context $0.08/M input $0.28/M output

Mistral Saba is a 24B-parameter language model specifically designed for the Middle East and South Asia, delivering accurate and contextually relevant responses while maintaining efficient performance. Trained on curated regional...

mistralai/mistral-saba 32.768K context $0.2/M input $0.6/M output

Ling-2.6-flash is an instant (instruct) model from inclusionAI with 104B total parameters and 7.4B active parameters, designed for real-world agents that require fast responses, strong execution, and high token efficiency....

inclusionai/ling-2.6-flash 262.144K context $0.01/M input $0.03/M output

Qwen2.5 72B is the latest series of Qwen large language models. Qwen2.5 brings the following improvements upon Qwen2: - Significantly more knowledge and has greatly improved capabilities in coding and...

qwen/qwen-2.5-72b-instruct 32.768K context $0.36/M input $0.4/M output

GPT-5.6 Luna Pro is the same underlying model as [GPT-5.6 Luna](https://openrouter.ai/openai/gpt-5.6-luna), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. Learn more in OpenAI's docs: https://developers.openai.com/api/docs/guides/reasoning#reasoning-mode

openai/gpt-5.6-luna-pro 1.05M context $0.2/M input $1.2/M output

No provider description is available for this model yet.

ollama/llama3.1 8.192K context Input not listed Output not listed

Qwen3-8B is a dense 8.2B parameter causal language model from the Qwen3 series, designed for both reasoning-heavy tasks and efficient dialogue. It supports seamless switching between "thinking" mode for math,...

qwen/qwen3-8b 131.072K context $0.117/M input $0.455/M output

No provider description is available for this model yet.

bedrock_mantle/xai.grok-4.6 500K context $2.2/M input $6.6/M output

Qwen3, the latest generation in the Qwen large language model series, features both dense and mixture-of-experts (MoE) architectures to excel in reasoning, multilingual support, and advanced agent tasks. Its unique...

qwen/qwen3-30b-a3b 40.96K context $0.12/M input $0.5/M output

Llama 4 Maverick 17B Instruct (128E) is a high-capacity multimodal language model from Meta, built on a mixture-of-experts (MoE) architecture with 128 experts and 17 billion active parameters per forward...

meta-llama/llama-4-maverick 128K context $0.2/M input $0.696/M output

May 28th update to the [original DeepSeek R1](/deepseek/deepseek-r1) Performance on par with [OpenAI o1](/openai/o1), but open-sourced and with fully open reasoning tokens. It's 671B parameters in size, with 37B active...

deepseek/deepseek-r1-0528 163.84K context $0.5/M input $2.15/M output

No provider description is available for this model yet.

openrouter/deepseek/deepseek-v4-pro 1.04858M context $1.32/M input $3.96/M output

ERNIE-4.5-VL-424B-A47B is a multimodal Mixture-of-Experts (MoE) model from Baidu’s ERNIE 4.5 series, featuring 424B total parameters with 47B active per token. It is trained jointly on text and image data...

baidu/ernie-4.5-vl-424b-a47b 123K context $0.42/M input $1.25/M output

Skyfall 36B v2 is an enhanced iteration of Mistral Small 2501, specifically fine-tuned for improved creativity, nuanced writing, role-playing, and coherent storytelling.

thedrummer/skyfall-36b-v2 32.768K context $0.55/M input $0.8/M output

Qwen3.6 27B is a dense 27-billion-parameter language model from the Qwen Team at Alibaba, released in April 2026. It features hybrid multimodal capabilities — accepting text, image, and video inputs...

qwen/qwen3.6-27b 262.144K context $0.3/M input $2/M output

No provider description is available for this model yet.

mistral/glm-5-2 1.04858M context $1.4/M input $4.4/M output

No provider description is available for this model yet.

mistral/zai-glm-5-2 1.04858M context $1.4/M input $4.4/M output

Qwen3-235B-A22B-Instruct-2507 is a multilingual, instruction-tuned mixture-of-experts language model based on the Qwen3-235B architecture, with 22B active parameters per forward pass. It is optimized for general-purpose text generation, including instruction following,...

qwen/qwen3-235b-a22b-2507 262.144K context $0.087/M input $0.35/M output

GPT-6 Astra Pro is the same underlying model as [GPT-6 Astra](https://openrouter.ai/openai/gpt-6-astra), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. Learn more in OpenAI's docs: https://developers.openai.com/api/docs/guides/reasoning#reasoning-mode

openai/gpt-6-astra-pro 1.05M context $10/M input $50/M output

No provider description is available for this model yet.

baseten/deepseek-ai/deepseek-v3.1 Not documented context $0.5/M input $1.5/M output

No provider description is available for this model yet.

baseten/deepseek-ai/deepseek-v3-0324 Not documented context $0.77/M input $0.77/M output

A cost-efficient version of GPT Audio. The new snapshot features an upgraded decoder for more natural sounding voices and maintains better voice consistency. Input is priced at $0.60 per million...

openai/gpt-audio-mini 128K context $0.6/M input $2.4/M output