No provider description is available for this model yet.

qwencloud/qwen-plus-2025-09-11 997.952K context Input not listed Output not listed

The Auto Router automatically selects the best model for your prompt, powered by the wisdom of the market. It routes you based on what the OpenRouter community collectively spends on...

openrouter/auto 2M context Input not listed Output not listed

Amazon Nova Pro 1.0 is a capable multimodal model from Amazon focused on providing a combination of accuracy, speed, and cost for a wide range of tasks. As of December...

amazon/nova-pro-v1 300K context $0.8/M input $3.2/M output

This is Mistral AI's flagship model, Mistral Large 2 (version mistral-large-2407). It's a proprietary weights-available model and excels at reasoning, code, JSON, chat, and more. Read the launch announcement [here](https://mistral.ai/news/mistral-large-2407/)....

mistralai/mistral-large-2407 131.072K context $2/M input $6/M output

No provider description is available for this model yet.

deepinfra/qwen/qwen3.8-27b 262.144K context $0.4/M input $3/M output

No provider description is available for this model yet.

openrouter/openai/gpt-5.5-pro 1.05M context $30/M input $180/M output

No provider description is available for this model yet.

openrouter/qwen/qwen3.6-27b 262.144K context $0.3/M input $2/M output

No provider description is available for this model yet.

qwencloud/qwen-plus-2025-07-28 997.952K context Input not listed Output not listed

No provider description is available for this model yet.

openrouter/qwen/qwen3.6-max-preview 262.144K context $1.027/M input $6.162/M output

No provider description is available for this model yet.

openrouter/qwen/qwen3.6-35b-a3b 262.144K context $0.1/M input $0.9/M output

No provider description is available for this model yet.

qwencloud/qwen-plus-2025-07-14 129.024K context $0.4/M input $1.2/M output

KAT-Coder-Air V2.5 is a flagship-level Agentic Coding model that can directly hand over an entire issue or an entire business workflow to it, allowing it to autonomously locate and make...

kwaipilot/kat-coder-air-v2.5:free 256K context Free input Free output

No provider description is available for this model yet.

deepinfra/minimaxai/minimax-m2.7 196.608K context $0.25/M input $1/M output

No provider description is available for this model yet.

novita/zai-org/glm-5.1 204.8K context $1.38/M input $4.4/M output

No provider description is available for this model yet.

openrouter/qwen/qwen3.6-flash 1M context $0.188/M input $1.125/M output

Hy3 is a 295B-parameter Mixture-of-Experts model from Tencent (21B active, 192 experts with top-8 routing) built for reasoning, agentic workflows, and real-world production use. It supports a configurable reasoning effort:...

tencent/hy3:free 262.144K context Free input Free output

No provider description is available for this model yet.

qwencloud/qwen-plus-2025-04-28 129.024K context $0.4/M input $1.2/M output

GPT-5.6 Luna Pro is the same underlying model as [GPT-5.6 Luna](https://openrouter.ai/openai/gpt-5.6-luna), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. Learn more in OpenAI's docs: https://developers.openai.com/api/docs/guides/reasoning#reasoning-mode

openai/gpt-5.6-luna-pro:batch 1.05M context $0.1/M input $0.6/M output

Claude Opus 4.8 is Anthropic's most capable generally available model in the Opus family. It supports text, image, and file inputs with text output, with reasoning support and a 1M-token...

anthropic/claude-opus-4.8 1M context $5/M input $25/M output

Mistral Small 4 is the next major release in the Mistral Small family, unifying the capabilities of several flagship Mistral models into a single system. It combines strong reasoning from...

mistralai/mistral-small-2603:batch 262.144K context $0.075/M input $0.3/M output

As a 30B-class SOTA model, GLM-4.7-Flash offers a new option that balances performance and efficiency. It is further optimized for agentic coding use cases, strengthening coding capabilities, long-horizon task planning,...

z-ai/glm-4.7-flash 131.072K context $0.061/M input $0.4/M output

No provider description is available for this model yet.

qwencloud/qwen-plus-2025-01-25 129.024K context $0.4/M input $1.2/M output

No provider description is available for this model yet.

deepinfra/openai/gpt-oss-120b-turbo 131.072K context $0.15/M input $0.6/M output

No provider description is available for this model yet.

openrouter/qwen/qwen3.7-max 1M context $1.475/M input $4.425/M output

Uncensored and creative writing model based on Mistral Small 3.2 24B with good recall, prompt adherence, and intelligence.

thedrummer/cydonia-24b-v4.1 131.072K context $0.3/M input $0.5/M output

OpenAI o4-mini-high is the same model as [o4-mini](/openai/o4-mini) with reasoning_effort set to high. OpenAI o4-mini is a compact reasoning model in the o-series, optimized for fast, cost-efficient performance while retaining...

openai/o4-mini-high 200K context $1.1/M input $4.4/M output

No provider description is available for this model yet.

qwencloud/qwen-plus 129.024K context $0.4/M input $1.2/M output

Mercury 2.5 is the fastest reasoning LLM, and the latest diffusion LLM (dLLM) from Inception. Instead of generating tokens sequentially, Mercury 2.5 produces and refines multiple tokens in parallel, achieving...

inception/mercury-2.5 260K context $0.04/M input $0.15/M output

Grok 4.3 is a reasoning model from SpaceXAI. It accepts text and image inputs with text output, and is suited for agentic workflows, instruction-following tasks, and applications requiring high factual...

x-ai/grok-4.3:batch 1M context $1/M input $2/M output