2,824 models

No provider description is available for this model yet.

azure/o1 200K context $15/M input $60/M output

No provider description is available for this model yet.

azure/gpt-5.4-nano-2026-03-17 272K context $0.2/M input $1.25/M output

No provider description is available for this model yet.

azure_ai/claude-opus-4-7 1M context $5/M input $25/M output

This model always redirects to the latest model in the GPT Luna family.

~openai/gpt-luna-latest 1.05M context $0.2/M input $1.2/M output

No provider description is available for this model yet.

azure_ai/claude-opus-4-6 1M context $5/M input $25/M output

GPT-6 Astra Pro is the same underlying model as [GPT-6 Astra](https://openrouter.ai/openai/gpt-6-astra), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. Learn more in OpenAI's docs: https://developers.openai.com/api/docs/guides/reasoning#reasoning-mode

openai/gpt-6-astra-pro 1.05M context $10/M input $50/M output

Qwen3-Coder-480B-A35B-Instruct is a Mixture-of-Experts (MoE) code generation model developed by the Qwen team. It is optimized for agentic coding tasks such as function calling, tool use, and long-context reasoning over...

qwen/qwen3-coder:free 262K context Free input Free output

No provider description is available for this model yet.

azure_ai/claude-opus-4-5 200K context $5/M input $25/M output

No provider description is available for this model yet.

azure_ai/claude-haiku-4-5 200K context $1/M input $5/M output

No provider description is available for this model yet.

azure/command-r-plus 128K context $3/M input $15/M output

Mistral's cutting-edge language model for coding released end of July 2025. Codestral specializes in low-latency, high-frequency tasks such as fill-in-the-middle (FIM), code correction and test generation. [Blog Post](https://mistral.ai/news/codestral-25-08)

mistralai/codestral-2508:batch 256K context $0.15/M input $0.45/M output

GPT-6 Astra Pro is the same underlying model as [GPT-6 Astra](https://openrouter.ai/openai/gpt-6-astra), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. Learn more in OpenAI's docs: https://developers.openai.com/api/docs/guides/reasoning#reasoning-mode

openai/gpt-6-astra-pro:batch 1.05M context $5/M input $25/M output

Mistral's cutting-edge language model for coding released end of July 2025. Codestral specializes in low-latency, high-frequency tasks such as fill-in-the-middle (FIM), code correction and test generation. [Blog Post](https://mistral.ai/news/codestral-25-08)

mistralai/codestral-2508 256K context $0.3/M input $0.9/M output

No provider description is available for this model yet.

cerebras/llama-3.3-70b 128K context $0.85/M input $1.2/M output

No provider description is available for this model yet.

azure/gpt-5.5-2026-04-23 1.05M context $5/M input $30/M output

No provider description is available for this model yet.

azure/eu/gpt-5.5 1.05M context $5.5/M input $33/M output

Gemma 4 31B Instruct is Google DeepMind's 30.7B dense multimodal model supporting text and image input with text output. Features a 256K token context window, configurable thinking/reasoning mode, native function...

google/gemma-4-31b-it:batch 262.144K context $0.39/M input $0.97/M output

No provider description is available for this model yet.

scx-ai/qwen3.8-max 1M context $1.65/M input $4.99/M output

No provider description is available for this model yet.

azure/us/gpt-5.5 1.05M context $5.5/M input $33/M output

No provider description is available for this model yet.

azure/gpt-5.5 1.05M context $5/M input $30/M output

Nex-N2.5 is an agentic model built to turn goals into working, verified outcomes. Its core strength is agentic coding within a visual feedback loop: it can explore codebases, implement multi-file...

nex-agi/nex-n2.5-mini:free 262.144K context Free input Free output

No provider description is available for this model yet.

azure/us/gpt-5.5-2026-04-23 1.05M context $5.5/M input $33/M output

No provider description is available for this model yet.

azure/eu/gpt-5.5-2026-04-23 1.05M context $5.5/M input $33/M output

Mercury 2.5 is the fastest reasoning LLM, and the latest diffusion LLM (dLLM) from Inception. Instead of generating tokens sequentially, Mercury 2.5 produces and refines multiple tokens in parallel, achieving...

inception/mercury-2.5-preview 260K context $0.2/M input $0.75/M output

No provider description is available for this model yet.

azure/gpt-5.4-mini 272K context $0.75/M input $4.5/M output

No provider description is available for this model yet.

cerebras/llama3.1-70b 128K context $0.6/M input $0.6/M output

No provider description is available for this model yet.

azure/gpt-5.4-mini-2026-03-17 272K context $0.75/M input $4.5/M output

DeepSeek V4 Pro 0813 is a large-scale mixture-of-experts model from DeepSeek. This is the GA release of DeepSeek V4 Pro.

deepseek/deepseek-v4-pro-0813:batch 1.04858M context $0.66/M input $1.98/M output

Qwen3.5-9B is a multimodal foundation model from the Qwen3.5 family, designed to deliver strong reasoning, coding, and visual understanding in an efficient 9B-parameter architecture. It uses a unified vision-language design...

qwen/qwen3.5-9b:batch 262.144K context $0.17/M input $0.25/M output

No provider description is available for this model yet.

azure/gpt-5.4-nano 272K context $0.2/M input $1.25/M output

Inkling Small is an open-weight multimodal mixture-of-experts model from Thinking Machines Lab, with 12B active parameters out of 276B total. It is positioned as the smaller, more efficient member of...

thinkingmachines/inkling-small:batch 524.288K context $0.5/M input $1.2/M output

No provider description is available for this model yet.

azure/eu/gpt-5.6-luna 1.05M context $1.1/M input $6.6/M output

Sonnet 4.6 is Anthropic's most capable Sonnet-class model yet, with frontier performance across coding, agents, and professional work. It excels at iterative development, complex codebase navigation, end-to-end project management with...

anthropic/claude-sonnet-4.6:batch 1M context $1.5/M input $7.5/M output

No provider description is available for this model yet.

azure/eu/gpt-5.6-terra 1.05M context $2.75/M input $16.5/M output