3,457 models

No provider description is available for this model yet.

azure_ai/deepseek-v3-0324 128K context $1.14/M input $4.56/M output

No provider description is available for this model yet.

openai/chatgpt-4o-latest 128K context $5/M input $15/M output

The simplest way to get free inference. openrouter/free is a router that selects free models at random from the models available on OpenRouter. The router smartly filters for models that...

openrouter/free 200K context Free input Free output

Relace Apply 3 is a specialized code-patching LLM that merges AI-suggested edits straight into your source files. It can apply updates from GPT-4o, Claude, and others into your files at...

relace/relace-apply-3 256K context $0.85/M input $1.25/M output

No provider description is available for this model yet.

azure_ai/deepseek-v3 128K context $1.14/M input $4.56/M output

GLM-5.1 delivers a major leap in coding capability, with particularly significant gains in handling long-horizon tasks. Unlike previous models built around minute-level interactions, GLM-5.1 can work independently and continuously on...

z-ai/glm-5.1 200K context $0.966/M input $3.036/M output

GPT-5.4 Pro is OpenAI's most advanced model, building on GPT-5.4's unified architecture with enhanced reasoning capabilities for complex, high-stakes tasks. It features a 1M+ token context window (922K input, 128K...

openai/gpt-5.4-pro:batch 1.05M context $15/M input $90/M output

No provider description is available for this model yet.

azure_ai/deepseek-r1 128K context $1.35/M input $5.4/M output

No provider description is available for this model yet.

azure_ai/deepseek-v3.2-speciale 163.84K context $0.58/M input $1.68/M output

No provider description is available for this model yet.

azure_ai/deepseek-v3.2 163.84K context $0.58/M input $1.68/M output

Coder‑Large is a 32 B‑parameter offspring of Qwen 2.5‑Instruct that has been further trained on permissively‑licensed GitHub, CodeSearchNet and synthetic bug‑fix corpora. It supports a 32k context window, enabling multi‑file...

arcee-ai/coder-large 32.768K context $0.5/M input $0.8/M output

GPT-5.6 Terra is a balanced model in OpenAI's GPT-5.6 series, positioned between the flagship Sol tier and the cost-efficient Luna tier. It is suited for everyday coding, reasoning, and agentic...

openai/gpt-5.6-terra:batch 1.05M context $1/M input $6/M output

NVIDIA Nemotron 3 Ultra is an open frontier-reasoning and orchestration model from NVIDIA, with 55B active parameters out of 550B total (MoE). Built on a hybrid Transformer-Mamba mixture-of-experts architecture, it...

nvidia/nemotron-3-ultra-550b-a55b:free 1M context Free input Free output

No provider description is available for this model yet.

vercel_ai_gateway/zai/glm-4.6 200K context $0.45/M input $1.8/M output

No provider description is available for this model yet.

vercel_ai_gateway/zai/glm-4.5-air 128K context $0.2/M input $1.1/M output

No provider description is available for this model yet.

vercel_ai_gateway/zai/glm-4.5 131.072K context $0.6/M input $2.2/M output

For tasks that demand low latency, GPT‑4.1 nano is the fastest and cheapest model in the GPT-4.1 series. It delivers exceptional performance at a small size with its 1 million...

openai/gpt-4.1-nano:batch 1.04758M context $0.05/M input $0.2/M output

No provider description is available for this model yet.

azure_ai/mai-ds-r1 128K context $1.35/M input $5.4/M output

Qwen3-Coder-Next is an open-weight causal language model optimized for coding agents and local development workflows. It uses a sparse MoE design with 80B total parameters and only 3B activated per...

qwen/qwen3-coder-next 262.144K context $0.12/M input $0.8/M output

Gemma 3n E4B-it is optimized for efficient execution on mobile and low-resource devices, such as phones, laptops, and tablets. It supports multimodal inputs—including text, visual data, and audio—enabling diverse tasks...

google/gemma-3n-e4b-it 32.768K context $0.06/M input $0.12/M output

No provider description is available for this model yet.

vercel_ai_gateway/xai/grok-4 256K context $3/M input $15/M output

No provider description is available for this model yet.

azure_ai/model-router 200K context $0.14/M input Output not listed

No provider description is available for this model yet.

azure_ai/phi-4-reasoning 32.768K context $0.125/M input $0.5/M output

No provider description is available for this model yet.

vercel_ai_gateway/xai/grok-3-mini 131.072K context $0.3/M input $0.5/M output

No provider description is available for this model yet.

vertex_ai/xai/grok-4.3 200K context $1.25/M input $2.5/M output

No provider description is available for this model yet.

azure_ai/grok-4-20-reasoning 262K context $1.25/M input $2.5/M output

No provider description is available for this model yet.

vertex_ai/xai/grok-4.6 524.288K context $2/M input $6/M output

No provider description is available for this model yet.

azure_ai/grok-4-20-non-reasoning 262K context $1.25/M input $2.5/M output

Qwen3 Coder Plus is Alibaba's proprietary version of the Open Source Qwen3 Coder 480B A35B. It is a powerful coding agent model specializing in autonomous programming via tool calling and...

qwen/qwen3-coder-plus 1M context $0.65/M input $3.25/M output

No provider description is available for this model yet.

gemini/lyria-3.5 1.04858M context Input not listed Output not listed

No provider description is available for this model yet.

vercel_ai_gateway/xai/grok-3-fast 131.072K context $5/M input $25/M output

No provider description is available for this model yet.

azure_ai/phi-4-mini-reasoning 131.072K context $0.08/M input $0.32/M output