1,424 models

No provider description is available for this model yet.

openrouter/minimax/minimax-m3 1.04858M context $0.3/M input $1.2/M output

No provider description is available for this model yet.

moonshot/kimi-k3 1.04858M context $3/M input $15/M output

No provider description is available for this model yet.

baseten/zai-org/glm-5.3 1.04858M context $1.4/M input $4.4/M output

No provider description is available for this model yet.

openrouter/x-ai/grok-build-0.1 256K context $1/M input $2/M output

No provider description is available for this model yet.

openrouter/x-ai/grok-4.6 500K context $2/M input $6/M output

No provider description is available for this model yet.

openrouter/x-ai/grok-4.5 500K context $2/M input $6/M output

No provider description is available for this model yet.

openrouter/x-ai/grok-4.3 1M context $1.25/M input $2.5/M output

No provider description is available for this model yet.

openrouter/x-ai/grok-4.20 1M context $1.25/M input $2.5/M output

No provider description is available for this model yet.

openrouter/openai/o4-mini 200K context $1.1/M input $4.4/M output

Mistral Small 4 is the next major release in the Mistral Small family, unifying the capabilities of several flagship Mistral models into a single system. It combines strong reasoning from...

mistralai/mistral-small-2603:batch 262.144K context $0.075/M input $0.3/M output

No provider description is available for this model yet.

deepinfra/moonshotai/kimi-k2.7-code 262.144K context $0.68/M input $3.4/M output

OpenAI o4-mini-high is the same model as [o4-mini](/openai/o4-mini) with reasoning_effort set to high. OpenAI o4-mini is a compact reasoning model in the o-series, optimized for fast, cost-efficient performance while retaining...

openai/o4-mini-high:batch 200K context $0.55/M input $2.2/M output

No provider description is available for this model yet.

openrouter/openai/o3 200K context $2/M input $8/M output

No provider description is available for this model yet.

openrouter/openai/gpt-5.6-terra 922K context $2/M input $12/M output

No provider description is available for this model yet.

openrouter/openai/gpt-5.6-luna 922K context $0.2/M input $1.2/M output

No provider description is available for this model yet.

openrouter/openai/gpt-5.5 1.05M context $5/M input $30/M output

Qwen 3.6 Plus builds on a hybrid architecture that combines efficient linear attention with sparse mixture-of-experts routing, enabling strong scalability and high-performance inference. Compared to the 3.5 series, it delivers...

qwen/qwen3.6-plus 1M context $0.325/M input $1.95/M output

No provider description is available for this model yet.

openrouter/openai/gpt-5.4-nano 272K context $0.2/M input $1.25/M output

No provider description is available for this model yet.

openrouter/openai/gpt-5.4-mini 272K context $0.75/M input $4.5/M output

No provider description is available for this model yet.

openrouter/openai/gpt-5.4 1.05M context $2.5/M input $15/M output

No provider description is available for this model yet.

openrouter/openai/gpt-5.3-codex 272K context $1.75/M input $14/M output

No provider description is available for this model yet.

deepinfra/google/gemma-4-31b-it 262.144K context $0.13/M input $0.38/M output

No provider description is available for this model yet.

openrouter/openai/gpt-5.1 272K context $1.25/M input $10/M output

No provider description is available for this model yet.

openrouter/openai/gpt-4o-mini 128K context $0.15/M input $0.6/M output

No provider description is available for this model yet.

openrouter/google/gemini-3.8-flash 1.04858M context $0.75/M input $3.75/M output

No provider description is available for this model yet.

openrouter/google/gemini-3.7-flash 1.04858M context $0.75/M input $3.75/M output

MiniMax-M3 is a multimodal foundation model from MiniMax. It supports text, image, and video inputs with text output, a 1M-token context window, and is suited for long-horizon agentic work, coding,...

minimax/minimax-m3:free 1.04858M context Free input Free output

No provider description is available for this model yet.

azure_ai/claude-fable-5-1 1M context $10/M input $50/M output

No provider description is available for this model yet.

openrouter/google/gemini-3.6-flash 1.04858M context $0.75/M input $3.75/M output

No provider description is available for this model yet.

cerebras/gemma-4-31b 131.072K context $0.99/M input $1.49/M output

No provider description is available for this model yet.

openrouter/google/gemini-3.5-flash 1.04858M context $1.5/M input $9/M output

No provider description is available for this model yet.

deepinfra/qwen/qwen3.5-9b 262.144K context $0.1/M input $0.15/M output

No provider description is available for this model yet.

deepinfra/qwen/qwen3.5-397b-a17b 262.144K context $0.45/M input $3/M output

NVIDIA Nemotron™ 3 Nano Omni is a 30B-A3B open multimodal model designed to function as a perception and context sub-agent in enterprise agent systems. It accepts text, image, video, and...

nvidia/nemotron-3-nano-omni-30b-a3b-reasoning:free 256K context Free input Free output

Llama Guard 4 is a Llama 4 Scout-derived multimodal pretrained model, fine-tuned for content safety classification. Similar to previous versions, it can be used to classify content in both LLM...

meta-llama/llama-guard-4-12b 163.84K context $0.18/M input $0.18/M output

No provider description is available for this model yet.

friendliai/zai-org/glm-5.3-flash 1.04858M context $0.15/M input $0.5/M output

This model always redirects to the latest model in the GPT Mini family.

~openai/gpt-mini-latest 400K context $0.75/M input $4.5/M output

This model always redirects to the latest model in the Gemini Pro family.

~google/gemini-pro-latest 1.04858M context $2/M input $12/M output