3,462 models

No provider description is available for this model yet.

openrouter/qwen/qwen3-32b 131.072K context $0.08/M input $0.28/M output

No provider description is available for this model yet.

openrouter/qwen/qwen3-235b-a22b 131.072K context $0.455/M input $1.82/M output

No provider description is available for this model yet.

openrouter/openai/o4-mini-high 200K context $1.1/M input $4.4/M output

No provider description is available for this model yet.

openrouter/openai/o1-pro 200K context $150/M input $600/M output

No provider description is available for this model yet.

openrouter/google/gemma-3-4b-it 131.072K context $0.05/M input $0.1/M output

No provider description is available for this model yet.

openrouter/google/gemma-3-27b-it 131.072K context $0.08/M input $0.45/M output

No provider description is available for this model yet.

openrouter/mistralai/mistral-saba 32.768K context $0.2/M input $0.6/M output

No provider description is available for this model yet.

openrouter/qwen/qwen-plus 1M context $0.26/M input $0.78/M output

No provider description is available for this model yet.

openrouter/minimax/minimax-01 1.00019M context $0.2/M input $1.1/M output

This model always redirects to the latest model in the DeepSeek Pro family.

~deepseek/deepseek-pro-latest 1.04858M context $0.579/M input $1.738/M output

No provider description is available for this model yet.

openrouter/mistralai/mistral-nemo 131.072K context $0.019/M input $0.03/M output

No provider description is available for this model yet.

openrouter/google/gemma-2-27b-it 8.192K context $0.65/M input $0.65/M output

No provider description is available for this model yet.

openrouter/openai/gpt-4-turbo 128K context $10/M input $30/M output

Gemini 3.6 Flash is a high-efficiency model from Google for coding, agentic workflows, and web and app development. It is designed to produce polished outputs with fewer unnecessary edits and...

google/gemini-3.6-flash:batch 1.04858M context $0.375/M input $1.875/M output

DeepSeek-V3.1 Terminus is an update to [DeepSeek V3.1](/deepseek/deepseek-chat-v3.1) that maintains the model's original capabilities while addressing issues reported by users, including language consistency and agent capabilities, further optimizing the model's...

deepseek/deepseek-v3.1-terminus 131.072K context $0.27/M input $1/M output

GPT-3.5 Turbo is OpenAI's fastest model. It can understand and generate natural language or code, and is optimized for chat and traditional completion tasks. Training data up to Sep 2021.

openai/gpt-3.5-turbo:batch 16.385K context $0.25/M input $0.75/M output

Opus 4.7 is the next generation of Anthropic's Opus family, built for long-running, asynchronous agents. Building on the coding and agentic strengths of Opus 4.6, it delivers stronger performance on...

anthropic/claude-opus-4.7:batch 1M context $2.5/M input $12.5/M output

Qwen3-14B is a dense 14.8B parameter causal language model from the Qwen3 series, designed for both complex reasoning and efficient dialogue. It supports seamless switching between a "thinking" mode for...

qwen/qwen3-14b 40.96K context $0.12/M input $0.24/M output

No provider description is available for this model yet.

deepinfra/microsoft/phi-4 16.384K context $0.07/M input $0.14/M output