1,564 models

No provider description is available for this model yet.

snowflake/claude-4-opus 200K context $5/M input $25/M output

No provider description is available for this model yet.

snowflake/claude-haiku-4-5 200K context $1/M input $5/M output

No provider description is available for this model yet.

snowflake/claude-3-7-sonnet 200K context $3/M input $15/M output

No provider description is available for this model yet.

snowflake/openai-gpt-4.1 300K context $2/M input $8/M output

No provider description is available for this model yet.

snowflake/openai-gpt-5 300K context $1.25/M input $10/M output

Sonnet 5 is Anthropic's most capable Sonnet-class model, with frontier performance across coding, agents, and professional work. It supports adaptive thinking with selectable reasoning effort levels (low, medium, high, max,...

anthropic/claude-sonnet-5:batch 1M context $1/M input $5/M output

GPT-5.5 is OpenAI’s frontier model designed for complex professional workloads, building on GPT-5.4 with stronger reasoning, higher reliability, and improved token efficiency on hard tasks. It features a 1M+ token...

openai/gpt-5.5:batch 1.05M context $2.5/M input $15/M output

GPT-5.6 Luna Pro is the same underlying model as [GPT-5.6 Luna](https://openrouter.ai/openai/gpt-5.6-luna), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. Learn more in OpenAI's docs: https://developers.openai.com/api/docs/guides/reasoning#reasoning-mode

openai/gpt-5.6-luna-pro:batch 1.05M context $0.1/M input $0.6/M output

Mistral Medium 3.1 is an updated version of Mistral Medium 3, which is a high-performance enterprise-grade language model designed to deliver frontier-level capabilities at significantly reduced operational cost. It balances...

mistralai/mistral-medium-3.1 131.072K context $0.4/M input $2/M output

GPT-5 Mini is a compact version of GPT-5, designed to handle lighter-weight reasoning tasks. It provides the same instruction-following and safety-tuning benefits as GPT-5, but with reduced latency and cost....

openai/gpt-5-mini:batch 400K context $0.125/M input $1/M output

Sonnet 4.6 is Anthropic's most capable Sonnet-class model yet, with frontier performance across coding, agents, and professional work. It excels at iterative development, complex codebase navigation, end-to-end project management with...

anthropic/claude-sonnet-4.6:batch 1M context $1.5/M input $7.5/M output

Claude Sonnet 4.5 is Anthropic’s most advanced Sonnet model to date, optimized for real-world agents and coding workflows. It delivers state-of-the-art performance on coding benchmarks such as SWE-bench Verified, with...

anthropic/claude-sonnet-4.5:batch 1M context $1.5/M input $7.5/M output

Gemini 2.5 Flash is Google's state-of-the-art workhorse model, specifically designed for advanced reasoning, coding, mathematics, and scientific tasks. It includes built-in "thinking" capabilities, enabling it to provide responses with greater...

google/gemini-2.5-flash:batch 1.04858M context $0.15/M input $1.25/M output

Fugu Max is the cost-performance model in Sakana AI's Fugu family. Rather than a single monolithic model, Fugu is a learned multi-agent orchestration system: a language model trained to route...

sakana/fugu-max 1M context $2/M input $6/M output

Qwen2.5-VL is proficient in recognizing common objects such as flowers, birds, fish, and insects. It is also highly capable of analyzing texts, charts, icons, graphics, and layouts within images.

qwen/qwen2.5-vl-72b-instruct 128K context $0.8/M input $1/M output

No provider description is available for this model yet.

openrouter/openai/gpt-5.6-sol 1.05M context $2/M input $10/M output

GLM-4.5V is a vision-language foundation model for multimodal agent applications. Built on a Mixture-of-Experts (MoE) architecture with 106B parameters and 12B activated parameters, it achieves state-of-the-art results in video understanding,...

z-ai/glm-4.5v 65.536K context $0.6/M input $1.8/M output

GPT-4o mini is OpenAI's newest model after [GPT-4 Omni](/models/openai/gpt-4o), supporting both text and image inputs with text outputs. As their most advanced small model, it is many multiples more affordable...

openai/gpt-4o-mini:batch 128K context $0.075/M input $0.3/M output

Claude Fable 5 is a Mythos-class model from Anthropic, built for autonomous knowledge work and coding. It supports text, image, and file inputs with text output, with reasoning support and...

anthropic/claude-fable-5:batch 1M context $5/M input $25/M output

GPT-5.6 Sol is the flagship model in OpenAI's GPT-5.6 series. It is suited for complex reasoning, coding, and agentic workflows, and is particularly strong at command-line and multi-step coding tasks...

openai/gpt-5.6-sol:batch 1.05M context $1/M input $5/M output

Mistral Large 3 2512 is Mistral’s most capable model to date, featuring a sparse mixture-of-experts architecture with 41B active parameters (675B total), and released under the Apache 2.0 license.

mistralai/mistral-large-2512 262.144K context $0.5/M input $1.5/M output

GPT-5.6 Terra Pro is the same underlying model as [GPT-5.6 Terra](https://openrouter.ai/openai/gpt-5.6-terra), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. Learn more in OpenAI's docs: https://developers.openai.com/api/docs/guides/reasoning#reasoning-mode

openai/gpt-5.6-terra-pro:batch 1.05M context $1/M input $6/M output

No provider description is available for this model yet.

databricks/databricks-kimi-k3 1M context $3/M input $15/M output

No provider description is available for this model yet.

mistral/ministral-14b-2512 262.144K context $0.2/M input $0.2/M output

No provider description is available for this model yet.

mistral/ministral-14b-latest 262.144K context $0.2/M input $0.2/M output

No provider description is available for this model yet.

mistral/ministral-3b-2512 131.072K context $0.1/M input $0.1/M output

No provider description is available for this model yet.

mistral/ministral-3b-latest 131.072K context $0.1/M input $0.1/M output

No provider description is available for this model yet.

mistral/mistral-medium-3 262.144K context $1.5/M input $7.5/M output

No provider description is available for this model yet.

together_ai/zai-org/glm-5.3-flash 1.04858M context $0.15/M input $0.5/M output

No provider description is available for this model yet.

novita/moonshotai/kimi-k3 1.04858M context $3/M input $15/M output

No provider description is available for this model yet.

novita/moonshotai/kimi-k2.7-code 262.144K context $0.95/M input $4/M output

No provider description is available for this model yet.

novita/minimax/minimax-m3 1M context $0.3/M input $1.2/M output

No provider description is available for this model yet.

novita/qwen/qwen3.8-max 1M context $2/M input $6/M output

No provider description is available for this model yet.

novita/stepfun/step-3.7-flash 262.144K context $0.2/M input $1.15/M output

Sonnet 4.6 is Anthropic's most capable Sonnet-class model yet, with frontier performance across coding, agents, and professional work. It excels at iterative development, complex codebase navigation, end-to-end project management with...

anthropic/claude-sonnet-4.6 1M context $3/M input $15/M output

No provider description is available for this model yet.

novita/xiaomimimo/mimo-v2.5 1.04858M context $0.168/M input $0.336/M output

No provider description is available for this model yet.

novita/qwen/qwen3.6-27b 262.144K context $0.6/M input $3.6/M output

No provider description is available for this model yet.

novita/moonshotai/kimi-k2.6 262.144K context $0.8/M input $3.4/M output

Claude Opus 4.5 is Anthropic’s frontier reasoning model optimized for complex software engineering, agentic workflows, and long-horizon computer use. It offers strong multimodal capabilities, competitive performance across real-world coding and...

anthropic/claude-opus-4.5:batch 200K context $2.5/M input $12.5/M output

No provider description is available for this model yet.

novita/zai-org/glm-5v-turbo 204.8K context $1.2/M input $4/M output

No provider description is available for this model yet.

novita/google/gemma-4-26b-a4b-it 262.144K context $0.13/M input $0.4/M output

No provider description is available for this model yet.

novita/google/gemma-4-31b-it 262.144K context $0.14/M input $0.4/M output

GPT-6 Astra Pro is the same underlying model as [GPT-6 Astra](https://openrouter.ai/openai/gpt-6-astra), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. Learn more in OpenAI's docs: https://developers.openai.com/api/docs/guides/reasoning#reasoning-mode

openai/gpt-6-astra-pro:batch 1.05M context $5/M input $25/M output

No provider description is available for this model yet.

novita/qwen/qwen3.5-27b 262.144K context $0.3/M input $2.4/M output

No provider description is available for this model yet.

novita/qwen/qwen3.5-122b-a10b 262.144K context $0.4/M input $3.2/M output

No provider description is available for this model yet.

novita/qwen/qwen3.5-35b-a3b 262.144K context $0.25/M input $2/M output

Inkling is an open-weight multimodal mixture-of-experts model from Thinking Machines Lab, with 41B active parameters out of 975B total. It is designed for general-purpose reasoning, coding, agentic and tool-use systems,...

thinkingmachines/inkling:free 1.04858M context Free input Free output

No provider description is available for this model yet.

novita/qwen/qwen3.5-397b-a17b 262.144K context $0.6/M input $3.6/M output

No provider description is available for this model yet.

novita/deepseek/deepseek-ocr-2 8.192K context $0.03/M input $0.03/M output