No provider description is available for this model yet.
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
GLM-5.3-Flash is a native multimodal model from Z.ai. It is suited for efficient coding and long-horizon agent tasks. Its hybrid sparse and linear attention architecture maintains accurate long-context behavior while...
Gemini 2.5 Pro is Google’s state-of-the-art AI model designed for advanced reasoning, coding, mathematics, and scientific tasks. It employs “thinking” capabilities, enabling it to reason through responses with enhanced accuracy...
GPT-5.6 Luna Pro is the same underlying model as [GPT-5.6 Luna](https://openrouter.ai/openai/gpt-5.6-luna), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. Learn more in OpenAI's docs: https://developers.openai.com/api/docs/guides/reasoning#reasoning-mode
GPT-4.1 Mini is a mid-sized model delivering performance competitive with GPT-4o at substantially lower latency and cost. It retains a 1 million token context window and scores 45.1% on hard...
Qwen-Plus, based on the Qwen2.5 foundation model, is a 131K context model with a balanced performance, speed, and cost combination.
Claude Opus 4.8 is Anthropic's most capable generally available model in the Opus family. It supports text, image, and file inputs with text output, with reasoning support and a 1M-token...
Ministral 8B is an 8B parameter model featuring a unique interleaved sliding-window attention pattern for faster, memory-efficient inference. Designed for edge use cases, it supports up to 128k context length...
Fugu Max is the cost-performance model in Sakana AI's Fugu family. Rather than a single monolithic model, Fugu is a learned multi-agent orchestration system: a language model trained to route...
No provider description is available for this model yet.
NVIDIA Nemotron 3.5 Lightning is an open mixture-of-experts model from NVIDIA, with 3B active parameters out of 30B total. It is suited for high-throughput agentic workloads and specialized tasks that...
GPT-4o mini is OpenAI's newest model after [GPT-4 Omni](/models/openai/gpt-4o), supporting both text and image inputs with text outputs. As their most advanced small model, it is many multiples more affordable...
GPT-5.6 Luna is a fast, cost-efficient model in OpenAI's GPT-5.6 series. It is suited for high-volume, latency-sensitive tasks such as chat, classification, and lightweight agentic workflows, providing capable reasoning for...
GPT-5.6 Sol is the flagship model in OpenAI's GPT-5.6 series. It is suited for complex reasoning, coding, and agentic workflows, and is particularly strong at command-line and multi-step coding tasks...
GPT-5 Pro is OpenAI’s most advanced model, offering major improvements in reasoning, code quality, and user experience. It is optimized for complex tasks that require step-by-step reasoning, instruction following, and...
Sonnet 4.6 is Anthropic's most capable Sonnet-class model yet, with frontier performance across coding, agents, and professional work. It excels at iterative development, complex codebase navigation, end-to-end project management with...
GPT-5.6 Terra Pro is the same underlying model as [GPT-5.6 Terra](https://openrouter.ai/openai/gpt-5.6-terra), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. Learn more in OpenAI's docs: https://developers.openai.com/api/docs/guides/reasoning#reasoning-mode
Mistral Large 3 2512 is Mistral’s most capable model to date, featuring a sparse mixture-of-experts architecture with 41B active parameters (675B total), and released under the Apache 2.0 license.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
MiniMax-M1 is a large-scale, open-weight reasoning model designed for extended context and high-efficiency inference. It leverages a hybrid Mixture-of-Experts (MoE) architecture paired with a custom "lightning attention" mechanism, allowing it...
No provider description is available for this model yet.
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| openai-gpt-4.1snowflake/openai-gpt-4.1 | 300K | $2 | $8 | — | |||
| openai-gpt-5snowflake/openai-gpt-5 | 300K | $1.25 | $10 | — | |||
| openai-gpt-5-minisnowflake/openai-gpt-5-mini | 1M | $0.3 | $1.2 | — | |||
| openai-gpt-5-nanosnowflake/openai-gpt-5-nano | 5M | $0.15 | $0.6 | — | |||
| llama4-mavericksnowflake/llama4-maverick | 128K | $0.24 | $0.97 | — | |||
| Qwen/Qwen3.5-397B-A17B-FP8tensormesh/qwen/qwen3.5-397b-a17b-fp8 | 262.144K | $0.6 | $3.6 | — | |||
| Qwen/Qwen3-Coder-480B-A35B-Instruct-FP8tensormesh/qwen/qwen3-coder-480b-a35b-instruct-fp8 | 262.144K | $0.45 | $1.8 | — | |||
| Qwen/Qwen3.6-27B-FP8tensormesh/qwen/qwen3.6-27b-fp8 | 262.144K | $0.32 | $3.2 | — | |||
| lukealonso/GLM-5.1-NVFP4-MTPtensormesh/lukealonso/glm-5.1-nvfp4-mtp | 202.752K | $1.4 | $4.4 | — | |||
| MiniMaxAI/MiniMax-M2.5tensormesh/minimaxai/minimax-m2.5 | 196.608K | $0.3 | $1.2 | — | |||
| openai/gpt-oss-120btensormesh/openai/gpt-oss-120b | 131.072K | $0.15 | $0.6 | — | |||
| openai/gpt-oss-20btensormesh/openai/gpt-oss-20b | 131.072K | $0.07 | $0.28 | — | |||
| deepseek-v4-protencent/deepseek-v4-pro | 1M | $0.435 | $0.87 | — | |||
| deepseek-v4-flashtencent/deepseek-v4-flash | 1M | $0.14 | $0.28 | — | |||
| ps/glm-4.5-airpinstripes/ps/glm-4.5-air | 128K | $0.125 | $0.45 | — | |||
| ps/qwen3.6-35b-a3bpinstripes/ps/qwen3.6-35b-a3b | 131.072K | $0.14 | $0.45 | — | |||
| ps/qwen3-30b-a3bpinstripes/ps/qwen3-30b-a3b | 131.072K | $0.09 | $0.2 | — | |||
| ps/qwen3-coder-30b-a3bpinstripes/ps/qwen3-coder-30b-a3b | 131.072K | $0.3 | $0.6 | — | |||
| ps/deepseek-v4-flashpinstripes/ps/deepseek-v4-flash | 163.84K | $0.1 | $0.2 | — | |||
| ps/minimax-m2.7pinstripes/ps/minimax-m2.7 | 1.00019M | $0.255 | $0.55 | — | |||
| gemma-4-26bdarkbloom/gemma-4-26b | 131.072K | $0.03 | $0.165 | — | |||
| gpt-oss-20bdarkbloom/gpt-oss-20b | 131.072K | $0.015 | $0.07 | — | |||
| Z.ai: GLM 5.3 Flash (batch)z-ai/glm-5.3-flash:batch | 1.04858M | $0.075 | $0.25 | — | |||
| Google: Gemini 2.5 Pro Preview 05-06google/gemini-2.5-pro-preview-05-06 | 1.04858M | $1.25 | $10 | — | |||
| OpenAI: GPT-5.6 Luna Pro (batch)openai/gpt-5.6-luna-pro:batch | 1.05M | $0.1 | $0.6 | — | |||
| OpenAI: GPT-4.1 Mini (batch)openai/gpt-4.1-mini:batch | 1.04758M | $0.2 | $0.8 | — | |||
| Qwen: Qwen-Plusqwen/qwen-plus | 1M | $0.26 | $0.78 | — | |||
| Anthropic: Claude Opus 4.8anthropic/claude-opus-4.8 | 1M | $5 | $25 | — | |||
| Mistral: Ministral 8Bmistralai/ministral-8b | 128K | $0.11 | $0.11 | — | |||
| Sakana: Fugu Maxsakana/fugu-max | 1M | $2 | $6 | — | |||
| openai/gpt-5.6-solopenrouter/openai/gpt-5.6-sol | 1.05M | $2 | $10 | — | |||
| NVIDIA: Nemotron 3.5 Lightning (free)nvidia/nemotron-3.5-lightning:free | 1M | Free | Free | — | |||
| OpenAI: GPT-4o-mini (batch)openai/gpt-4o-mini:batch | 128K | $0.075 | $0.3 | — | |||
| OpenAI: GPT-5.6 Luna (batch)openai/gpt-5.6-luna:batch | 1.05M | $0.1 | $0.6 | — | |||
| OpenAI: GPT-5.6 Sol (batch)openai/gpt-5.6-sol:batch | 1.05M | $1 | $5 | — | |||
| OpenAI: GPT-5 Pro (batch)openai/gpt-5-pro:batch | 400K | $7.5 | $60 | — | |||
| Anthropic: Claude Sonnet 4.6 (batch)anthropic/claude-sonnet-4.6:batch | 1M | $1.5 | $7.5 | — | |||
| OpenAI: GPT-5.6 Terra Pro (batch)openai/gpt-5.6-terra-pro:batch | 1.05M | $1 | $6 | — | |||
| Mistral: Mistral Large 3 2512mistralai/mistral-large-2512 | 262.144K | $0.5 | $1.5 | — | |||
| databricks-glm-5-2databricks/databricks-glm-5-2 | 1M | $1.4 | $4.4 | — | |||
| databricks-kimi-k3databricks/databricks-kimi-k3 | 1M | $3 | $15 | — | |||
| accounts/fireworks/models/deepseek-v4-pro-0813fireworks_ai/accounts/fireworks/models/deepseek-v4-pro-0813 | 1.04858M | $1.32 | $3.96 | — | |||
| ministral-14b-2512mistral/ministral-14b-2512 | 262.144K | $0.2 | $0.2 | — | |||
| ministral-14b-latestmistral/ministral-14b-latest | 262.144K | $0.2 | $0.2 | — | |||
| ministral-3b-2512mistral/ministral-3b-2512 | 131.072K | $0.1 | $0.1 | — | |||
| ministral-3b-latestmistral/ministral-3b-latest | 131.072K | $0.1 | $0.1 | — | |||
| mistral-medium-3mistral/mistral-medium-3 | 262.144K | $1.5 | $7.5 | — | |||
| zai-org/GLM-5.3-Flashtogether_ai/zai-org/glm-5.3-flash | 1.04858M | $0.15 | $0.5 | — | |||
| MiniMax: MiniMax M1minimax/minimax-m1 | 1M | $0.55 | $2.2 | — | |||
| glm-5.3zai/glm-5.3 | 1M | $1.4 | $4.4 | — |