No provider description is available for this model yet.
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
No provider description is available for this model yet.
GLM-5V-Turbo is Z.ai’s first native multimodal agent foundation model, built for vision-based coding and agent-driven tasks. It natively handles image, video, and text inputs, excels at long-horizon planning, complex coding,...
No provider description is available for this model yet.
No provider description is available for this model yet.
GPT-5.4 mini brings the core capabilities of GPT-5.4 to a faster, more efficient model optimized for high-throughput workloads. It supports text and image inputs with strong performance across reasoning, coding,...
No provider description is available for this model yet.
Grok 4.20 Multi-Agent is a variant of SpaceXAI’s Grok 4.20 designed for collaborative, agent-based workflows. Multiple agents operate in parallel to conduct deep research, coordinate tool use, and synthesize information...
No provider description is available for this model yet.
No provider description is available for this model yet.
GPT-5.4 Pro is OpenAI's most advanced model, building on GPT-5.4's unified architecture with enhanced reasoning capabilities for complex, high-stakes tasks. It features a 1M+ token context window (922K input, 128K...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Sakana Namazu is a Japanese-specialized reasoning model from Sakana AI, based on Kimi K2.6 with additional training for Japanese language and business contexts. It is suited for Japanese instruction following,...
No provider description is available for this model yet.
Qwen3.8 2.4T A95B is an open-weight sparse mixture-of-experts model from Qwen and the open-weight variant of [Qwen3.8 Max](/qwen/qwen3.8-max), with 95 billion active parameters out of 2.4 trillion total. It is...
NVIDIA Nemotron 3 Nano 30B A3B is a small language MoE model with highest compute efficiency and accuracy for developers to build specialized agentic AI systems. The model is fully...
No provider description is available for this model yet.
The relace-search model uses 4-12 `view_file` and `grep` tools in parallel to explore a codebase and return relevant files to the user request. In contrast to RAG, relace-search performs agentic...
No provider description is available for this model yet.
MoonshotAI: Kimi K2.7 Code is a coding-focused model in Moonshot AI's Kimi K2 family, built to complete end-to-end programming tasks reliably over long contexts. It uses a native multimodal mixture-of-experts...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Qwen3.8 Max 0902 is an updated snapshot of Qwen3.8 Max from Alibaba's Qwen team. It is a 2.4-trillion-parameter mixture-of-experts model that accepts text, image, and video input and returns text,...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Opus 4.6 is Anthropic’s strongest model for coding and long-running professional tasks. It is built for agents that operate across entire workflows rather than single prompts, making it especially effective...
No provider description is available for this model yet.
No provider description is available for this model yet.
GPT-6 Astra Pro is the same underlying model as [GPT-6 Astra](https://openrouter.ai/openai/gpt-6-astra), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. Learn more in OpenAI's docs: https://developers.openai.com/api/docs/guides/reasoning#reasoning-mode
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Qwen3-Coder-480B-A35B-Instruct is a Mixture-of-Experts (MoE) code generation model developed by the Qwen team. It is optimized for agentic coding tasks such as function calling, tool use, and long-context reasoning over...
No provider description is available for this model yet.
The simplest way to get free inference. openrouter/free is a router that selects free models at random from the models available on OpenRouter. The router smartly filters for models that...
Nova 2 Lite is a fast, cost-effective reasoning model for everyday workloads that can process text, images, and videos to generate text. Nova 2 Lite demonstrates standout capabilities in processing...
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| xai.grok-4.6bedrock_mantle/xai.grok-4.6 | 500K | $2.2 | $6.6 | — | |||
| claude-haiku-4-5azure_ai/claude-haiku-4-5 | 200K | $1 | $5 | — | |||
| Z.ai: GLM 5V Turboz-ai/glm-5v-turbo | 202.752K | $1.2 | $4 | — | |||
| claude-opus-4-5azure_ai/claude-opus-4-5 | 200K | $5 | $25 | — | |||
| claude-opus-4-6azure_ai/claude-opus-4-6 | 1M | $5 | $25 | — | |||
| OpenAI: GPT-5.4 Mini (batch)openai/gpt-5.4-mini:batch | 400K | $0.375 | $2.25 | — | |||
| claude-opus-4-7azure_ai/claude-opus-4-7 | 1M | $5 | $25 | — | |||
| SpaceXAI: Grok 4.20 Multi-Agentx-ai/grok-4.20-multi-agent | 2M | $1.25 | $2.5 | — | |||
| claude-fable-5azure_ai/claude-fable-5 | 1M | $10 | $50 | — | |||
| claude-opus-5azure_ai/claude-opus-5 | 1M | $5 | $25 | — | |||
| OpenAI: GPT-5.4 Pro (batch)openai/gpt-5.4-pro:batch | 1.05M | $15 | $90 | — | |||
| claude-opus-4-8azure_ai/claude-opus-4-8 | 1M | $5 | $25 | — | |||
| us-gov-west-1/anthropic.claude-3-7-sonnet-20250219-v1:0bedrock/us-gov-west-1/anthropic.claude-3-7-sonnet-20250219-v1:0 | 200K | $3.6 | $18 | — | |||
| global.openai.gpt-5.6-terrabedrock_converse/global.openai.gpt-5.6-terra | 1M | $2 | $12 | — | |||
| deepseek/deepseek-v4-pro-0813openrouter/deepseek/deepseek-v4-pro-0813 | 1.04858M | $1.32 | $3.96 | — | |||
| Sakana: Namazusakana/namazu | 262.144K | $0.95 | $4 | — | |||
| claude-sonnet-4-5azure_ai/claude-sonnet-4-5 | 200K | $3 | $15 | — | |||
| Qwen: Qwen3.8 2.4T A95B (batch)qwen/qwen3.8-2.4t-a95b:batch | 1.01M | $2 | $6 | — | |||
| NVIDIA: Nemotron 3 Nano 30B A3B (free)nvidia/nemotron-3-nano-30b-a3b:free | 256K | Free | Free | — | |||
| claude-sonnet-5azure_ai/claude-sonnet-5 | 1M | $2 | $10 | — | |||
| Relace: Relace Searchrelace/relace-search | 256K | $1 | $3 | — | |||
| gemini-3.8-flashvertex_ai/gemini-3.8-flash | 1.04858M | $0.75 | $3.75 | — | |||
| MoonshotAI: Kimi K2.7 Code (batch)moonshotai/kimi-k2.7-code:batch | 262.144K | $0.95 | $4 | — | |||
| gpt-5.5azure_ai/gpt-5.5 | 1.05M | $5 | $30 | — | |||
| gemini-3.8-flashgemini/gemini-3.8-flash | 1.04858M | $0.75 | $3.75 | — | |||
| us-gov-west-1/claude-sonnet-4-5-20250929-v1:0bedrock/us-gov-west-1/claude-sonnet-4-5-20250929-v1:0 | 200K | $3.6 | $18 | — | |||
| deepseek-v4-flash-0731dashscope/deepseek-v4-flash-0731 | 1M | $0.2 | $0.4 | — | |||
| deepseek-v4-flashdashscope/deepseek-v4-flash | 1M | $0.2 | $0.4 | — | |||
| deepseek/deepseek-v4-proopenrouter/deepseek/deepseek-v4-pro | 1.04858M | $1.32 | $3.96 | — | |||
| Qwen: Qwen3.8 Max (0902)qwen/qwen3.8-max-0902 | 1M | $2 | $6 | — | |||
| us.openai.gpt-5.6-terrabedrock_converse/us.openai.gpt-5.6-terra | 1M | $2.2 | $13.2 | — | |||
| gpt-5.5-2026-04-23azure_ai/gpt-5.5-2026-04-23 | 1.05M | $5 | $30 | — | |||
| anthropic/claude-opus-5openrouter/anthropic/claude-opus-5 | 1M | $5 | $25 | — | |||
| Anthropic: Claude Opus 4.6anthropic/claude-opus-4.6 | 1M | $5 | $25 | — | |||
| @cf/nvidia/nemotron-3-120b-a12bcloudflare/@cf/nvidia/nemotron-3-120b-a12b | 256K | $0.5 | $1.5 | — | |||
| grok-build-latestxai/grok-build-latest | 500K | $2 | $6 | — | |||
| OpenAI: GPT-6 Astra Pro (batch)openai/gpt-6-astra-pro:batch | 1.05M | $5 | $25 | — | |||
| accounts/fireworks/models/inklingfireworks_ai/accounts/fireworks/models/inkling | 1.04858M | $1 | $4.05 | — | |||
| glm-5.2scaleway/glm-5.2 | 256K | $1.8 | $5.5 | — | |||
| deepseek-v4-flash-0731scaleway/deepseek-v4-flash-0731 | 256K | $0.4 | $0.8 | — | |||
| glm-5-2mistral/glm-5-2 | 1.04858M | $1.4 | $4.4 | — | |||
| kimi-k2.7-codeazure_ai/kimi-k2.7-code | 262.144K | $0.95 | $4 | — | |||
| us-gov-west-1/nvidia.nemotron-nano-3-30bbedrock/us-gov-west-1/nvidia.nemotron-nano-3-30b | 262.144K | $0.072 | $0.288 | — | |||
| us-gov-west-1/nvidia.nemotron-super-3-120bbedrock/us-gov-west-1/nvidia.nemotron-super-3-120b | 256K | $0.18 | $0.78 | — | |||
| global.openai.gpt-5.6-solbedrock_converse/global.openai.gpt-5.6-sol | 1M | $5 | $30 | — | |||
| us-gov-west-1/anthropic.claude-opus-4-8bedrock/us-gov-west-1/anthropic.claude-opus-4-8 | 1M | $6 | $30 | — | |||
| Qwen: Qwen3 Coder 480B A35B (free)qwen/qwen3-coder:free | 262K | Free | Free | — | |||
| zai-glm-5-2mistral/zai-glm-5-2 | 1.04858M | $1.4 | $4.4 | — | |||
| Free Models Routeropenrouter/free | 200K | Free | Free | — | |||
| Amazon: Nova 2 Liteamazon/nova-2-lite-v1 | 1M | $0.3 | $2.5 | — |