No provider description is available for this model yet.
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
GPT-5.1 is the latest frontier-grade model in the GPT-5 series, offering stronger general-purpose reasoning, improved instruction adherence, and a more natural conversational style compared to GPT-5. It uses adaptive reasoning...
No provider description is available for this model yet.
No provider description is available for this model yet.
Mistral's cutting-edge language model for coding released end of July 2025. Codestral specializes in low-latency, high-frequency tasks such as fill-in-the-middle (FIM), code correction and test generation. [Blog Post](https://mistral.ai/news/codestral-25-08)
Qwen3.6 Flash is a fast, efficient language model from Alibaba's Qwen 3.6 series. It supports text, image, and video input with a 1M token context window. Tiered pricing kicks in...
This model always redirects to the latest model in the DeepSeek V4 Flash family.
Kimi K2 0905 is the September update of [Kimi K2 0711](moonshotai/kimi-k2). It is a large-scale Mixture-of-Experts (MoE) language model developed by Moonshot AI, featuring 1 trillion total parameters with 32...
Seed 1.6 Flash is an ultra-fast multimodal deep thinking model by ByteDance Seed, supporting both text and visual understanding. It features a 256k context window and can generate outputs of...
Seed 1.6 is a general-purpose model released by the ByteDance Seed team. It incorporates multimodal capabilities and adaptive deep thinking with a 256K context window.
GPT-5.6 Luna Pro is the same underlying model as [GPT-5.6 Luna](https://openrouter.ai/openai/gpt-5.6-luna), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. Learn more in OpenAI's docs: https://developers.openai.com/api/docs/guides/reasoning#reasoning-mode
*Ling-3.0-flash* is a *124B-parameter Mixture-of-Experts (MoE) model*, with approximately *5.1B parameters activated per token*. The model is designed with *token efficiency and production-scale agentic inference* as key priorities, enabling developers...
Fast-mode variant of [Opus 5](/anthropic/claude-opus-5) - identical capabilities with higher output speed at 2x pricing relative to regular Opus 5. Learn more in Anthropic's docs: https://platform.claude.com/docs/en/build-with-claude/fast-mode Note: As of September...
Qwen3.7 Flash is a vision-language reasoning model from Alibaba. It is suited for multimodal agents, visual coding, search, and computer interaction, with strengths in object recognition, spatial understanding, and real-world...
Laguna S 2.1 is the latest coding agent model from [Poolside](<https://poolside.ai/>). Laguna S 2.1 is a 118B total parameter model with 8B active parameters, scoring 70.2% on Terminal-Bench 2.1 and...
GPT-5.6 Terra Pro is the same underlying model as [GPT-5.6 Terra](https://openrouter.ai/openai/gpt-5.6-terra), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. Learn more in OpenAI's docs: https://developers.openai.com/api/docs/guides/reasoning#reasoning-mode
GPT-5.6 Sol Pro is the same underlying model as [GPT-5.6 Sol](https://openrouter.ai/openai/gpt-5.6-sol), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. Learn more in OpenAI's docs: https://developers.openai.com/api/docs/guides/reasoning#reasoning-mode
Fast-mode variant of [Opus 4.8](/anthropic/claude-opus-4.8) - identical capabilities with higher output speed at 2x pricing relative to regular Opus 4.8. Learn more in Anthropic's docs: https://platform.claude.com/docs/en/build-with-claude/fast-mode Note: As of September...
KAT-Coder-Air V2.5 is a flagship-level Agentic Coding model that can directly hand over an entire issue or an entire business workflow to it, allowing it to autonomously locate and make...
Auto Router (Beta) is a task-aware router from OpenRouter. It classifies each request, then routes it the [most popular model](/rankings#task-spend) for that task based on aggregate spend, filtered by your...
KAT-Coder-Pro V2.5 is a flagship-level Agentic Coding model that can directly hand over an entire issue or an entire business workflow to it, allowing it to autonomously locate and make...
This model always redirects to the latest Grok model from xAI.
North Mini Code is Cohere's first agentic coding model and the debut of its North family. A sparse mixture-of-experts model with 30B total parameters and 3B active, it is optimized...
No provider description is available for this model yet.
GLM 5.2 is a large-scale reasoning model from Z.ai. It supports text input and output with a 1M-token context window, and is suited for long-horizon agent workflows, project-level software engineering,...
Fusion turns your prompt into a small multi-model deliberation. A panel of expert models (see below) analyzes your prompt in parallel with web search and web fetch enabled, then a...
This model always redirects to the latest model in the Claude Fable family.
Qwen3.7-Max is the flagship model in Alibaba's Qwen3.7 series. It supports text input and output and is designed for agent-centric workloads, with particular strengths in coding, office and productivity tasks,...
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| moonshotai/kimi-k2-thinkingopenrouter/moonshotai/kimi-k2-thinking | 262.144K | $0.6 | $2.5 | — | |||
| qwen/qwen3-vl-8b-instructopenrouter/qwen/qwen3-vl-8b-instruct | 262.144K | $0.117 | $0.455 | — | |||
| qwen/qwen3-vl-30b-a3b-thinkingopenrouter/qwen/qwen3-vl-30b-a3b-thinking | 262.144K | $0.2 | $2.4 | — | |||
| qwen/qwen3-vl-30b-a3b-instructopenrouter/qwen/qwen3-vl-30b-a3b-instruct | 262.144K | $0.15 | $0.6 | — | |||
| openai/gpt-5-proopenrouter/openai/gpt-5-pro | 400K | $15 | $120 | — | |||
| qwen/qwen3-vl-235b-a22b-instructopenrouter/qwen/qwen3-vl-235b-a22b-instruct | 262.144K | $0.21 | $1.9 | — | |||
| qwen/qwen3-maxopenrouter/qwen/qwen3-max | 262.144K | $0.78 | $3.9 | — | |||
| qwen/qwen3-coder-flashopenrouter/qwen/qwen3-coder-flash | 1M | $0.195 | $0.975 | — | |||
| qwen/qwen3-next-80b-a3b-thinkingopenrouter/qwen/qwen3-next-80b-a3b-thinking | 262.144K | $0.15 | $1.2 | — | |||
| qwen/qwen3-next-80b-a3b-instructopenrouter/qwen/qwen3-next-80b-a3b-instruct | 262.144K | $0.1 | $1.1 | — | |||
| qwen/qwen-plus-2025-07-28openrouter/qwen/qwen-plus-2025-07-28 | 1M | $0.26 | $0.78 | — | |||
| moonshotai/kimi-k2-0905openrouter/moonshotai/kimi-k2-0905 | 262.144K | $0.6 | $2.5 | — | |||
| mistralai/codestral-2508openrouter/mistralai/codestral-2508 | 256K | $0.3 | $0.9 | — | |||
| qwen/qwen3-coder-30b-a3b-instructopenrouter/qwen/qwen3-coder-30b-a3b-instruct | 262.144K | $0.07 | $0.28 | — | |||
| qwen/qwen3-30b-a3b-instruct-2507openrouter/qwen/qwen3-30b-a3b-instruct-2507 | 262.144K | $0.048 | $0.193 | — | |||
| minimax/minimax-m1openrouter/minimax/minimax-m1 | 1M | $0.55 | $2.2 | — | |||
| openai/o3-proopenrouter/openai/o3-pro | 200K | $20 | $80 | — | |||
| google/gemini-2.5-pro-previewopenrouter/google/gemini-2.5-pro-preview | 1.04858M | $1.25 | $10 | — | |||
| google/gemini-2.5-pro-preview-05-06openrouter/google/gemini-2.5-pro-preview-05-06 | 1.04858M | $1.25 | $10 | — | |||
| openai/o4-mini-highopenrouter/openai/o4-mini-high | 200K | $1.1 | $4.4 | — | |||
| meta-llama/llama-4-maverickopenrouter/meta-llama/llama-4-maverick | 1.04858M | $0.2 | $0.696 | — | |||
| meta-llama/llama-4-scoutopenrouter/meta-llama/llama-4-scout | 1.31072M | $0.1 | $0.3 | — | |||
| openai/o1-proopenrouter/openai/o1-pro | 200K | $150 | $600 | — | |||
| OpenAI: GPT-5.1 (batch)openai/gpt-5.1:batch | 400K | $0.625 | $5 | — | |||
| qwen/qwen-plusopenrouter/qwen/qwen-plus | 1M | $0.26 | $0.78 | — | |||
| minimax/minimax-01openrouter/minimax/minimax-01 | 1.00019M | $0.2 | $1.1 | — | |||
| Mistral: Codestral 2508 (batch)mistralai/codestral-2508:batch | 256K | $0.15 | $0.45 | — | |||
| Qwen: Qwen3.6 Flashqwen/qwen3.6-flash | 1M | $0.188 | $1.125 | — | |||
| DeepSeek: DeepSeek V4 Flash Latest~deepseek/deepseek-v4-flash-latest | 1.04858M | $0.035 | $0.106 | — | |||
| MoonshotAI: Kimi K2 0905moonshotai/kimi-k2-0905 | 262.144K | $0.6 | $2.5 | — | |||
| ByteDance Seed: Seed 1.6 Flashbytedance-seed/seed-1.6-flash | 262.144K | $0.075 | $0.3 | — | |||
| ByteDance Seed: Seed 1.6bytedance-seed/seed-1.6 | 262.144K | $0.25 | $2 | — | |||
| OpenAI: GPT-5.6 Luna Proopenai/gpt-5.6-luna-pro | 1.05M | $0.2 | $1.2 | — | |||
| Ling-3.0-flash (free)inclusionai/ling-3.0-flash:free | 262.144K | Free | Free | — | |||
| Claude Opus 5 (Fast)anthropic/claude-opus-5-fast | 1M | $10 | $50 | — | |||
| Qwen: Qwen3.7 Flashqwen/qwen3.7-flash | 1M | $0.03 | $0.13 | — | |||
| Poolside: Laguna S 2.1 (free)poolside/laguna-s-2.1:free | 262.144K | Free | Free | — | |||
| OpenAI: GPT-5.6 Terra Proopenai/gpt-5.6-terra-pro | 1.05M | $2 | $12 | — | |||
| OpenAI: GPT-5.6 Sol Proopenai/gpt-5.6-sol-pro | 1.05M | $2 | $10 | — | |||
| Anthropic: Claude Opus 4.8 (Fast)anthropic/claude-opus-4.8-fast | 1M | $10 | $50 | — | |||
| Kwaipilot: KAT-Coder-Air V2.5kwaipilot/kat-coder-air-v2.5 | 256K | $0.15 | $0.6 | — | |||
| Auto Router (Beta)openrouter/auto-beta | 2M | — | — | — | |||
| Kwaipilot: KAT-Coder-Pro V2.5kwaipilot/kat-coder-pro-v2.5 | 262.144K | $0.74 | $2.96 | — | |||
| xAI: Grok Latest~x-ai/grok-latest | 500K | $2 | $6 | — | |||
| Cohere: North Mini Code (free)cohere/north-mini-code:free | 256K | Free | Free | — | |||
| us-east-2/minimax.minimax-m2.5bedrock/us-east-2/minimax.minimax-m2.5 | 1M | $0.3 | $1.2 | — | |||
| Z.ai: GLM 5.2z-ai/glm-5.2 | 202.752K | $0.6 | $2 | — | |||
| OpenRouter: Fusionopenrouter/fusion | 1M | — | — | — | |||
| Anthropic: Claude Fable Latest~anthropic/claude-fable-latest | 1M | $10 | $50 | — | |||
| Qwen: Qwen3.7 Maxqwen/qwen3.7-max | 1M | $1.475 | $4.425 | — |