No provider description is available for this model yet.
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
No provider description is available for this model yet.
No provider description is available for this model yet.
Qwen3.7 Flash is a vision-language reasoning model from Alibaba. It is suited for multimodal agents, visual coding, search, and computer interaction, with strengths in object recognition, spatial understanding, and real-world...
*Ling-3.0-flash* is a *124B-parameter Mixture-of-Experts (MoE) model*, with approximately *5.1B parameters activated per token*. The model is designed with *token efficiency and production-scale agentic inference* as key priorities, enabling developers...
No provider description is available for this model yet.
MiniMax-M3 is a multimodal foundation model from MiniMax. It supports text, image, and video inputs with text output, a 1M-token context window, and is suited for long-horizon agentic work, coding,...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
GPT-5.6 Luna Pro is the same underlying model as [GPT-5.6 Luna](https://openrouter.ai/openai/gpt-5.6-luna), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. Learn more in OpenAI's docs: https://developers.openai.com/api/docs/guides/reasoning#reasoning-mode
No provider description is available for this model yet.
Seed 1.6 is a general-purpose model released by the ByteDance Seed team. It incorporates multimodal capabilities and adaptive deep thinking with a 256K context window.
Seed 1.6 Flash is an ultra-fast multimodal deep thinking model by ByteDance Seed, supporting both text and visual understanding. It features a 256k context window and can generate outputs of...
Kimi K2 0905 is the September update of [Kimi K2 0711](moonshotai/kimi-k2). It is a large-scale Mixture-of-Experts (MoE) language model developed by Moonshot AI, featuring 1 trillion total parameters with 32...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
This model always redirects to the latest model in the DeepSeek V4 Flash family.
Qwen3.6 Flash is a fast, efficient language model from Alibaba's Qwen 3.6 series. It supports text, image, and video input with a 1M token context window. Tiered pricing kicks in...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Inkling is an open-weight multimodal mixture-of-experts model from Thinking Machines Lab, with 41B active parameters out of 975B total. It is designed for general-purpose reasoning, coding, agentic and tool-use systems,...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Ox Alpha is a reasoning model designed for coding, sustained agentic work, and production workloads. It is suited for long-horizon software engineering, complex reasoning, and workflows that combine text with...
Qwen Plus 0728, based on the Qwen3 foundation model, is a 1 million context hybrid reasoning model with a balanced performance, speed, and cost combination.
No provider description is available for this model yet.
A balanced model in the Ministral 3 family, Ministral 3 8B is a powerful, efficient tiny language model with vision capabilities.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| eu-north-1/minimax.minimax-m2.5bedrock/eu-north-1/minimax.minimax-m2.5 | 1M | $0.36 | $1.44 | — | |||
| eu-north-1/moonshotai.kimi-k2.5bedrock/eu-north-1/moonshotai.kimi-k2.5 | 262.144K | $0.72 | $3.6 | — | |||
| qwen-turbo-latestqwen_ai_platform/qwen-turbo-latest | 1M | $0.05 | $0.2 | — | |||
| Qwen: Qwen3.7 Flashqwen/qwen3.7-flash | 1M | $0.03 | $0.13 | — | |||
| Ling-3.0-flash (free)inclusionai/ling-3.0-flash:free | 262.144K | Free | Free | — | |||
| eu-central-1/minimax.minimax-m2.5bedrock/eu-central-1/minimax.minimax-m2.5 | 1M | $0.36 | $1.44 | — | |||
| MiniMax: MiniMax M3 (batch)minimax/minimax-m3:batch | 524.288K | $0.3 | $1.2 | — | |||
| eu-west-1/minimax.minimax-m2.5bedrock/eu-west-1/minimax.minimax-m2.5 | 1M | $0.36 | $1.44 | — | |||
| gemma-4-26b-a4b-itgemini/gemma-4-26b-a4b-it | 262.144K | — | — | — | |||
| moonshotai/kimi-k2.6novita/moonshotai/kimi-k2.6 | 262.144K | $0.8 | $3.4 | — | |||
| OpenAI: GPT-5.6 Luna Proopenai/gpt-5.6-luna-pro | 1.05M | $0.2 | $1.2 | — | |||
| eu-west-2/minimax.minimax-m2.5bedrock/eu-west-2/minimax.minimax-m2.5 | 1M | $0.47 | $1.86 | — | |||
| ByteDance Seed: Seed 1.6bytedance-seed/seed-1.6 | 262.144K | $0.25 | $2 | — | |||
| ByteDance Seed: Seed 1.6 Flashbytedance-seed/seed-1.6-flash | 262.144K | $0.075 | $0.3 | — | |||
| MoonshotAI: Kimi K2 0905moonshotai/kimi-k2-0905 | 262.144K | $0.6 | $2.5 | — | |||
| eu-south-1/qwen.qwen3-coder-nextbedrock/eu-south-1/qwen.qwen3-coder-next | 262.144K | $0.6 | $1.44 | — | |||
| qwen-turbo-2025-04-28qwen_ai_platform/qwen-turbo-2025-04-28 | 1M | $0.05 | $0.2 | — | |||
| databricks-glm-5-3-flashdatabricks/databricks-glm-5-3-flash | 1.04858M | — | — | — | |||
| DeepSeek: DeepSeek V4 Flash Latest~deepseek/deepseek-v4-flash-latest | 1.04858M | $0.03 | $0.07 | — | |||
| Qwen: Qwen3.6 Flashqwen/qwen3.6-flash | 1M | $0.188 | $1.125 | — | |||
| qwen-turbo-2024-11-01qwen_ai_platform/qwen-turbo-2024-11-01 | 1M | $0.05 | $0.2 | — | |||
| qwen-plus-latestqwen_ai_platform/qwen-plus-latest | 997.952K | — | — | — | |||
| invoke/anthropic.claude-3-5-sonnet-20240620-v1:0bedrock/invoke/anthropic.claude-3-5-sonnet-20240620-v1:0 | 200K | $3 | $15 | — | |||
| Thinking Machines: Inkling (free)thinkingmachines/inkling:free | 1.04858M | Free | Free | — | |||
| google/gemma-4-31B-it-turbodeepinfra/google/gemma-4-31b-it-turbo | 262.144K | $0.09 | $0.34 | — | |||
| qwen/qwen3.6-27bnovita/qwen/qwen3.6-27b | 262.144K | $0.6 | $3.6 | — | |||
| ministral-14b-latestmistral/ministral-14b-latest | 262.144K | $0.2 | $0.2 | — | |||
| ministral-14b-2512mistral/ministral-14b-2512 | 262.144K | $0.2 | $0.2 | — | |||
| qwen-plus-2025-09-11qwen_ai_platform/qwen-plus-2025-09-11 | 997.952K | — | — | — | |||
| qwen-plus-2025-07-28qwen_ai_platform/qwen-plus-2025-07-28 | 997.952K | — | — | — | |||
| Qwen/Qwen3-Maxdeepinfra/qwen/qwen3-max | 256K | $1.2 | $6 | — | |||
| qwen-flash-2025-07-28qwen_ai_platform/qwen-flash-2025-07-28 | 997.952K | — | — | — | |||
| Ox Alphastealth/ox-alpha | 1.04858M | — | — | — | |||
| Qwen: Qwen Plus 0728 (thinking)qwen/qwen-plus-2025-07-28:thinking | 1M | $0.26 | $0.78 | — | |||
| XiaomiMiMo/MiMo-V2.5deepinfra/xiaomimimo/mimo-v2.5 | 262.144K | $0.4 | $2 | — | |||
| Mistral: Ministral 3 8B 2512 (batch)mistralai/ministral-8b-2512:batch | 262.144K | $0.075 | $0.075 | — | |||
| qwen-flashqwen_ai_platform/qwen-flash | 997.952K | — | — | — | |||
| sa-east-1/minimax.minimax-m2.5bedrock/sa-east-1/minimax.minimax-m2.5 | 1M | $0.36 | $1.44 | — | |||
| sa-east-1/moonshotai.kimi-k2-thinkingbedrock/sa-east-1/moonshotai.kimi-k2-thinking | 262.144K | $0.73 | $3.03 | — | |||
| qwen-coderqwen_ai_platform/qwen-coder | 1M | $0.3 | $1.5 | — | |||
| google/gemini-3.5-flashdeepinfra/google/gemini-3.5-flash | 1M | $1.5 | $9 | — | |||
| sa-east-1/moonshotai.kimi-k2.5bedrock/sa-east-1/moonshotai.kimi-k2.5 | 262.144K | $0.72 | $3.6 | — | |||
| kimi-k2.7-codeqwen_ai_platform/kimi-k2.7-code | 229.376K | $0.95 | $4 | — | |||
| glm-5.2qwen_ai_platform/glm-5.2 | 1.04858M | $1.4 | $4.4 | — | |||
| anthropic/claude-sonnet-4-6deepinfra/anthropic/claude-sonnet-4-6 | 1M | $3 | $15 | — | |||
| chat-latestopenai/chat-latest | 400K | $5 | $30 | — | |||
| xiaomimimo/mimo-v2.5-pronovita/xiaomimimo/mimo-v2.5-pro | 1.04858M | $0.522 | $1.044 | — | |||
| glm-5-2mistral/glm-5-2 | 1.04858M | $1.4 | $4.4 | — | |||
| anthropic/claude-opus-5openrouter/anthropic/claude-opus-5 | 1M | $5 | $25 | — | |||
| accounts/fireworks/models/deepseek-v4-pro-0813fireworks_ai/accounts/fireworks/models/deepseek-v4-pro-0813 | 1.04858M | $1.32 | $3.96 | — |