Grok 4.20 is a reasoning model from SpaceXAI with industry-leading speed and agentic tool calling capabilities. It combines the lowest hallucination rate on the market with strict prompt adherance, delivering...
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
Opus 4.7 is the next generation of Anthropic's Opus family, built for long-running, asynchronous agents. Building on the coding and agentic strengths of Opus 4.6, it delivers stronger performance on...
No provider description is available for this model yet.
No provider description is available for this model yet.
Qwen Plus 0728, based on the Qwen3 foundation model, is a 1 million context hybrid reasoning model with a balanced performance, speed, and cost combination.
No provider description is available for this model yet.
No provider description is available for this model yet.
Qwen3.5-9B is a multimodal foundation model from the Qwen3.5 family, designed to deliver strong reasoning, coding, and visual understanding in an efficient 9B-parameter architecture. It uses a unified vision-language design...
No provider description is available for this model yet.
No provider description is available for this model yet.
Inkling is an open-weight multimodal mixture-of-experts model from Thinking Machines Lab, with 41B active parameters out of 975B total. It is designed for general-purpose reasoning, coding, agentic and tool-use systems,...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Mistral Medium 3 is a high-performance enterprise-grade language model designed to deliver frontier-level capabilities at significantly reduced operational cost. It balances state-of-the-art reasoning and multimodal performance with 8× lower cost...
No provider description is available for this model yet.
No provider description is available for this model yet.
Mistral Medium 3.1 is an updated version of Mistral Medium 3, which is a high-performance enterprise-grade language model designed to deliver frontier-level capabilities at significantly reduced operational cost. It balances...
OpenAI o4-mini-high is the same model as [o4-mini](/openai/o4-mini) with reasoning_effort set to high. OpenAI o4-mini is a compact reasoning model in the o-series, optimized for fast, cost-efficient performance while retaining...
No provider description is available for this model yet.
Qwen-Plus, based on the Qwen2.5 foundation model, is a 131K context model with a balanced performance, speed, and cost combination.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Sonnet 4.6 is Anthropic's most capable Sonnet-class model yet, with frontier performance across coding, agents, and professional work. It excels at iterative development, complex codebase navigation, end-to-end project management with...
No provider description is available for this model yet.
GPT-5-Nano is the smallest and fastest variant in the GPT-5 system, optimized for developer tools, rapid interactions, and ultra-low latency environments. While limited in reasoning depth compared to its larger...
Qwen3-235B-A22B-Thinking-2507 is a high-performance, open-weight Mixture-of-Experts (MoE) language model optimized for complex reasoning tasks. It activates 22B of its 235B parameters per forward pass and natively supports up to 262,144...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
GPT-6 Astra Pro is the same underlying model as [GPT-6 Astra](https://openrouter.ai/openai/gpt-6-astra), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. Learn more in OpenAI's docs: https://developers.openai.com/api/docs/guides/reasoning#reasoning-mode
No provider description is available for this model yet.
MiniMax-M2.7 is a next-generation large language model designed for autonomous, real-world productivity and continuous improvement. Built to actively participate in its own evolution, M2.7 integrates advanced agentic capabilities through multi-agent...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| SpaceXAI: Grok 4.20x-ai/grok-4.20 | 2M | $1.25 | $2.5 | — | |||
| Anthropic: Claude Opus 4.7anthropic/claude-opus-4.7 | 1M | $5 | $25 | — | |||
| deepseek/deepseek-v4-flash-vision-expnovita/deepseek/deepseek-v4-flash-vision-exp | 1.04858M | $0.44 | $1.32 | — | |||
| deepseek/deepseek-v4-flash-0731novita/deepseek/deepseek-v4-flash-0731 | 1.04858M | $0.44 | $1.32 | — | |||
| Qwen: Qwen Plus 0728qwen/qwen-plus-2025-07-28 | 1M | $0.26 | $0.78 | — | |||
| mindai/macaron-v1-ventinovita/mindai/macaron-v1-venti | 1.04858M | $1.5 | $4.5 | — | |||
| minimax/minimax-m3novita/minimax/minimax-m3 | 1M | $0.3 | $1.2 | — | |||
| Qwen: Qwen3.5-9B (batch)qwen/qwen3.5-9b:batch | 262.144K | $0.17 | $0.25 | — | |||
| deepseek/deepseek-v4-flashnovita/deepseek/deepseek-v4-flash | 1.04858M | $0.14 | $0.28 | — | |||
| deepseek/deepseek-v4-pronovita/deepseek/deepseek-v4-pro | 1.04858M | $1.6 | $3.2 | — | |||
| Thinking Machines: Inkling (batch)thinkingmachines/inkling:batch | 524.288K | $1 | $4.05 | — | |||
| inclusionai/ling-3.0-flash-fastnovita/inclusionai/ling-3.0-flash-fast | 262.144K | $0.06 | $0.18 | — | |||
| qwen/qwen3.8-maxnovita/qwen/qwen3.8-max | 1M | $2 | $6 | — | |||
| inclusionai/ling-3.0-flashnovita/inclusionai/ling-3.0-flash | 262.144K | $0.06 | $0.18 | — | |||
| mindai/macaron-v1-tallnovita/mindai/macaron-v1-tall | 262.144K | $0.45 | $2.6 | — | |||
| stepfun/step-3.7-flashnovita/stepfun/step-3.7-flash | 262.144K | $0.2 | $1.15 | — | |||
| Mistral: Mistral Medium 3mistralai/mistral-medium-3 | 131.072K | $0.4 | $2 | — | |||
| nvidia/nemotron-3-nano-30b-a3bnovita/nvidia/nemotron-3-nano-30b-a3b | 262.144K | $0.05 | $0.2 | — | |||
| baidu/cobuddynovita/baidu/cobuddy | 131.072K | $0.28 | $1.13 | — | |||
| Mistral: Mistral Medium 3.1 (batch)mistralai/mistral-medium-3.1:batch | 131.072K | $0.2 | $1 | — | |||
| OpenAI: o4 Mini Highopenai/o4-mini-high | 200K | $1.1 | $4.4 | — | |||
| xiaomimimo/mimo-v2.5novita/xiaomimimo/mimo-v2.5 | 1.04858M | $0.168 | $0.336 | — | |||
| Qwen: Qwen-Plusqwen/qwen-plus | 1M | $0.26 | $0.78 | — | |||
| qwen/qwen3.7-maxnovita/qwen/qwen3.7-max | 1M | $1.25 | $3.75 | — | |||
| xiaomimimo/mimo-v2.5-pronovita/xiaomimimo/mimo-v2.5-pro | 1.04858M | $0.522 | $1.044 | — | |||
| qwen/qwen3.6-27bnovita/qwen/qwen3.6-27b | 262.144K | $0.6 | $3.6 | — | |||
| moonshotai/kimi-k2.6novita/moonshotai/kimi-k2.6 | 262.144K | $0.8 | $3.4 | — | |||
| zai-org/glm-5.1novita/zai-org/glm-5.1 | 204.8K | $1.38 | $4.4 | — | |||
| Anthropic: Claude Sonnet 4.6anthropic/claude-sonnet-4.6 | 1M | $3 | $15 | — | |||
| minimax/minimax-m2.7-highspeednovita/minimax/minimax-m2.7-highspeed | 204.8K | $0.6 | $2.4 | — | |||
| OpenAI: GPT-5 Nano (batch)openai/gpt-5-nano:batch | 400K | $0.025 | $0.2 | — | |||
| Qwen: Qwen3 235B A22B Thinking 2507qwen/qwen3-235b-a22b-thinking-2507 | 131.072K | $0.23 | $2.3 | — | |||
| zai-org/glm-5v-turbonovita/zai-org/glm-5v-turbo | 204.8K | $1.2 | $4 | — | |||
| google/gemma-4-26b-a4b-itnovita/google/gemma-4-26b-a4b-it | 262.144K | $0.13 | $0.4 | — | |||
| google/gemma-4-31b-itnovita/google/gemma-4-31b-it | 262.144K | $0.14 | $0.4 | — | |||
| OpenAI: GPT-6 Astra Pro (batch)openai/gpt-6-astra-pro:batch | 1.05M | $5 | $25 | — | |||
| zai-org/glm-5-turbonovita/zai-org/glm-5-turbo | 202.8K | $1.2 | $4 | — | |||
| MiniMax: MiniMax M2.7 (free)minimax/minimax-m2.7:free | 196.608K | Free | Free | — | |||
| minimax/minimax-m2.7novita/minimax/minimax-m2.7 | 204.8K | $0.3 | $1.2 | — | |||
| minimax/minimax-m2.5-highspeednovita/minimax/minimax-m2.5-highspeed | 204.8K | $0.6 | $2.4 | — | |||
| qwen/qwen3.5-27bnovita/qwen/qwen3.5-27b | 262.144K | $0.3 | $2.4 | — | |||
| qwen/qwen3.5-122b-a10bnovita/qwen/qwen3.5-122b-a10b | 262.144K | $0.4 | $3.2 | — | |||
| qwen/qwen3.5-35b-a3bnovita/qwen/qwen3.5-35b-a3b | 262.144K | $0.25 | $2 | — | |||
| databricks-claude-opus-4-6databricks/databricks-claude-opus-4-6 | 1M | $5 | $25 | — | |||
| databricks-claude-sonnet-4-6databricks/databricks-claude-sonnet-4-6 | 1M | $3 | $15 | — | |||
| databricks-gemini-3-1-flash-litedatabricks/databricks-gemini-3-1-flash-lite | 1.04858M | $0.312 | $1.875 | — | |||
| qwen/qwen3.5-397b-a17bnovita/qwen/qwen3.5-397b-a17b | 262.144K | $0.6 | $3.6 | — | |||
| minimax/minimax-m2.5novita/minimax/minimax-m2.5 | 204.8K | $0.3 | $1.2 | — | |||
| zai-org/glm-5novita/zai-org/glm-5 | 202.8K | $1 | $3.2 | — | |||
| qwen/qwen3-coder-nextnovita/qwen/qwen3-coder-next | 262.144K | $0.2 | $1.5 | — |