No provider description is available for this model yet.
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
MiniMax-M1 is a large-scale, open-weight reasoning model designed for extended context and high-efficiency inference. It leverages a hybrid Mixture-of-Experts (MoE) architecture paired with a custom "lightning attention" mechanism, allowing it...
No provider description is available for this model yet.
Grok 4.20 is a reasoning model from SpaceXAI with industry-leading speed and agentic tool calling capabilities. It combines the lowest hallucination rate on the market with strict prompt adherance, delivering...
Opus 4.7 is the next generation of Anthropic's Opus family, built for long-running, asynchronous agents. Building on the coding and agentic strengths of Opus 4.6, it delivers stronger performance on...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Qwen3.5-9B is a multimodal foundation model from the Qwen3.5 family, designed to deliver strong reasoning, coding, and visual understanding in an efficient 9B-parameter architecture. It uses a unified vision-language design...
No provider description is available for this model yet.
Inkling is an open-weight multimodal mixture-of-experts model from Thinking Machines Lab, with 41B active parameters out of 975B total. It is designed for general-purpose reasoning, coding, agentic and tool-use systems,...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Sonnet 4.6 is Anthropic's most capable Sonnet-class model yet, with frontier performance across coding, agents, and professional work. It excels at iterative development, complex codebase navigation, end-to-end project management with...
Gemini 3.1 Pro Preview is Google’s frontier reasoning model, delivering enhanced software engineering performance, improved agentic reliability, and more efficient token usage across complex workflows. Building on the multimodal foundation...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
GPT-6 Astra Pro is the same underlying model as [GPT-6 Astra](https://openrouter.ai/openai/gpt-6-astra), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. Learn more in OpenAI's docs: https://developers.openai.com/api/docs/guides/reasoning#reasoning-mode
No provider description is available for this model yet.
Claude Opus 5 is Anthropic’s flagship model for demanding reasoning, coding, and long-horizon agentic work. It is particularly strong at end-to-end software tasks, code review and bug finding, visual analysis...
MiniMax-M2.7 is a next-generation large language model designed for autonomous, real-world productivity and continuous improvement. Built to actively participate in its own evolution, M2.7 integrates advanced agentic capabilities through multi-agent...
Gemini 3.8 Flash is Google's most intelligent Flash model with significant gains from 3.7 Flash across software engineering, agentic tasks, and multi-step reasoning.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| ministral-3b-latestmistral/ministral-3b-latest | 131.072K | $0.1 | $0.1 | — | |||
| mistral-medium-3mistral/mistral-medium-3 | 262.144K | $1.5 | $7.5 | — | |||
| zai-org/GLM-5.3-Flashtogether_ai/zai-org/glm-5.3-flash | 1.04858M | $0.15 | $0.5 | — | |||
| glm-5.3zai/glm-5.3 | 1M | $1.4 | $4.4 | — | |||
| minimax-m3tencent/minimax-m3 | 1M | $0.3 | $1.2 | — | |||
| zai-org/glm-5.3novita/zai-org/glm-5.3 | 1.04858M | $1.4 | $4.4 | — | |||
| deepseek/deepseek-v4-pro-0813novita/deepseek/deepseek-v4-pro-0813 | 1.04858M | $1.32 | $3.96 | — | |||
| moonshotai/kimi-k3novita/moonshotai/kimi-k3 | 1.04858M | $3 | $15 | — | |||
| tencent/hy3novita/tencent/hy3 | 262.144K | $0.14 | $0.58 | — | |||
| zai-org/glm-5.2novita/zai-org/glm-5.2 | 1.04858M | $1.4 | $4.4 | — | |||
| MiniMax: MiniMax M1minimax/minimax-m1 | 1M | $0.55 | $2.2 | — | |||
| moonshotai/kimi-k2.7-codenovita/moonshotai/kimi-k2.7-code | 262.144K | $0.95 | $4 | — | |||
| SpaceXAI: Grok 4.20x-ai/grok-4.20 | 2M | $1.25 | $2.5 | — | |||
| Anthropic: Claude Opus 4.7anthropic/claude-opus-4.7 | 1M | $5 | $25 | — | |||
| deepseek/deepseek-v4-flash-vision-expnovita/deepseek/deepseek-v4-flash-vision-exp | 1.04858M | $0.44 | $1.32 | — | |||
| deepseek/deepseek-v4-flash-0731novita/deepseek/deepseek-v4-flash-0731 | 1.04858M | $0.44 | $1.32 | — | |||
| mindai/macaron-v1-ventinovita/mindai/macaron-v1-venti | 1.04858M | $1.5 | $4.5 | — | |||
| Qwen: Qwen3.5-9B (batch)qwen/qwen3.5-9b:batch | 262.144K | $0.17 | $0.25 | — | |||
| minimax/minimax-m3novita/minimax/minimax-m3 | 1M | $0.3 | $1.2 | — | |||
| Thinking Machines: Inkling (batch)thinkingmachines/inkling:batch | 524.288K | $1 | $4.05 | — | |||
| deepseek/deepseek-v4-flashnovita/deepseek/deepseek-v4-flash | 1.04858M | $0.14 | $0.28 | — | |||
| deepseek/deepseek-v4-pronovita/deepseek/deepseek-v4-pro | 1.04858M | $1.6 | $3.2 | — | |||
| inclusionai/ling-3.0-flash-fastnovita/inclusionai/ling-3.0-flash-fast | 262.144K | $0.06 | $0.18 | — | |||
| qwen/qwen3.8-maxnovita/qwen/qwen3.8-max | 1M | $2 | $6 | — | |||
| inclusionai/ling-3.0-flashnovita/inclusionai/ling-3.0-flash | 262.144K | $0.06 | $0.18 | — | |||
| mindai/macaron-v1-tallnovita/mindai/macaron-v1-tall | 262.144K | $0.45 | $2.6 | — | |||
| stepfun/step-3.7-flashnovita/stepfun/step-3.7-flash | 262.144K | $0.2 | $1.15 | — | |||
| nvidia/nemotron-3-nano-30b-a3bnovita/nvidia/nemotron-3-nano-30b-a3b | 262.144K | $0.05 | $0.2 | — | |||
| baidu/cobuddynovita/baidu/cobuddy | 131.072K | $0.28 | $1.13 | — | |||
| xiaomimimo/mimo-v2.5novita/xiaomimimo/mimo-v2.5 | 1.04858M | $0.168 | $0.336 | — | |||
| qwen/qwen3.7-maxnovita/qwen/qwen3.7-max | 1M | $1.25 | $3.75 | — | |||
| xiaomimimo/mimo-v2.5-pronovita/xiaomimimo/mimo-v2.5-pro | 1.04858M | $0.522 | $1.044 | — | |||
| qwen/qwen3.6-27bnovita/qwen/qwen3.6-27b | 262.144K | $0.6 | $3.6 | — | |||
| moonshotai/kimi-k2.6novita/moonshotai/kimi-k2.6 | 262.144K | $0.8 | $3.4 | — | |||
| zai-org/glm-5.1novita/zai-org/glm-5.1 | 204.8K | $1.38 | $4.4 | — | |||
| minimax/minimax-m2.7-highspeednovita/minimax/minimax-m2.7-highspeed | 204.8K | $0.6 | $2.4 | — | |||
| Anthropic: Claude Sonnet 4.6anthropic/claude-sonnet-4.6 | 1M | $3 | $15 | — | |||
| Google: Gemini 3.1 Pro Preview (batch)google/gemini-3.1-pro-preview:batch | 1.04858M | $1 | $6 | — | |||
| zai-org/glm-5v-turbonovita/zai-org/glm-5v-turbo | 204.8K | $1.2 | $4 | — | |||
| google/gemma-4-26b-a4b-itnovita/google/gemma-4-26b-a4b-it | 262.144K | $0.13 | $0.4 | — | |||
| google/gemma-4-31b-itnovita/google/gemma-4-31b-it | 262.144K | $0.14 | $0.4 | — | |||
| OpenAI: GPT-6 Astra Pro (batch)openai/gpt-6-astra-pro:batch | 1.05M | $5 | $25 | — | |||
| zai-org/glm-5-turbonovita/zai-org/glm-5-turbo | 202.8K | $1.2 | $4 | — | |||
| Claude Opus 5 (batch)anthropic/claude-opus-5:batch | 1M | $2.5 | $12.5 | — | |||
| MiniMax: MiniMax M2.7 (free)minimax/minimax-m2.7:free | 196.608K | Free | Free | — | |||
| Google: Gemini 3.8 Flash (batch)google/gemini-3.8-flash:batch | 1.04858M | $0.375 | $1.875 | — | |||
| minimax/minimax-m2.7novita/minimax/minimax-m2.7 | 204.8K | $0.3 | $1.2 | — | |||
| minimax/minimax-m2.5-highspeednovita/minimax/minimax-m2.5-highspeed | 204.8K | $0.6 | $2.4 | — | |||
| qwen/qwen3.5-27bnovita/qwen/qwen3.5-27b | 262.144K | $0.3 | $2.4 | — | |||
| qwen/qwen3.5-122b-a10bnovita/qwen/qwen3.5-122b-a10b | 262.144K | $0.4 | $3.2 | — |