No provider description is available for this model yet.
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
Granite 4.2 8B is a dense reasoning model from IBM. It is suited for mathematics, code generation, multilingual dialogue, and agentic workflows that need multi-step reasoning. It supports full, low-effort,...
GPT-5.4 is OpenAI’s latest frontier model, unifying the Codex and GPT lines into a single system. It features a 1M+ token context window (922K input, 128K output) with support for...
Gemini 3 Flash Preview is a high speed, high value thinking model designed for agentic workflows, multi turn chat, and coding assistance. It delivers near Pro level reasoning and tool...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
GPT-5 Pro is OpenAI’s most advanced model, offering major improvements in reasoning, code quality, and user experience. It is optimized for complex tasks that require step-by-step reasoning, instruction following, and...
No provider description is available for this model yet.
No provider description is available for this model yet.
Mistral Large 3 2512 is Mistral’s most capable model to date, featuring a sparse mixture-of-experts architecture with 41B active parameters (675B total), and released under the Apache 2.0 license.
No provider description is available for this model yet.
GPT-5.6 Terra Pro is the same underlying model as [GPT-5.6 Terra](https://openrouter.ai/openai/gpt-5.6-terra), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. Learn more in OpenAI's docs: https://developers.openai.com/api/docs/guides/reasoning#reasoning-mode
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
MiniMax-M1 is a large-scale, open-weight reasoning model designed for extended context and high-efficiency inference. It leverages a hybrid Mixture-of-Experts (MoE) architecture paired with a custom "lightning attention" mechanism, allowing it...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
The latest GPT-4 Turbo model with vision capabilities. Vision requests can now use JSON mode and function calling. Training data: up to December 2023.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Inkling is an open-weight multimodal mixture-of-experts model from Thinking Machines Lab, with 41B active parameters out of 975B total. It is designed for general-purpose reasoning, coding, agentic and tool-use systems,...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| glm-4.5zai/glm-4.5 | 128K | $0.6 | $2.2 | — | |||
| IBM: Granite 4.2 8Bibm-granite/granite-4.2-8b | 131.072K | $0.06 | $0.25 | — | |||
| OpenAI: GPT-5.4 (batch)openai/gpt-5.4:batch | 1.05M | $1.25 | $7.5 | — | |||
| Google: Gemini 3 Flash Preview (batch)google/gemini-3-flash-preview:batch | 1.04858M | $0.25 | $1.5 | — | |||
| grok-4-1-fast-reasoningazure_ai/grok-4-1-fast-reasoning | 131.072K | $0.2 | $0.5 | — | |||
| moonshotai/kimi-k2-thinking-maasvertex_ai-moonshot_models/moonshotai/kimi-k2-thinking-maas | 256K | $0.6 | $2.5 | — | |||
| minimaxai/minimax-m2-maasvertex_ai-minimax_models/minimaxai/minimax-m2-maas | 196.608K | $0.3 | $1.2 | — | |||
| OpenAI: GPT-5 Pro (batch)openai/gpt-5-pro:batch | 400K | $7.5 | $60 | — | |||
| MiniMaxAI/MiniMax-M2.1gmi/minimaxai/minimax-m2.1 | 196.608K | $0.3 | $1.2 | — | |||
| meta/llama3-8b-instruct-maasvertex_ai-llama_models/meta/llama3-8b-instruct-maas | 32K | — | — | — | |||
| Mistral: Mistral Large 3 2512mistralai/mistral-large-2512 | 262.144K | $0.5 | $1.5 | — | |||
| databricks-glm-5-2databricks/databricks-glm-5-2 | 1M | $1.4 | $4.4 | — | |||
| OpenAI: GPT-5.6 Terra Pro (batch)openai/gpt-5.6-terra-pro:batch | 1.05M | $1 | $6 | — | |||
| databricks-kimi-k3databricks/databricks-kimi-k3 | 1M | $3 | $15 | — | |||
| accounts/fireworks/models/deepseek-v4-pro-0813fireworks_ai/accounts/fireworks/models/deepseek-v4-pro-0813 | 1.04858M | $1.32 | $3.96 | — | |||
| ministral-14b-2512mistral/ministral-14b-2512 | 262.144K | $0.2 | $0.2 | — | |||
| ministral-14b-latestmistral/ministral-14b-latest | 262.144K | $0.2 | $0.2 | — | |||
| ministral-3b-2512mistral/ministral-3b-2512 | 131.072K | $0.1 | $0.1 | — | |||
| ministral-3b-latestmistral/ministral-3b-latest | 131.072K | $0.1 | $0.1 | — | |||
| mistral-medium-3mistral/mistral-medium-3 | 262.144K | $1.5 | $7.5 | — | |||
| voxtral-small-2507mistral/voxtral-small-2507 | 32.768K | $0.1 | $0.4 | — | |||
| zai-org/GLM-5.3-Flashtogether_ai/zai-org/glm-5.3-flash | 1.04858M | $0.15 | $0.5 | — | |||
| MiniMax: MiniMax M1minimax/minimax-m1 | 1M | $0.55 | $2.2 | — | |||
| glm-5.3zai/glm-5.3 | 1M | $1.4 | $4.4 | — | |||
| minimax-m3tencent/minimax-m3 | 1M | $0.3 | $1.2 | — | |||
| zai-org/glm-5.3novita/zai-org/glm-5.3 | 1.04858M | $1.4 | $4.4 | — | |||
| deepseek/deepseek-v4-pro-0813novita/deepseek/deepseek-v4-pro-0813 | 1.04858M | $1.32 | $3.96 | — | |||
| moonshotai/kimi-k3novita/moonshotai/kimi-k3 | 1.04858M | $3 | $15 | — | |||
| tencent/hy3novita/tencent/hy3 | 262.144K | $0.14 | $0.58 | — | |||
| zai-org/glm-5.2novita/zai-org/glm-5.2 | 1.04858M | $1.4 | $4.4 | — | |||
| OpenAI: GPT-4 Turbo (batch)openai/gpt-4-turbo:batch | 128K | $5 | $15 | — | |||
| meta/llama3-70b-instruct-maasvertex_ai-llama_models/meta/llama3-70b-instruct-maas | 32K | — | — | — | |||
| moonshotai/kimi-k2.7-codenovita/moonshotai/kimi-k2.7-code | 262.144K | $0.95 | $4 | — | |||
| grok-4-1-fast-non-reasoningazure_ai/grok-4-1-fast-non-reasoning | 131.072K | $0.2 | $0.5 | — | |||
| meta/llama3-405b-instruct-maasvertex_ai-llama_models/meta/llama3-405b-instruct-maas | 32K | — | — | — | |||
| deepseek/deepseek-v4-flash-vision-expnovita/deepseek/deepseek-v4-flash-vision-exp | 1.04858M | $0.44 | $1.32 | — | |||
| deepseek/deepseek-v4-flash-0731novita/deepseek/deepseek-v4-flash-0731 | 1.04858M | $0.44 | $1.32 | — | |||
| meta/llama-4-scout-17b-16e-instruct-maasvertex_ai-llama_models/meta/llama-4-scout-17b-16e-instruct-maas | 10M | $0.25 | $0.7 | — | |||
| mindai/macaron-v1-ventinovita/mindai/macaron-v1-venti | 1.04858M | $1.5 | $4.5 | — | |||
| grok-4-fast-reasoningazure_ai/grok-4-fast-reasoning | 131.072K | $0.2 | $0.5 | — | |||
| minimax/minimax-m3novita/minimax/minimax-m3 | 1M | $0.3 | $1.2 | — | |||
| deepseek/deepseek-v4-flashnovita/deepseek/deepseek-v4-flash | 1.04858M | $0.14 | $0.28 | — | |||
| deepseek/deepseek-v4-pronovita/deepseek/deepseek-v4-pro | 1.04858M | $1.6 | $3.2 | — | |||
| Thinking Machines: Inkling (batch)thinkingmachines/inkling:batch | 524.288K | $1 | $4.05 | — | |||
| moonshotai/Kimi-K2-Thinkinggmi/moonshotai/kimi-k2-thinking | 262.144K | $0.8 | $1.2 | — | |||
| inclusionai/ling-3.0-flash-fastnovita/inclusionai/ling-3.0-flash-fast | 262.144K | $0.06 | $0.18 | — | |||
| qwen/qwen3.8-maxnovita/qwen/qwen3.8-max | 1M | $2 | $6 | — | |||
| inclusionai/ling-3.0-flashnovita/inclusionai/ling-3.0-flash | 262.144K | $0.06 | $0.18 | — | |||
| mindai/macaron-v1-tallnovita/mindai/macaron-v1-tall | 262.144K | $0.45 | $2.6 | — | |||
| eu-south-1/qwen.qwen3-coder-nextbedrock/eu-south-1/qwen.qwen3-coder-next | 262.144K | $0.6 | $1.44 | — |