3,455 models

No provider description is available for this model yet.

zai/glm-4.5 128K context $0.6/M input $2.2/M output
Open weights

Granite 4.2 8B is a dense reasoning model from IBM. It is suited for mathematics, code generation, multilingual dialogue, and agentic workflows that need multi-step reasoning. It supports full, low-effort,...

ibm-granite/granite-4.2-8b 131.072K context $0.06/M input $0.25/M output

GPT-5.4 is OpenAI’s latest frontier model, unifying the Codex and GPT lines into a single system. It features a 1M+ token context window (922K input, 128K output) with support for...

openai/gpt-5.4:batch 1.05M context $1.25/M input $7.5/M output

Gemini 3 Flash Preview is a high speed, high value thinking model designed for agentic workflows, multi turn chat, and coding assistance. It delivers near Pro level reasoning and tool...

google/gemini-3-flash-preview:batch 1.04858M context $0.25/M input $1.5/M output

No provider description is available for this model yet.

azure_ai/grok-4-1-fast-reasoning 131.072K context $0.2/M input $0.5/M output

GPT-5 Pro is OpenAI’s most advanced model, offering major improvements in reasoning, code quality, and user experience. It is optimized for complex tasks that require step-by-step reasoning, instruction following, and...

openai/gpt-5-pro:batch 400K context $7.5/M input $60/M output

No provider description is available for this model yet.

gmi/minimaxai/minimax-m2.1 196.608K context $0.3/M input $1.2/M output

Mistral Large 3 2512 is Mistral’s most capable model to date, featuring a sparse mixture-of-experts architecture with 41B active parameters (675B total), and released under the Apache 2.0 license.

mistralai/mistral-large-2512 262.144K context $0.5/M input $1.5/M output

No provider description is available for this model yet.

databricks/databricks-glm-5-2 1M context $1.4/M input $4.4/M output

GPT-5.6 Terra Pro is the same underlying model as [GPT-5.6 Terra](https://openrouter.ai/openai/gpt-5.6-terra), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. Learn more in OpenAI's docs: https://developers.openai.com/api/docs/guides/reasoning#reasoning-mode

openai/gpt-5.6-terra-pro:batch 1.05M context $1/M input $6/M output

No provider description is available for this model yet.

databricks/databricks-kimi-k3 1M context $3/M input $15/M output

No provider description is available for this model yet.

mistral/ministral-14b-2512 262.144K context $0.2/M input $0.2/M output

No provider description is available for this model yet.

mistral/ministral-14b-latest 262.144K context $0.2/M input $0.2/M output

No provider description is available for this model yet.

mistral/ministral-3b-2512 131.072K context $0.1/M input $0.1/M output

No provider description is available for this model yet.

mistral/ministral-3b-latest 131.072K context $0.1/M input $0.1/M output

No provider description is available for this model yet.

mistral/mistral-medium-3 262.144K context $1.5/M input $7.5/M output

No provider description is available for this model yet.

mistral/voxtral-small-2507 32.768K context $0.1/M input $0.4/M output

No provider description is available for this model yet.

together_ai/zai-org/glm-5.3-flash 1.04858M context $0.15/M input $0.5/M output

MiniMax-M1 is a large-scale, open-weight reasoning model designed for extended context and high-efficiency inference. It leverages a hybrid Mixture-of-Experts (MoE) architecture paired with a custom "lightning attention" mechanism, allowing it...

minimax/minimax-m1 1M context $0.55/M input $2.2/M output

No provider description is available for this model yet.

zai/glm-5.3 1M context $1.4/M input $4.4/M output

No provider description is available for this model yet.

tencent/minimax-m3 1M context $0.3/M input $1.2/M output

No provider description is available for this model yet.

novita/zai-org/glm-5.3 1.04858M context $1.4/M input $4.4/M output

No provider description is available for this model yet.

novita/moonshotai/kimi-k3 1.04858M context $3/M input $15/M output

No provider description is available for this model yet.

novita/tencent/hy3 262.144K context $0.14/M input $0.58/M output

No provider description is available for this model yet.

novita/zai-org/glm-5.2 1.04858M context $1.4/M input $4.4/M output

The latest GPT-4 Turbo model with vision capabilities. Vision requests can now use JSON mode and function calling. Training data: up to December 2023.

openai/gpt-4-turbo:batch 128K context $5/M input $15/M output

No provider description is available for this model yet.

novita/moonshotai/kimi-k2.7-code 262.144K context $0.95/M input $4/M output

No provider description is available for this model yet.

novita/mindai/macaron-v1-venti 1.04858M context $1.5/M input $4.5/M output

No provider description is available for this model yet.

azure_ai/grok-4-fast-reasoning 131.072K context $0.2/M input $0.5/M output

No provider description is available for this model yet.

novita/minimax/minimax-m3 1M context $0.3/M input $1.2/M output

No provider description is available for this model yet.

novita/deepseek/deepseek-v4-flash 1.04858M context $0.14/M input $0.28/M output

No provider description is available for this model yet.

novita/deepseek/deepseek-v4-pro 1.04858M context $1.6/M input $3.2/M output

Inkling is an open-weight multimodal mixture-of-experts model from Thinking Machines Lab, with 41B active parameters out of 975B total. It is designed for general-purpose reasoning, coding, agentic and tool-use systems,...

thinkingmachines/inkling:batch 524.288K context $1/M input $4.05/M output

No provider description is available for this model yet.

gmi/moonshotai/kimi-k2-thinking 262.144K context $0.8/M input $1.2/M output

No provider description is available for this model yet.

novita/qwen/qwen3.8-max 1M context $2/M input $6/M output

No provider description is available for this model yet.

novita/inclusionai/ling-3.0-flash 262.144K context $0.06/M input $0.18/M output

No provider description is available for this model yet.

novita/mindai/macaron-v1-tall 262.144K context $0.45/M input $2.6/M output