3,233 models

No provider description is available for this model yet.

friendliai/zai-org/glm-5.3-flash 1.04858M context $0.15/M input $0.5/M output

No provider description is available for this model yet.

friendliai/zai-org/glm-5.3 1.04858M context $1.26/M input $3.96/M output

No provider description is available for this model yet.

gigachat/gigachat-2 128K context Input not listed Output not listed

No provider description is available for this model yet.

zai/glm-4.7-flash 200K context Input not listed Output not listed

No provider description is available for this model yet.

zai/glm-5.2 1M context $1.4/M input $4.4/M output

GPT Chat Latest points to OpenAI's stable API alias `chat-latest` that always resolves to the latest Instant chat model used in ChatGPT. As OpenAI rolls out new Instant model updates...

openai/gpt-chat-latest 400K context $5/M input $30/M output

No provider description is available for this model yet.

cerebras/gemma-4-31b 131.072K context $0.99/M input $1.49/M output

Transform your natural language requests into structured OpenRouter API request objects. Describe what you want to accomplish with AI models, and Body Builder will construct the appropriate API calls. Example:...

openrouter/bodybuilder 128K context Input not listed Output not listed

Fusion turns your prompt into a small multi-model deliberation. A panel of expert models (see below) analyzes your prompt in parallel with web search and web fetch enabled, then a...

openrouter/fusion 1M context Input not listed Output not listed

MiniMax-M3 is a multimodal foundation model from MiniMax. It supports text, image, and video inputs with text output, a 1M-token context window, and is suited for long-horizon agentic work, coding,...

minimax/minimax-m3:free 1.04858M context Free input Free output

Nex-N2.5 is an agentic model built to turn goals into working, verified outcomes. Its core strength is agentic coding within a visual feedback loop: it can explore codebases, implement multi-file...

nex-agi/nex-n2.5-pro:free 262.144K context Free input Free output

No provider description is available for this model yet.

zai/glm-4.7 200K context $0.6/M input $2.2/M output

No provider description is available for this model yet.

zai/glm-5-code 200K context $1.2/M input $5/M output

No provider description is available for this model yet.

zai/glm-5.1 200K context $1.4/M input $4.4/M output

No provider description is available for this model yet.

zai/glm-5 200K context $1/M input $3.2/M output

No provider description is available for this model yet.

gmi/zai-org/glm-4.7-fp8 202.752K context $0.4/M input $2/M output

No provider description is available for this model yet.

bedrock_converse/zai.glm-5 200K context $1/M input $3.2/M output

No provider description is available for this model yet.

bedrock_converse/zai.glm-4.7 200K context $0.6/M input $2.2/M output

No provider description is available for this model yet.

xai/grok-code-fast-1-0825 256K context $0.2/M input $1.5/M output

Muse Glimmer 30B is a dense, open-weight multimodal model from Meta Superintelligence Labs, distilled from Muse Spark and optimized for autonomous agents on consumer hardware. It is suited for long-horizon...

meta/muse-glimmer-30b:batch 131.072K context $0.175/M input $0.75/M output

No provider description is available for this model yet.

xai/grok-code-fast-1 256K context $0.2/M input $1.5/M output

No provider description is available for this model yet.

xai/grok-code-fast 256K context $0.2/M input $1.5/M output

DeepSeek V4 Flash 0731 is a sparse mixture-of-experts model from DeepSeek, with 13B active parameters out of 284B total. This re-post-trained revision is suited for coding, reasoning, and agent workflows....

deepseek/deepseek-v4-flash-0731:batch 1.04858M context $0.11/M input $0.33/M output

No provider description is available for this model yet.

xai/grok-beta 131.072K context $5/M input $15/M output

No provider description is available for this model yet.

xai/grok-4.5-latest 500K context $2/M input $6/M output

No provider description is available for this model yet.

xai/grok-4.3-latest 1M context $1.25/M input $2.5/M output

GLM-5.3-Flash is a native multimodal model from Z.ai. It is suited for efficient coding and long-horizon agent tasks. Its hybrid sparse and linear attention architecture maintains accurate long-context behavior while...

z-ai/glm-5.3-flash 1.04858M context $0.15/M input $0.5/M output

Gemini 3.1 Flash Lite is Google’s GA high-efficiency multimodal model optimized for low-latency, high-volume workloads. It supports text, image, video, audio, and PDF inputs, and is designed for lightweight agentic...

google/gemini-3.1-flash-lite:batch 1.04858M context $0.125/M input $0.75/M output
Open weights

Granite 4.1 8B is a dense, decoder-only 8-billion-parameter language model from IBM, part of the Granite 4.1 family. It supports a 131K-token context window and is designed for enterprise tasks...

ibm-granite/granite-4.1-8b 131.072K context $0.05/M input $0.1/M output

No provider description is available for this model yet.

xai/grok-4-1-fast-non-reasoning 2M context $0.2/M input $0.5/M output