Gemma 4 26B A4B IT is an instruction-tuned Mixture-of-Experts (MoE) model from Google DeepMind. Despite 25.2B total parameters, only 3.8B activate per token during inference — delivering near-31B quality at...

google/gemma-4-26b-a4b-it:free 262.144K context Free input Free output

Gemma 4 31B Instruct is Google DeepMind's 30.7B dense multimodal model supporting text and image input with text output. Features a 256K token context window, configurable thinking/reasoning mode, native function...

google/gemma-4-31b-it:free 262.144K context Free input Free output

GPT-5 Image Mini combines OpenAI's advanced language capabilities, powered by [GPT-5 Mini](https://openrouter.ai/openai/gpt-5-mini), with GPT Image 1 Mini for efficient image generation. This natively multimodal model features superior instruction following, text...

openai/gpt-5-image-mini 400K context $2.5/M input $2/M output

GLM-5 Turbo is a new model from Z.ai designed for fast inference and strong performance in agent-driven environments such as OpenClaw scenarios. It is deeply optimized for real-world agent workflows...

z-ai/glm-5-turbo 202.752K context $1.2/M input $4/M output

The Qwen3.5 native vision-language Flash models are built on a hybrid architecture that integrates a linear attention mechanism with a sparse mixture-of-experts model, achieving higher inference efficiency. Compared to the...

qwen/qwen3.5-flash-02-23 1M context $0.065/M input $0.26/M output

The Qwen3.5 series 397B-A17B native vision-language model is built on a hybrid architecture that integrates a linear attention mechanism with a sparse mixture-of-experts model, achieving higher inference efficiency. It delivers...

qwen/qwen3.5-397b-a17b 262.144K context $0.55/M input $3.5/M output

MiniMax-M2.5 is a SOTA large language model designed for real-world productivity. Trained in a diverse range of complex real-world digital working environments, M2.5 builds upon the coding expertise of M2.1...

minimax/minimax-m2.5 200K context $0.27/M input $1.08/M output

Qwen3-Coder-Next is an open-weight causal language model optimized for coding agents and local development workflows. It uses a sparse MoE design with 80B total parameters and only 3B activated per...

qwen/qwen3-coder-next 262.144K context $0.12/M input $0.8/M output

The simplest way to get free inference. openrouter/free is a router that selects free models at random from the models available on OpenRouter. The router smartly filters for models that...

openrouter/free 200K context Free input Free output

Opus 4.6 is Anthropic’s strongest model for coding and long-running professional tasks. It is built for agents that operate across entire workflows rather than single prompts, making it especially effective...

anthropic/claude-opus-4.6 1M context $5/M input $25/M output

NVIDIA Nemotron 3 Nano 30B A3B is a small language MoE model with highest compute efficiency and accuracy for developers to build specialized agentic AI systems. The model is fully...

nvidia/nemotron-3-nano-30b-a3b:free 256K context Free input Free output

The largest model in the Ministral 3 family, Ministral 3 14B offers frontier capabilities and performance comparable to its larger Mistral Small 3.2 24B counterpart. A powerful and efficient language...

mistralai/ministral-14b-2512 262.144K context $0.2/M input $0.2/M output

Claude Opus 4.5 is Anthropic’s frontier reasoning model optimized for complex software engineering, agentic workflows, and long-horizon computer use. It offers strong multimodal capabilities, competitive performance across real-world coding and...

anthropic/claude-opus-4.5 200K context $5/M input $25/M output

The preview GPT-4 model with improved instruction following, JSON mode, reproducible outputs, parallel function calling, and more. Training data: up to Dec 2023. **Note:** heavily rate limited by OpenAI while...

openai/gpt-4-turbo-preview 128K context $10/M input $30/M output

No provider description is available for this model yet.

nebius/deepseek-ai/deepseek-v4-pro 1.04858M context $1.75/M input $3.5/M output

No provider description is available for this model yet.

nebius/minimaxai/minimax-m2.5 196.608K context $0.3/M input $1.2/M output

No provider description is available for this model yet.

nebius/minimaxai/minimax-m3 1.04858M context $0.3/M input $1.2/M output

No provider description is available for this model yet.

nebius/moonshotai/kimi-k2.6 262.144K context $0.95/M input $4/M output

No provider description is available for this model yet.

nebius/moonshotai/kimi-k2.7-code 262.144K context $0.95/M input $4/M output

No provider description is available for this model yet.

nebius/moonshotai/kimi-k3 1.024M context $3/M input $15/M output

No provider description is available for this model yet.

nebius/nousresearch/hermes-4-405b 131.072K context $1/M input $3/M output

No provider description is available for this model yet.

nebius/nousresearch/hermes-4-70b 131.072K context $0.13/M input $0.4/M output

No provider description is available for this model yet.

nebius/nvidia/nemotron-3-nano-omni 262.144K context $0.06/M input $0.24/M output

No provider description is available for this model yet.

nebius/openai/gpt-oss-120b 131.072K context $0.15/M input $0.6/M output

No provider description is available for this model yet.

nebius/openbmb/minicpm-v-4_5 32K context $0.658/M input $1.11/M output

No provider description is available for this model yet.

nebius/qwen/qwen3.5-397b-a17b 262.144K context $0.6/M input $3.6/M output

No provider description is available for this model yet.

nebius/zai-org/glm-5.1 202.752K context $1.4/M input $4.4/M output

No provider description is available for this model yet.

nebius/zai-org/glm-5.2 1.04858M context $1.4/M input $4.4/M output

No provider description is available for this model yet.

nebius/zai-org/glm-5.3-flash 1.024M context $0.15/M input $0.5/M output

No provider description is available for this model yet.

watsonx/bigscience/mt0-xxl 4.096K context $1.908/M input $1.908/M output

No provider description is available for this model yet.

anthropic/claude-mythos-5-1 1M context $10/M input $50/M output

No provider description is available for this model yet.

gemini/lyria-3.5-clip-preview 131.072K context Input not listed Output not listed

No provider description is available for this model yet.

gemini/lyria-3.5-pro-preview 131.072K context Input not listed Output not listed

The latest GPT-4 Turbo model with vision capabilities. Vision requests can now use JSON mode and function calling. Training data: up to December 2023.

openai/gpt-4-turbo:batch 128K context $5/M input $15/M output