No provider description is available for this model yet.

gemini/gemini-pro-latest 1.04858M context $1.25/M input $10/M output

No provider description is available for this model yet.

bedrock_mantle/google.gemma-4-31b 256K context $0.14/M input $0.4/M output

No provider description is available for this model yet.

snowflake/claude-4-sonnet 200K context $3/M input $15/M output

No provider description is available for this model yet.

bedrock/us-east-1/zai.glm-5 200K context $1/M input $3.2/M output

No provider description is available for this model yet.

bedrock/us-west-2/zai.glm-5 200K context $1/M input $3.2/M output

No provider description is available for this model yet.

snowflake/claude-sonnet-4-5 200K context $3/M input $15/M output

No provider description is available for this model yet.

snowflake/claude-sonnet-4-6 200K context $3/M input $15/M output

No provider description is available for this model yet.

snowflake/claude-4-opus 200K context $5/M input $25/M output

No provider description is available for this model yet.

snowflake/claude-haiku-4-5 200K context $1/M input $5/M output

No provider description is available for this model yet.

snowflake/claude-3-7-sonnet 200K context $3/M input $15/M output

No provider description is available for this model yet.

snowflake/openai-gpt-4.1 300K context $2/M input $8/M output

No provider description is available for this model yet.

snowflake/openai-gpt-5 300K context $1.25/M input $10/M output

No provider description is available for this model yet.

snowflake/openai-gpt-5-mini 1M context $0.3/M input $1.2/M output

No provider description is available for this model yet.

snowflake/openai-gpt-5-nano 5M context $0.15/M input $0.6/M output

No provider description is available for this model yet.

tensormesh/qwen/qwen3.6-27b-fp8 262.144K context $0.32/M input $3.2/M output

No provider description is available for this model yet.

tencent/deepseek-v4-pro 1M context $0.435/M input $0.87/M output

No provider description is available for this model yet.

tencent/deepseek-v4-flash 1M context $0.14/M input $0.28/M output

No provider description is available for this model yet.

pinstripes/ps/minimax-m2.7 1.00019M context $0.255/M input $0.55/M output

Claude Fable 5.1 improves on Claude Fable 5 across the board, with the biggest gains in agentic coding, long-running agentic workflows, and knowledge work: long code refactors, front-end and visual...

anthropic/claude-fable-5.1 1M context $10/M input $50/M output

Fugu Max is the cost-performance model in Sakana AI's Fugu family. Rather than a single monolithic model, Fugu is a learned multi-agent orchestration system: a language model trained to route...

sakana/fugu-max 1M context $2/M input $6/M output

GPT-5.4 Pro is OpenAI's most advanced model, building on GPT-5.4's unified architecture with enhanced reasoning capabilities for complex, high-stakes tasks. It features a 1M+ token context window (922K input, 128K...

openai/gpt-5.4-pro:batch 1.05M context $15/M input $90/M output

No provider description is available for this model yet.

openrouter/openai/gpt-5.6-sol 1.05M context $2/M input $10/M output

NVIDIA Nemotron 3.5 Lightning is an open mixture-of-experts model from NVIDIA, with 3B active parameters out of 30B total. It is suited for high-throughput agentic workloads and specialized tasks that...

nvidia/nemotron-3.5-lightning:free 1M context Free input Free output

GLM-5.3 is a large-scale reasoning model from Z.ai, built for complex software engineering and long-horizon agent tasks. It supports text input and output with a 1M-token context window, and improves...

z-ai/glm-5.3:batch 1.04858M context $0.7/M input $2.2/M output

GLM-5.3-Flash is a native multimodal model from Z.ai. It is suited for efficient coding and long-horizon agent tasks. Its hybrid sparse and linear attention architecture maintains accurate long-context behavior while...

z-ai/glm-5.3-flash:batch 1.04858M context $0.075/M input $0.25/M output

Gemini 3.5 Flash Lite is a high-efficiency model from Google with upgraded agentic capabilities. It is suited for subagents that execute focused tasks within complex, multi-agent workflows.

google/gemini-3.5-flash-lite:batch 1.04858M context $0.15/M input $1.25/M output

GPT-5.6 Sol is the flagship model in OpenAI's GPT-5.6 series. It is suited for complex reasoning, coding, and agentic workflows, and is particularly strong at command-line and multi-step coding tasks...

openai/gpt-5.6-sol:batch 1.05M context $1/M input $5/M output

GPT-5.4 is OpenAI’s latest frontier model, unifying the Codex and GPT lines into a single system. It features a 1M+ token context window (922K input, 128K output) with support for...

openai/gpt-5.4:batch 1.05M context $1.25/M input $7.5/M output

No provider description is available for this model yet.

databricks/databricks-glm-5-2 1M context $1.4/M input $4.4/M output

GPT-5.6 Terra Pro is the same underlying model as [GPT-5.6 Terra](https://openrouter.ai/openai/gpt-5.6-terra), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. Learn more in OpenAI's docs: https://developers.openai.com/api/docs/guides/reasoning#reasoning-mode

openai/gpt-5.6-terra-pro:batch 1.05M context $1/M input $6/M output

No provider description is available for this model yet.

databricks/databricks-kimi-k3 1M context $3/M input $15/M output

No provider description is available for this model yet.

mistral/ministral-14b-2512 262.144K context $0.2/M input $0.2/M output

No provider description is available for this model yet.

mistral/ministral-14b-latest 262.144K context $0.2/M input $0.2/M output

No provider description is available for this model yet.

mistral/mistral-medium-3 262.144K context $1.5/M input $7.5/M output

No provider description is available for this model yet.

together_ai/zai-org/glm-5.3-flash 1.04858M context $0.15/M input $0.5/M output