3,455 models

No provider description is available for this model yet.

snowflake/claude-4-sonnet 200K context $3/M input $15/M output

No provider description is available for this model yet.

bedrock/us-east-1/zai.glm-5 200K context $1/M input $3.2/M output

No provider description is available for this model yet.

bedrock/us-west-2/zai.glm-5 200K context $1/M input $3.2/M output

No provider description is available for this model yet.

snowflake/claude-sonnet-4-5 200K context $3/M input $15/M output

No provider description is available for this model yet.

snowflake/claude-sonnet-4-6 200K context $3/M input $15/M output

No provider description is available for this model yet.

snowflake/claude-4-opus 200K context $5/M input $25/M output

No provider description is available for this model yet.

snowflake/claude-haiku-4-5 200K context $1/M input $5/M output

No provider description is available for this model yet.

snowflake/claude-3-7-sonnet 200K context $3/M input $15/M output

No provider description is available for this model yet.

snowflake/openai-gpt-4.1 300K context $2/M input $8/M output

No provider description is available for this model yet.

snowflake/openai-gpt-5 300K context $1.25/M input $10/M output

No provider description is available for this model yet.

snowflake/openai-gpt-5-mini 1M context $0.3/M input $1.2/M output

No provider description is available for this model yet.

snowflake/openai-gpt-5-nano 5M context $0.15/M input $0.6/M output

No provider description is available for this model yet.

snowflake/llama4-maverick 128K context $0.24/M input $0.97/M output

No provider description is available for this model yet.

tensormesh/qwen/qwen3.6-27b-fp8 262.144K context $0.32/M input $3.2/M output

No provider description is available for this model yet.

tensormesh/moonshotai/kimi-k2.6 32.768K context $0.96/M input $4/M output

No provider description is available for this model yet.

tensormesh/minimaxai/minimax-m2.5 196.608K context $0.3/M input $1.2/M output

No provider description is available for this model yet.

tensormesh/google/gemma-4-31b-it 32.768K context $0.14/M input $0.56/M output

No provider description is available for this model yet.

tensormesh/openai/gpt-oss-120b 131.072K context $0.15/M input $0.6/M output

No provider description is available for this model yet.

tensormesh/openai/gpt-oss-20b 131.072K context $0.07/M input $0.28/M output

No provider description is available for this model yet.

tencent/deepseek-v4-pro 1M context $0.435/M input $0.87/M output

No provider description is available for this model yet.

tencent/deepseek-v4-flash 1M context $0.14/M input $0.28/M output

No provider description is available for this model yet.

pinstripes/ps/glm-4.5-air 128K context $0.125/M input $0.45/M output

No provider description is available for this model yet.

pinstripes/ps/qwen3.6-35b-a3b 131.072K context $0.14/M input $0.45/M output

No provider description is available for this model yet.

pinstripes/ps/qwen3-30b-a3b 131.072K context $0.09/M input $0.2/M output

No provider description is available for this model yet.

pinstripes/ps/qwen3-coder-30b-a3b 131.072K context $0.3/M input $0.6/M output

No provider description is available for this model yet.

pinstripes/ps/deepseek-v4-flash 163.84K context $0.1/M input $0.2/M output

No provider description is available for this model yet.

pinstripes/ps/minimax-m2.7 1.00019M context $0.255/M input $0.55/M output

No provider description is available for this model yet.

darkbloom/gemma-4-26b 131.072K context $0.03/M input $0.165/M output

No provider description is available for this model yet.

darkbloom/gpt-oss-20b 131.072K context $0.015/M input $0.07/M output

GLM-5.3-Flash is a native multimodal model from Z.ai. It is suited for efficient coding and long-horizon agent tasks. Its hybrid sparse and linear attention architecture maintains accurate long-context behavior while...

z-ai/glm-5.3-flash:batch 1.04858M context $0.075/M input $0.25/M output

The latest GPT-4 Turbo model with vision capabilities. Vision requests can now use JSON mode and function calling. Training data: up to December 2023.

openai/gpt-4-turbo:batch 128K context $5/M input $15/M output

LFM2.5-2.6B is a compact reasoning model from Liquid AI. It is suited for agent workflows, data extraction, RAG, and long-context processing. Liquid advises against using it for agentic coding or...

liquid/lfm-2.5-2.6b:free 65.536K context Free input Free output

GPT-5.4 is OpenAI’s latest frontier model, unifying the Codex and GPT lines into a single system. It features a 1M+ token context window (922K input, 128K output) with support for...

openai/gpt-5.4:batch 1.05M context $1.25/M input $7.5/M output

GPT-5.6 Luna Pro is the same underlying model as [GPT-5.6 Luna](https://openrouter.ai/openai/gpt-5.6-luna), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. Learn more in OpenAI's docs: https://developers.openai.com/api/docs/guides/reasoning#reasoning-mode

openai/gpt-5.6-luna-pro:batch 1.05M context $0.1/M input $0.6/M output

GPT-4.1 Mini is a mid-sized model delivering performance competitive with GPT-4o at substantially lower latency and cost. It retains a 1 million token context window and scores 45.1% on hard...

openai/gpt-4.1-mini:batch 1.04758M context $0.2/M input $0.8/M output

GLM 5.2 is a large-scale reasoning model from Z.ai. It supports text input and output with a 1M-token context window, and is suited for long-horizon agent workflows, project-level software engineering,...

z-ai/glm-5.2:batch 1.04858M context $0.7/M input $2.2/M output

GLM-5V-Turbo is Z.ai’s first native multimodal agent foundation model, built for vision-based coding and agent-driven tasks. It natively handles image, video, and text inputs, excels at long-horizon planning, complex coding,...

z-ai/glm-5v-turbo 202.752K context $1.2/M input $4/M output

Qwen-Plus, based on the Qwen2.5 foundation model, is a 131K context model with a balanced performance, speed, and cost combination.

qwen/qwen-plus 1M context $0.26/M input $0.78/M output

Claude Sonnet 4.5 is Anthropic’s most advanced Sonnet model to date, optimized for real-world agents and coding workflows. It delivers state-of-the-art performance on coding benchmarks such as SWE-bench Verified, with...

anthropic/claude-sonnet-4.5 1M context $3/M input $15/M output

Ministral 8B is an 8B parameter model featuring a unique interleaved sliding-window attention pattern for faster, memory-efficient inference. Designed for edge use cases, it supports up to 128k context length...

mistralai/ministral-8b 128K context $0.11/M input $0.11/M output

Fugu Max is the cost-performance model in Sakana AI's Fugu family. Rather than a single monolithic model, Fugu is a learned multi-agent orchestration system: a language model trained to route...

sakana/fugu-max 1M context $2/M input $6/M output