1,118 models

No provider description is available for this model yet.

openrouter/minimax/minimax-m2.1 204K context $0.27/M input $1.2/M output

No provider description is available for this model yet.

mistral/ministral-14b-2512 262.144K context $0.2/M input $0.2/M output

No provider description is available for this model yet.

databricks/databricks-kimi-k3 1M context $3/M input $15/M output

Fugu Ultra v2 is the higher-performance model in Sakana AI's Fugu family. Rather than a single monolithic model, Fugu is a learned multi-agent orchestration system: a language model trained to...

sakana/fugu-ultra-v2 1M context $5/M input $30/M output

No provider description is available for this model yet.

openrouter/openai/gpt-5.6-sol 1.05M context $2/M input $10/M output

Fugu Max is the cost-performance model in Sakana AI's Fugu family. Rather than a single monolithic model, Fugu is a learned multi-agent orchestration system: a language model trained to route...

sakana/fugu-max 1M context $2/M input $6/M output

No provider description is available for this model yet.

openrouter/z-ai/glm-4.7-flash 200K context $0.07/M input $0.4/M output

No provider description is available for this model yet.

openrouter/openai/gpt-5.2-pro 272K context $21/M input $168/M output

GPT-5 Pro is OpenAI’s most advanced model, offering major improvements in reasoning, code quality, and user experience. It is optimized for complex tasks that require step-by-step reasoning, instruction following, and...

openai/gpt-5-pro:batch 400K context $7.5/M input $60/M output

Sonnet 4.6 is Anthropic's most capable Sonnet-class model yet, with frontier performance across coding, agents, and professional work. It excels at iterative development, complex codebase navigation, end-to-end project management with...

anthropic/claude-sonnet-4.6 1M context $3/M input $15/M output

Sonnet 5 is Anthropic's most capable Sonnet-class model, with frontier performance across coding, agents, and professional work. It supports adaptive thinking with selectable reasoning effort levels (low, medium, high, max,...

anthropic/claude-sonnet-5:batch 1M context $1/M input $5/M output

Gemini 2.5 Flash-Lite is a lightweight reasoning model in the Gemini 2.5 family, optimized for ultra-low latency and cost efficiency. It offers improved throughput, faster token generation, and better performance...

google/gemini-2.5-flash-lite:batch 1.04858M context $0.05/M input $0.2/M output

Claude Sonnet 4.5 is Anthropic’s most advanced Sonnet model to date, optimized for real-world agents and coding workflows. It delivers state-of-the-art performance on coding benchmarks such as SWE-bench Verified, with...

anthropic/claude-sonnet-4.5:batch 1M context $1.5/M input $7.5/M output

Claude Sonnet 4.5 is Anthropic’s most advanced Sonnet model to date, optimized for real-world agents and coding workflows. It delivers state-of-the-art performance on coding benchmarks such as SWE-bench Verified, with...

anthropic/claude-sonnet-4.5 1M context $3/M input $15/M output

MiniMax-M3 is a multimodal foundation model from MiniMax. It supports text, image, and video inputs with text output, a 1M-token context window, and is suited for long-horizon agentic work, coding,...

minimax/minimax-m3:free 1.04858M context Free input Free output

No provider description is available for this model yet.

openrouter/z-ai/glm-4.7 202.752K context $0.4/M input $1.5/M output

GPT-5.6 Terra is a balanced model in OpenAI's GPT-5.6 series, positioned between the flagship Sol tier and the cost-efficient Luna tier. It is suited for everyday coding, reasoning, and agentic...

openai/gpt-5.6-terra:batch 1.05M context $1/M input $6/M output

For tasks that demand low latency, GPT‑4.1 nano is the fastest and cheapest model in the GPT-4.1 series. It delivers exceptional performance at a small size with its 1 million...

openai/gpt-4.1-nano:batch 1.04758M context $0.05/M input $0.2/M output

Claude Haiku 4.5 is Anthropic’s fastest and most efficient model, delivering near-frontier intelligence at a fraction of the cost and latency of larger Claude models. Matching Claude Sonnet 4’s performance...

anthropic/claude-haiku-4.5:batch 200K context $0.5/M input $2.5/M output

Claude Opus 4.5 is Anthropic’s frontier reasoning model optimized for complex software engineering, agentic workflows, and long-horizon computer use. It offers strong multimodal capabilities, competitive performance across real-world coding and...

anthropic/claude-opus-4.5:batch 200K context $2.5/M input $12.5/M output

No provider description is available for this model yet.

friendliai/zai-org/glm-5.3-flash 1.04858M context $0.15/M input $0.5/M output

No provider description is available for this model yet.

qwen_ai_platform/qwen3.8-max 991.808K context $2/M input $6/M output

GPT-5.1 is the latest frontier-grade model in the GPT-5 series, offering stronger general-purpose reasoning, improved instruction adherence, and a more natural conversational style compared to GPT-5. It uses adaptive reasoning...

openai/gpt-5.1:batch 400K context $0.625/M input $5/M output

GPT-5 Mini is a compact version of GPT-5, designed to handle lighter-weight reasoning tasks. It provides the same instruction-following and safety-tuning benefits as GPT-5, but with reduced latency and cost....

openai/gpt-5-mini:batch 400K context $0.125/M input $1/M output

No provider description is available for this model yet.

qwen_ai_platform/qwen3.7-plus 991.808K context Input not listed Output not listed

GPT-5-Nano is the smallest and fastest variant in the GPT-5 system, optimized for developer tools, rapid interactions, and ultra-low latency environments. While limited in reasoning depth compared to its larger...

openai/gpt-5-nano:batch 400K context $0.025/M input $0.2/M output

GPT-5.4 nano is the most lightweight and cost-efficient variant of the GPT-5.4 family, optimized for speed-critical and high-volume tasks. It supports text and image inputs and is designed for low-latency...

openai/gpt-5.4-nano:batch 400K context $0.1/M input $0.625/M output

No provider description is available for this model yet.

snowflake/openai-gpt-5 300K context $1.25/M input $10/M output

Gemini 2.5 Pro is Google’s state-of-the-art AI model designed for advanced reasoning, coding, mathematics, and scientific tasks. It employs “thinking” capabilities, enabling it to reason through responses with enhanced accuracy...

google/gemini-2.5-pro-preview-05-06 1.04858M context $1.25/M input $10/M output

GPT-4.1 is a flagship large language model optimized for advanced instruction following, real-world software engineering, and long-context reasoning. It supports a 1 million token context window and outperforms GPT-4o and...

openai/gpt-4.1:batch 1.04758M context $1/M input $4/M output

Gemini 2.5 Flash is Google's state-of-the-art workhorse model, specifically designed for advanced reasoning, coding, mathematics, and scientific tasks. It includes built-in "thinking" capabilities, enabling it to provide responses with greater...

google/gemini-2.5-flash:batch 1.04858M context $0.15/M input $1.25/M output

No provider description is available for this model yet.

qwen_ai_platform/qwen3.5-plus 991.808K context Input not listed Output not listed