1,808 models

No provider description is available for this model yet.

qwen_ai_platform/qwen3.7-plus 991.808K context Input not listed Output not listed

GPT-5-Nano is the smallest and fastest variant in the GPT-5 system, optimized for developer tools, rapid interactions, and ultra-low latency environments. While limited in reasoning depth compared to its larger...

openai/gpt-5-nano:batch 400K context $0.025/M input $0.2/M output

Fugu Max is the cost-performance model in Sakana AI's Fugu family. Rather than a single monolithic model, Fugu is a learned multi-agent orchestration system: a language model trained to route...

sakana/fugu-max 1M context $2/M input $6/M output

Claude Opus 4.8 is Anthropic's most capable generally available model in the Opus family. It supports text, image, and file inputs with text output, with reasoning support and a 1M-token...

anthropic/claude-opus-4.8:batch 1M context $2.5/M input $12.5/M output

No provider description is available for this model yet.

qwen_ai_platform/qwen3.7-max 991.808K context $2.5/M input $7.5/M output

Qwen3.8 2.4T A95B is an open-weight sparse mixture-of-experts model from Qwen and the open-weight variant of [Qwen3.8 Max](/qwen/qwen3.8-max), with 95 billion active parameters out of 2.4 trillion total. It is...

qwen/qwen3.8-2.4t-a95b 1M context $2/M input $6/M output

No provider description is available for this model yet.

qwen_ai_platform/qwen3.5-plus 991.808K context Input not listed Output not listed

GPT-5.4 nano is the most lightweight and cost-efficient variant of the GPT-5.4 family, optimized for speed-critical and high-volume tasks. It supports text and image inputs and is designed for low-latency...

openai/gpt-5.4-nano:batch 400K context $0.1/M input $0.625/M output

Opus 4.7 is the next generation of Anthropic's Opus family, built for long-running, asynchronous agents. Building on the coding and agentic strengths of Opus 4.6, it delivers stronger performance on...

anthropic/claude-opus-4.7:batch 1M context $2.5/M input $12.5/M output

Gemini 2.5 Flash is Google's state-of-the-art workhorse model, specifically designed for advanced reasoning, coding, mathematics, and scientific tasks. It includes built-in "thinking" capabilities, enabling it to provide responses with greater...

google/gemini-2.5-flash:batch 1.04858M context $0.15/M input $1.25/M output

No provider description is available for this model yet.

qwen_ai_platform/qwen3-vl-plus 260.096K context Input not listed Output not listed

Claude Haiku 4.5 is Anthropic’s fastest and most efficient model, delivering near-frontier intelligence at a fraction of the cost and latency of larger Claude models. Matching Claude Sonnet 4’s performance...

anthropic/claude-haiku-4.5:batch 200K context $0.5/M input $2.5/M output

GLM-5V-Turbo is Z.ai’s first native multimodal agent foundation model, built for vision-based coding and agent-driven tasks. It natively handles image, video, and text inputs, excels at long-horizon planning, complex coding,...

z-ai/glm-5v-turbo 202.752K context $1.2/M input $4/M output

GPT-5.5 is OpenAI’s frontier model designed for complex professional workloads, building on GPT-5.4 with stronger reasoning, higher reliability, and improved token efficiency on hard tasks. It features a 1M+ token...

openai/gpt-5.5:batch 1.05M context $2.5/M input $15/M output

No provider description is available for this model yet.

azure_ai/gpt-chat-latest 272K context $5/M input $30/M output

No provider description is available for this model yet.

azure_ai/model-router 200K context $0.14/M input Output not listed

No provider description is available for this model yet.

vertex_ai/xai/grok-4.3 200K context $1.25/M input $2.5/M output

No provider description is available for this model yet.

azure_ai/grok-4-20-reasoning 262K context $1.25/M input $2.5/M output

No provider description is available for this model yet.

vertex_ai/xai/grok-4.6 524.288K context $2/M input $6/M output

No provider description is available for this model yet.

azure_ai/grok-4-20-non-reasoning 262K context $1.25/M input $2.5/M output

No provider description is available for this model yet.

gemini/lyria-3.5 1.04858M context Input not listed Output not listed

MiniMax-M3 is a multimodal foundation model from MiniMax. It supports text, image, and video inputs with text output, a 1M-token context window, and is suited for long-horizon agentic work, coding,...

minimax/minimax-m3 524.288K context $0.3/M input $1.2/M output

No provider description is available for this model yet.

azure/gpt-6-astra 922K context $10/M input $50/M output

No provider description is available for this model yet.

azure/us/gpt-6-astra 922K context $11/M input $55/M output

No provider description is available for this model yet.

azure_ai/codestral-2501 256K context $0.3/M input $0.9/M output

No provider description is available for this model yet.

azure_ai/mai-thinking-1 256K context $2/M input $8/M output

No provider description is available for this model yet.

azure_ai/grok-4.6 200K context $2/M input $6/M output

No provider description is available for this model yet.

databricks/databricks-glm-5-3 1.04858M context $1.4/M input $4.4/M output

No provider description is available for this model yet.

databricks/databricks-grok-4-6 500K context $2.5/M input $7.5/M output

No provider description is available for this model yet.

databricks/databricks-inkling 1M context $1/M input $4.05/M output

No provider description is available for this model yet.

openrouter/z-ai/glm-5 202.752K context $0.8/M input $2.56/M output