No provider description is available for this model yet.

azure/us/gpt-5.4 1.05M context $2.75/M input $16.5/M output

No provider description is available for this model yet.

azure/eu/gpt-5.4 1.05M context $2.75/M input $16.5/M output

Jamba Large 1.7 is the latest model in the Jamba open family, offering improvements in grounding, instruction-following, and overall efficiency. Built on a hybrid SSM-Transformer architecture with a 256K context...

ai21/jamba-large-1.7 256K context $2/M input $8/M output

MiniMax-M2 is a compact, high-efficiency large language model optimized for end-to-end coding and agentic workflows. With 10 billion activated parameters (230 billion total), it delivers near-frontier intelligence across general reasoning,...

minimax/minimax-m2 204.8K context $0.255/M input $1.02/M output

GPT-5.4 Pro is OpenAI's most advanced model, building on GPT-5.4's unified architecture with enhanced reasoning capabilities for complex, high-stakes tasks. It features a 1M+ token context window (922K input, 128K...

openai/gpt-5.4-pro:batch 1.05M context $15/M input $90/M output

No provider description is available for this model yet.

together_ai/zai-org/glm-5.2 1.04858M context $1.4/M input $4.4/M output

Qwen3.8 Max (0803) is the August 3, 2026 checkpoint of Qwen3.8 Max, the flagship model in Alibaba's Qwen3.8 series and the general-availability successor to the Qwen3.8 Max Preview. It is...

qwen/qwen3.8-max 1M context $2/M input $6/M output

No provider description is available for this model yet.

azure/gpt-5.4-2026-03-05 1.05M context $2.5/M input $15/M output

No provider description is available for this model yet.

azure/us/gpt-5.4-2026-03-05 1.05M context $2.75/M input $16.5/M output

No provider description is available for this model yet.

azure/eu/gpt-5.4-2026-03-05 1.05M context $2.75/M input $16.5/M output

No provider description is available for this model yet.

azure/gpt-5.6 1.05M context $5/M input $30/M output

No provider description is available for this model yet.

azure/gpt-5.6-sol 1.05M context $5/M input $30/M output

No provider description is available for this model yet.

azure/gpt-5.6-terra 1.05M context $2.5/M input $15/M output

No provider description is available for this model yet.

azure/gpt-5.6-luna 1.05M context $1/M input $6/M output

No provider description is available for this model yet.

azure/us/gpt-5.6 1.05M context $5.5/M input $33/M output

MiniMax-M3 is a multimodal foundation model from MiniMax. It supports text, image, and video inputs with text output, a 1M-token context window, and is suited for long-horizon agentic work, coding,...

minimax/minimax-m3 524.288K context $0.3/M input $1.2/M output

No provider description is available for this model yet.

azure/us/gpt-5.6-sol 1.05M context $5.5/M input $33/M output

No provider description is available for this model yet.

azure/us/gpt-5.6-terra 1.05M context $2.75/M input $16.5/M output

No provider description is available for this model yet.

azure/us/gpt-5.6-luna 1.05M context $1.1/M input $6.6/M output

No provider description is available for this model yet.

azure/eu/gpt-5.6 1.05M context $5.5/M input $33/M output

No provider description is available for this model yet.

azure/eu/gpt-5.6-sol 1.05M context $5.5/M input $33/M output

No provider description is available for this model yet.

azure/eu/gpt-5.6-terra 1.05M context $2.75/M input $16.5/M output

No provider description is available for this model yet.

azure/eu/gpt-5.6-luna 1.05M context $1.1/M input $6.6/M output

No provider description is available for this model yet.

azure/gpt-5.5 1.05M context $5/M input $30/M output

No provider description is available for this model yet.

azure/us/gpt-5.5 1.05M context $5.5/M input $33/M output

No provider description is available for this model yet.

azure/eu/gpt-5.5 1.05M context $5.5/M input $33/M output

No provider description is available for this model yet.

azure_ai/claude-opus-5 1M context $5/M input $25/M output

Sonnet 4.6 is Anthropic's most capable Sonnet-class model yet, with frontier performance across coding, agents, and professional work. It excels at iterative development, complex codebase navigation, end-to-end project management with...

anthropic/claude-sonnet-4.6:batch 1M context $1.5/M input $7.5/M output

Nex-N2.5 is an agentic model built to turn goals into working, verified outcomes. Its core strength is agentic coding within a visual feedback loop: it can explore codebases, implement multi-file...

nex-agi/nex-n2.5-mini:free 262.144K context Free input Free output

No provider description is available for this model yet.

azure/us/gpt-5.5-2026-04-23 1.05M context $5.5/M input $33/M output

No provider description is available for this model yet.

azure/eu/gpt-5.5-2026-04-23 1.05M context $5.5/M input $33/M output

Mercury 2.5 is the fastest reasoning LLM, and the latest diffusion LLM (dLLM) from Inception. Instead of generating tokens sequentially, Mercury 2.5 produces and refines multiple tokens in parallel, achieving...

inception/mercury-2.5-preview 260K context $0.2/M input $0.75/M output

No provider description is available for this model yet.

azure/gpt-5.4-mini 272K context $0.75/M input $4.5/M output

Gemini 2.5 Pro is Google’s state-of-the-art AI model designed for advanced reasoning, coding, mathematics, and scientific tasks. It employs “thinking” capabilities, enabling it to reason through responses with enhanced accuracy...

google/gemini-2.5-pro:batch 1.04858M context $0.625/M input $5/M output

No provider description is available for this model yet.

azure/gpt-5.2-2025-12-11 272K context $1.75/M input $14/M output

No provider description is available for this model yet.

azure/gpt-5.4-mini-2026-03-17 272K context $0.75/M input $4.5/M output

No provider description is available for this model yet.

azure/gpt-5.4-nano 272K context $0.2/M input $1.25/M output

Qwen3-Max is an updated release built on the Qwen3 series, offering major improvements in reasoning, instruction following, multilingual support, and long-tail knowledge coverage compared to the January 2025 version. It...

qwen/qwen3-max 262.144K context $0.78/M input $3.9/M output

No provider description is available for this model yet.

azure/gpt-5.4-nano-2026-03-17 272K context $0.2/M input $1.25/M output

Grok 4.20 Multi-Agent is a variant of SpaceXAI’s Grok 4.20 designed for collaborative, agent-based workflows. Multiple agents operate in parallel to conduct deep research, coordinate tool use, and synthesize information...

x-ai/grok-4.20-multi-agent 2M context $1.25/M input $2.5/M output

No provider description is available for this model yet.

azure/o1 200K context $15/M input $60/M output

No provider description is available for this model yet.

azure/o1-2024-12-17 200K context $15/M input $60/M output

No provider description is available for this model yet.

azure/gpt-5.2 272K context $1.75/M input $14/M output