No provider description is available for this model yet.
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
No provider description is available for this model yet.
Jamba Large 1.7 is the latest model in the Jamba open family, offering improvements in grounding, instruction-following, and overall efficiency. Built on a hybrid SSM-Transformer architecture with a 256K context...
MiniMax-M2 is a compact, high-efficiency large language model optimized for end-to-end coding and agentic workflows. With 10 billion activated parameters (230 billion total), it delivers near-frontier intelligence across general reasoning,...
No provider description is available for this model yet.
No provider description is available for this model yet.
GPT-5.4 Pro is OpenAI's most advanced model, building on GPT-5.4's unified architecture with enhanced reasoning capabilities for complex, high-stakes tasks. It features a 1M+ token context window (922K input, 128K...
No provider description is available for this model yet.
Qwen3.8 Max (0803) is the August 3, 2026 checkpoint of Qwen3.8 Max, the flagship model in Alibaba's Qwen3.8 series and the general-availability successor to the Qwen3.8 Max Preview. It is...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
MiniMax-M3 is a multimodal foundation model from MiniMax. It supports text, image, and video inputs with text output, a 1M-token context window, and is suited for long-horizon agentic work, coding,...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Sonnet 4.6 is Anthropic's most capable Sonnet-class model yet, with frontier performance across coding, agents, and professional work. It excels at iterative development, complex codebase navigation, end-to-end project management with...
Nex-N2.5 is an agentic model built to turn goals into working, verified outcomes. Its core strength is agentic coding within a visual feedback loop: it can explore codebases, implement multi-file...
No provider description is available for this model yet.
No provider description is available for this model yet.
Mercury 2.5 is the fastest reasoning LLM, and the latest diffusion LLM (dLLM) from Inception. Instead of generating tokens sequentially, Mercury 2.5 produces and refines multiple tokens in parallel, achieving...
No provider description is available for this model yet.
Gemini 2.5 Pro is Google’s state-of-the-art AI model designed for advanced reasoning, coding, mathematics, and scientific tasks. It employs “thinking” capabilities, enabling it to reason through responses with enhanced accuracy...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Qwen3-Max is an updated release built on the Qwen3 series, offering major improvements in reasoning, instruction following, multilingual support, and long-tail knowledge coverage compared to the January 2025 version. It...
No provider description is available for this model yet.
Grok 4.20 Multi-Agent is a variant of SpaceXAI’s Grok 4.20 designed for collaborative, agent-based workflows. Multiple agents operate in parallel to conduct deep research, coordinate tool use, and synthesize information...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| us/gpt-5.4azure/us/gpt-5.4 | 1.05M | $2.75 | $16.5 | — | |||
| eu/gpt-5.4azure/eu/gpt-5.4 | 1.05M | $2.75 | $16.5 | — | |||
| AI21: Jamba Large 1.7ai21/jamba-large-1.7 | 256K | $2 | $8 | — | |||
| MiniMax: MiniMax M2minimax/minimax-m2 | 204.8K | $0.255 | $1.02 | — | |||
| thinkingmachines/Inklingtogether_ai/thinkingmachines/inkling | 524.288K | $1 | $4.05 | — | |||
| thinkingmachines/Inkling-Smalltogether_ai/thinkingmachines/inkling-small | 524.288K | $0.5 | $1.2 | — | |||
| OpenAI: GPT-5.4 Pro (batch)openai/gpt-5.4-pro:batch | 1.05M | $15 | $90 | — | |||
| zai-org/GLM-5.2together_ai/zai-org/glm-5.2 | 1.04858M | $1.4 | $4.4 | — | |||
| Qwen: Qwen3.8 Max (0803)qwen/qwen3.8-max | 1M | $2 | $6 | — | |||
| gpt-5.4-2026-03-05azure/gpt-5.4-2026-03-05 | 1.05M | $2.5 | $15 | — | |||
| us/gpt-5.4-2026-03-05azure/us/gpt-5.4-2026-03-05 | 1.05M | $2.75 | $16.5 | — | |||
| eu/gpt-5.4-2026-03-05azure/eu/gpt-5.4-2026-03-05 | 1.05M | $2.75 | $16.5 | — | |||
| gpt-5.6azure/gpt-5.6 | 1.05M | $5 | $30 | — | |||
| gpt-5.6-solazure/gpt-5.6-sol | 1.05M | $5 | $30 | — | |||
| gpt-5.6-terraazure/gpt-5.6-terra | 1.05M | $2.5 | $15 | — | |||
| gpt-5.6-lunaazure/gpt-5.6-luna | 1.05M | $1 | $6 | — | |||
| us/gpt-5.6azure/us/gpt-5.6 | 1.05M | $5.5 | $33 | — | |||
| MiniMax: MiniMax M3minimax/minimax-m3 | 524.288K | $0.3 | $1.2 | — | |||
| us-west-2/moonshotai.kimi-k2-thinkingbedrock/us-west-2/moonshotai.kimi-k2-thinking | 262.144K | $0.6 | $2.5 | — | |||
| us-west-2/qwen.qwen3-coder-nextbedrock/us-west-2/qwen.qwen3-coder-next | 262.144K | $0.5 | $1.2 | — | |||
| us/gpt-5.6-solazure/us/gpt-5.6-sol | 1.05M | $5.5 | $33 | — | |||
| us/gpt-5.6-terraazure/us/gpt-5.6-terra | 1.05M | $2.75 | $16.5 | — | |||
| us/gpt-5.6-lunaazure/us/gpt-5.6-luna | 1.05M | $1.1 | $6.6 | — | |||
| eu/gpt-5.6azure/eu/gpt-5.6 | 1.05M | $5.5 | $33 | — | |||
| eu/gpt-5.6-solazure/eu/gpt-5.6-sol | 1.05M | $5.5 | $33 | — | |||
| nvidia/nemotron-3-ultra-550b-a55btogether_ai/nvidia/nemotron-3-ultra-550b-a55b | 512.288K | $0.6 | $3.6 | — | |||
| eu/gpt-5.6-terraazure/eu/gpt-5.6-terra | 1.05M | $2.75 | $16.5 | — | |||
| eu/gpt-5.6-lunaazure/eu/gpt-5.6-luna | 1.05M | $1.1 | $6.6 | — | |||
| us.anthropic.claude-3-5-haiku-20241022-v1:0bedrock/us.anthropic.claude-3-5-haiku-20241022-v1:0 | 200K | $0.8 | $4 | — | |||
| gpt-5.5azure/gpt-5.5 | 1.05M | $5 | $30 | — | |||
| us/gpt-5.5azure/us/gpt-5.5 | 1.05M | $5.5 | $33 | — | |||
| eu/gpt-5.5azure/eu/gpt-5.5 | 1.05M | $5.5 | $33 | — | |||
| claude-opus-5azure_ai/claude-opus-5 | 1M | $5 | $25 | — | |||
| Anthropic: Claude Sonnet 4.6 (batch)anthropic/claude-sonnet-4.6:batch | 1M | $1.5 | $7.5 | — | |||
| Nex AGI: Nex-N2.5-Mini (free)nex-agi/nex-n2.5-mini:free | 262.144K | Free | Free | — | |||
| us/gpt-5.5-2026-04-23azure/us/gpt-5.5-2026-04-23 | 1.05M | $5.5 | $33 | — | |||
| eu/gpt-5.5-2026-04-23azure/eu/gpt-5.5-2026-04-23 | 1.05M | $5.5 | $33 | — | |||
| Inception: Mercury 2.5 Previewinception/mercury-2.5-preview | 260K | $0.2 | $0.75 | — | |||
| gpt-5.4-miniazure/gpt-5.4-mini | 272K | $0.75 | $4.5 | — | |||
| Google: Gemini 2.5 Pro (batch)google/gemini-2.5-pro:batch | 1.04858M | $0.625 | $5 | — | |||
| gpt-5.2-2025-12-11azure/gpt-5.2-2025-12-11 | 272K | $1.75 | $14 | — | |||
| gpt-5.4-mini-2026-03-17azure/gpt-5.4-mini-2026-03-17 | 272K | $0.75 | $4.5 | — | |||
| gpt-5.4-nanoazure/gpt-5.4-nano | 272K | $0.2 | $1.25 | — | |||
| Qwen: Qwen3 Maxqwen/qwen3-max | 262.144K | $0.78 | $3.9 | — | |||
| gpt-5.4-nano-2026-03-17azure/gpt-5.4-nano-2026-03-17 | 272K | $0.2 | $1.25 | — | |||
| SpaceXAI: Grok 4.20 Multi-Agentx-ai/grok-4.20-multi-agent | 2M | $1.25 | $2.5 | — | |||
| o1azure/o1 | 200K | $15 | $60 | — | |||
| o1-2024-12-17azure/o1-2024-12-17 | 200K | $15 | $60 | — | |||
| gpt-5.2azure/gpt-5.2 | 272K | $1.75 | $14 | — | |||
| us-west-2/minimax.minimax-m2.5bedrock/us-west-2/minimax.minimax-m2.5 | 1M | $0.3 | $1.2 | — |