No provider description is available for this model yet.
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Jamba Large 1.7 is the latest model in the Jamba open family, offering improvements in grounding, instruction-following, and overall efficiency. Built on a hybrid SSM-Transformer architecture with a 256K context...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Mistral Medium 3.5 is a dense 128B instruction-following model from Mistral AI. It supports text and image inputs with text output, and is designed for agentic workflows, coding, and complex...
MiniMax-M3 is a multimodal foundation model from MiniMax. It supports text, image, and video inputs with text output, a 1M-token context window, and is suited for long-horizon agentic work, coding,...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Qwen3.8 Max (0803) is the August 3, 2026 checkpoint of Qwen3.8 Max, the flagship model in Alibaba's Qwen3.8 series and the general-availability successor to the Qwen3.8 Max Preview. It is...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Mercury 2 is an extremely fast reasoning LLM, and the first reasoning diffusion LLM (dLLM). Instead of generating tokens sequentially, Mercury 2 produces and refines multiple tokens in parallel, achieving...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Qwen-Plus, based on the Qwen2.5 foundation model, is a 131K context model with a balanced performance, speed, and cost combination.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| gpt-5-chatazure/gpt-5-chat | 128K | $1.25 | $10 | — | |||
| gpt-5-chat-latestazure/gpt-5-chat-latest | 128K | $1.25 | $10 | — | |||
| google/gemma-4-31B-ittogether_ai/google/gemma-4-31b-it | 262.144K | $0.39 | $0.97 | — | |||
| meta-llama/Llama-Guard-4-12Btogether_ai/meta-llama/llama-guard-4-12b | 1.04858M | $0.2 | $0.2 | — | |||
| @cf/zai-org/glm-4.7-flashcloudflare/@cf/zai-org/glm-4.7-flash | 131.072K | $0.061 | $0.4 | — | |||
| gpt-5-miniazure/gpt-5-mini | 272K | $0.25 | $2 | — | |||
| gpt-5-mini-2025-08-07azure/gpt-5-mini-2025-08-07 | 272K | $0.25 | $2 | — | |||
| gpt-5-nanoazure/gpt-5-nano | 272K | $0.05 | $0.4 | — | |||
| us-west-2/minimax.minimax-m2.1bedrock/us-west-2/minimax.minimax-m2.1 | 196K | $0.3 | $1.2 | — | |||
| meta-models/Muse-Glimmer-30Btogether_ai/meta-models/muse-glimmer-30b | 131.072K | $0.35 | $1.5 | — | |||
| moonshotai/Kimi-K2.7-Codetogether_ai/moonshotai/kimi-k2.7-code | 262.144K | $0.95 | $4 | — | |||
| gpt-5-nano-2025-08-07azure/gpt-5-nano-2025-08-07 | 272K | $0.05 | $0.4 | — | |||
| gpt-5.1azure/gpt-5.1 | 272K | $1.25 | $10 | — | |||
| moonshotai/Kimi-K3together_ai/moonshotai/kimi-k3 | 1.04858M | $3 | $15 | — | |||
| gpt-5.1-chatazure/gpt-5.1-chat | 128K | $1.25 | $10 | — | |||
| us-west-2/minimax.minimax-m2.5bedrock/us-west-2/minimax.minimax-m2.5 | 1M | $0.3 | $1.2 | — | |||
| gpt-5.2azure/gpt-5.2 | 272K | $1.75 | $14 | — | |||
| gpt-5.2-2025-12-11azure/gpt-5.2-2025-12-11 | 272K | $1.75 | $14 | — | |||
| gpt-5.2-chatazure/gpt-5.2-chat | 128K | $1.75 | $14 | — | |||
| nvidia/nemotron-3-ultra-550b-a55btogether_ai/nvidia/nemotron-3-ultra-550b-a55b | 512.288K | $0.6 | $3.6 | — | |||
| us-west-2/moonshotai.kimi-k2-thinkingbedrock/us-west-2/moonshotai.kimi-k2-thinking | 262.144K | $0.6 | $2.5 | — | |||
| AI21: Jamba Large 1.7ai21/jamba-large-1.7 | 256K | $2 | $8 | — | |||
| gpt-5.2-chat-2025-12-11azure/gpt-5.2-chat-2025-12-11 | 128K | $1.75 | $14 | — | |||
| gpt-5.3-chatazure/gpt-5.3-chat | 128K | $1.75 | $14 | — | |||
| gpt-5.4azure/gpt-5.4 | 1.05M | $2.5 | $15 | — | |||
| pearl-ai/gemma-4-31b-ittogether_ai/pearl-ai/gemma-4-31b-it | 262.144K | $0.28 | $0.86 | — | |||
| us/gpt-5.4azure/us/gpt-5.4 | 1.05M | $2.75 | $16.5 | — | |||
| eu/gpt-5.4azure/eu/gpt-5.4 | 1.05M | $2.75 | $16.5 | — | |||
| Mistral: Mistral Medium 3.5 (batch)mistralai/mistral-medium-3-5:batch | 262.144K | $0.75 | $3.75 | — | |||
| MiniMax: MiniMax M3 (batch)minimax/minimax-m3:batch | 524.288K | $0.3 | $1.2 | — | |||
| thinkingmachines/Inklingtogether_ai/thinkingmachines/inkling | 524.288K | $1 | $4.05 | — | |||
| thinkingmachines/Inkling-Smalltogether_ai/thinkingmachines/inkling-small | 524.288K | $0.5 | $1.2 | — | |||
| us-west-2/moonshotai.kimi-k2.5bedrock/us-west-2/moonshotai.kimi-k2.5 | 262.144K | $0.6 | $3 | — | |||
| zai-org/GLM-5.2together_ai/zai-org/glm-5.2 | 1.04858M | $1.4 | $4.4 | — | |||
| Qwen: Qwen3.8 Max (0803)qwen/qwen3.8-max | 1M | $2 | $6 | — | |||
| gpt-5.4-2026-03-05azure/gpt-5.4-2026-03-05 | 1.05M | $2.5 | $15 | — | |||
| us/gpt-5.4-2026-03-05azure/us/gpt-5.4-2026-03-05 | 1.05M | $2.75 | $16.5 | — | |||
| eu/gpt-5.4-2026-03-05azure/eu/gpt-5.4-2026-03-05 | 1.05M | $2.75 | $16.5 | — | |||
| gpt-5.6azure/gpt-5.6 | 1.05M | $5 | $30 | — | |||
| gpt-5.6-solazure/gpt-5.6-sol | 1.05M | $5 | $30 | — | |||
| Inception: Mercury 2inception/mercury-2 | 128K | $0.25 | $0.75 | — | |||
| gpt-5.6-terraazure/gpt-5.6-terra | 1.05M | $2.5 | $15 | — | |||
| gpt-5.6-lunaazure/gpt-5.6-luna | 1.05M | $1 | $6 | — | |||
| us/gpt-5.6azure/us/gpt-5.6 | 1.05M | $5.5 | $33 | — | |||
| us-west-2/qwen.qwen3-coder-nextbedrock/us-west-2/qwen.qwen3-coder-next | 262.144K | $0.5 | $1.2 | — | |||
| Qwen: Qwen-Plusqwen/qwen-plus | 1M | $0.26 | $0.78 | — | |||
| us/gpt-5.6-solazure/us/gpt-5.6-sol | 1.05M | $5.5 | $33 | — | |||
| us/gpt-5.6-terraazure/us/gpt-5.6-terra | 1.05M | $2.75 | $16.5 | — | |||
| us/gpt-5.6-lunaazure/us/gpt-5.6-luna | 1.05M | $1.1 | $6.6 | — | |||
| eu/gpt-5.6azure/eu/gpt-5.6 | 1.05M | $5.5 | $33 | — |