No provider description is available for this model yet.
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Mistral Medium 3.5 is a dense 128B instruction-following model from Mistral AI. It supports text and image inputs with text output, and is designed for agentic workflows, coding, and complex...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Qwen3.8 Max (0803) is the August 3, 2026 checkpoint of Qwen3.8 Max, the flagship model in Alibaba's Qwen3.8 series and the general-availability successor to the Qwen3.8 Max Preview. It is...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Mercury 2 is an extremely fast reasoning LLM, and the first reasoning diffusion LLM (dLLM). Instead of generating tokens sequentially, Mercury 2 produces and refines multiple tokens in parallel, achieving...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Qwen-Plus, based on the Qwen2.5 foundation model, is a 131K context model with a balanced performance, speed, and cost combination.
GPT Chat Latest points to OpenAI's stable API alias `chat-latest` that always resolves to the latest Instant chat model used in ChatGPT. As OpenAI rolls out new Instant model updates...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Claude Opus 4 is benchmarked as the world’s best coding model, at time of release, bringing sustained performance on complex, long-running tasks and agent workflows. It sets new benchmarks in...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| eu/gpt-5.4azure/eu/gpt-5.4 | 1.05M | $2.75 | $16.5 | — | |||
| microsoft/Dayhoff-3b-UR90-20350microsoft/Dayhoff-3b-UR90-20350 | Not documented | — | — | — | |||
| microsoft/Dayhoff-3b-UR90-10000microsoft/Dayhoff-3b-UR90-10000 | Not documented | — | — | — | |||
| microsoft/Dayhoff-3b-UR90-10microsoft/Dayhoff-3b-UR90-10 | Not documented | — | — | — | |||
| microsoft/Dayhoff-170M-GRS-SS-86000microsoft/Dayhoff-170M-GRS-SS-86000 | Not documented | — | — | — | |||
| microsoft/Dayhoff-170M-GRS-SS-38000microsoft/Dayhoff-170M-GRS-SS-38000 | Not documented | — | — | — | |||
| @cf/meta-llama/llama-2-7b-chat-hf-loracloudflare/@cf/meta-llama/llama-2-7b-chat-hf-lora | 8.192K | — | — | — | |||
| Mistral: Mistral Medium 3.5 (batch)mistralai/mistral-medium-3-5:batch | 262.144K | $0.75 | $3.75 | — | |||
| microsoft/Mage-Flowmicrosoft/Mage-Flow | Not documented | — | — | — | |||
| thinkingmachines/Inklingtogether_ai/thinkingmachines/inkling | 524.288K | $1 | $4.05 | — | |||
| meta-llama/Llama-2-13b-chat-hfmeta-llama/Llama-2-13b-chat-hf | Not documented | — | — | — | |||
| us-west-2/moonshotai.kimi-k2.5bedrock/us-west-2/moonshotai.kimi-k2.5 | 262.144K | $0.6 | $3 | — | |||
| zai-org/GLM-5.2together_ai/zai-org/glm-5.2 | 1.04858M | $1.4 | $4.4 | — | |||
| Qwen: Qwen3.8 Max (0803)qwen/qwen3.8-max | 1M | $2 | $6 | — | |||
| gpt-5.4-2026-03-05azure/gpt-5.4-2026-03-05 | 1.05M | $2.5 | $15 | — | |||
| us/gpt-5.4-2026-03-05azure/us/gpt-5.4-2026-03-05 | 1.05M | $2.75 | $16.5 | — | |||
| eu/gpt-5.4-2026-03-05azure/eu/gpt-5.4-2026-03-05 | 1.05M | $2.75 | $16.5 | — | |||
| gpt-5.6azure/gpt-5.6 | 1.05M | $5 | $30 | — | |||
| gpt-5.6-solazure/gpt-5.6-sol | 1.05M | $5 | $30 | — | |||
| Inception: Mercury 2inception/mercury-2 | 128K | $0.25 | $0.75 | — | |||
| eu-west-3/mistral.mistral-large-2402-v1:0bedrock/eu-west-3/mistral.mistral-large-2402-v1:0 | 32K | $10.4 | $31.2 | — | |||
| gpt-5.6-terraazure/gpt-5.6-terra | 1.05M | $2.5 | $15 | — | |||
| gpt-5.6-lunaazure/gpt-5.6-luna | 1.05M | $1 | $6 | — | |||
| Qwen: Qwen-Plusqwen/qwen-plus | 1M | $0.26 | $0.78 | — | |||
| OpenAI: GPT Chat Latestopenai/gpt-chat-latest | 400K | $5 | $30 | — | |||
| us-west-2/qwen.qwen3-coder-nextbedrock/us-west-2/qwen.qwen3-coder-next | 262.144K | $0.5 | $1.2 | — | |||
| davinci-002text-completion-openai/davinci-002 | 16.384K | $2 | $2 | — | |||
| mistral-7B-Instruct-v0.1ollama/mistral-7b-instruct-v0.1 | 8.192K | — | — | — | |||
| mistral-7B-Instruct-v0.2ollama/mistral-7b-instruct-v0.2 | 32.768K | — | — | — | |||
| global/gpt-4o-2024-08-06azure/global/gpt-4o-2024-08-06 | 128K | $2.5 | $10 | — | |||
| us/gpt-5.6-solazure/us/gpt-5.6-sol | 1.05M | $5.5 | $33 | — | |||
| us/gpt-5.6-terraazure/us/gpt-5.6-terra | 1.05M | $2.75 | $16.5 | — | |||
| mistral-large-instruct-2407ollama/mistral-large-instruct-2407 | 65.536K | — | — | — | |||
| mixtral-8x22B-Instruct-v0.1ollama/mixtral-8x22b-instruct-v0.1 | 65.536K | — | — | — | |||
| nvidia/MiniMax-M2.7-DFlashnvidia/MiniMax-M2.7-DFlash | Not documented | — | — | — | |||
| us/gpt-5.6-lunaazure/us/gpt-5.6-luna | 1.05M | $1.1 | $6.6 | — | |||
| eu/gpt-5.6azure/eu/gpt-5.6 | 1.05M | $5.5 | $33 | — | |||
| Anthropic: Claude Opus 4anthropic/claude-opus-4 | 200K | $15 | $75 | — | |||
| microsoft/Dayhoff-170M-UR90-HL-28000microsoft/Dayhoff-170M-UR90-HL-28000 | Not documented | — | — | — | |||
| microsoft/Dayhoff-170M-UR90-HL-26000microsoft/Dayhoff-170M-UR90-HL-26000 | Not documented | — | — | — | |||
| mixtral-8x7b-instructperplexity/mixtral-8x7b-instruct | 4.096K | $0.07 | $0.28 | — | |||
| microsoft/Dayhoff-170M-UR90-HL-24000microsoft/Dayhoff-170M-UR90-HL-24000 | Not documented | — | — | — | |||
| microsoft/Dayhoff-170M-UR90-HL-22000microsoft/Dayhoff-170M-UR90-HL-22000 | Not documented | — | — | — | |||
| microsoft/Dayhoff-170M-UR90-HL-20000microsoft/Dayhoff-170M-UR90-HL-20000 | Not documented | — | — | — | |||
| eu/gpt-5.6-solazure/eu/gpt-5.6-sol | 1.05M | $5.5 | $33 | — | |||
| pplx-70b-onlineperplexity/pplx-70b-online | 4.096K | — | $2.8 | — | |||
| pplx-7b-chatperplexity/pplx-7b-chat | 8.192K | $0.07 | $0.28 | — | |||
| microsoft/Dayhoff-170M-UR90-HL-18000microsoft/Dayhoff-170M-UR90-HL-18000 | Not documented | — | — | — | |||
| eu/gpt-5.6-terraazure/eu/gpt-5.6-terra | 1.05M | $2.75 | $16.5 | — | |||
| nvidia/Nemotron-Labs-Audex-30B-A3Bnvidia/Nemotron-Labs-Audex-30B-A3B | Not documented | — | — | — |