No provider description is available for this model yet.
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Qwen3.8 Max (0803) is the August 3, 2026 checkpoint of Qwen3.8 Max, the flagship model in Alibaba's Qwen3.8 series and the general-availability successor to the Qwen3.8 Max Preview. It is...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
MiniMax-M3 is a multimodal foundation model from MiniMax. It supports text, image, and video inputs with text output, a 1M-token context window, and is suited for long-horizon agentic work, coding,...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Compared with GLM-4.5, this generation brings several key improvements: Longer context window: The context window has been expanded from 128K to 200K tokens, enabling the model to handle more complex...
No provider description is available for this model yet.
No provider description is available for this model yet.
Mercury 2.5 is the fastest reasoning LLM, and the latest diffusion LLM (dLLM) from Inception. Instead of generating tokens sequentially, Mercury 2.5 produces and refines multiple tokens in parallel, achieving...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Hy-MT2-7B is a 7B-parameter translation model from Tencent. It supports 33 language pairs and five Chinese dialect and minority-language pairs, with workflows for structured, delimiter-based, contextual, glossary-based, and style-guided translation.
DeepSeek V4 Pro 0813 is a large-scale mixture-of-experts model from DeepSeek. This is the GA release of DeepSeek V4 Pro.
No provider description is available for this model yet.
Qwen3.8 Max 0902 is an updated snapshot of Qwen3.8 Max from Alibaba's Qwen team. It is a 2.4-trillion-parameter mixture-of-experts model that accepts text, image, and video input and returns text,...
No provider description is available for this model yet.
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| eu/gpt-5.4azure/eu/gpt-5.4 | 1.05M | $2.75 | $16.5 | — | |||
| @cf/meta-llama/llama-2-7b-chat-hf-loracloudflare/@cf/meta-llama/llama-2-7b-chat-hf-lora | 8.192K | — | — | — | |||
| thinkingmachines/Inklingtogether_ai/thinkingmachines/inkling | 524.288K | $1 | $4.05 | — | |||
| thinkingmachines/Inkling-Smalltogether_ai/thinkingmachines/inkling-small | 524.288K | $0.5 | $1.2 | — | |||
| us-west-2/moonshotai.kimi-k2.5bedrock/us-west-2/moonshotai.kimi-k2.5 | 262.144K | $0.6 | $3 | — | |||
| zai-org/GLM-5.2together_ai/zai-org/glm-5.2 | 1.04858M | $1.4 | $4.4 | — | |||
| Qwen: Qwen3.8 Max (0803)qwen/qwen3.8-max | 1M | $2 | $6 | — | |||
| gpt-5.4-2026-03-05azure/gpt-5.4-2026-03-05 | 1.05M | $2.5 | $15 | — | |||
| us/gpt-5.4-2026-03-05azure/us/gpt-5.4-2026-03-05 | 1.05M | $2.75 | $16.5 | — | |||
| eu/gpt-5.4-2026-03-05azure/eu/gpt-5.4-2026-03-05 | 1.05M | $2.75 | $16.5 | — | |||
| gpt-5.6azure/gpt-5.6 | 1.05M | $5 | $30 | — | |||
| gpt-5.6-solazure/gpt-5.6-sol | 1.05M | $5 | $30 | — | |||
| gpt-5.6-terraazure/gpt-5.6-terra | 1.05M | $2.5 | $15 | — | |||
| gpt-5.6-lunaazure/gpt-5.6-luna | 1.05M | $1 | $6 | — | |||
| us/gpt-5.6azure/us/gpt-5.6 | 1.05M | $5.5 | $33 | — | |||
| us-west-2/qwen.qwen3-coder-nextbedrock/us-west-2/qwen.qwen3-coder-next | 262.144K | $0.5 | $1.2 | — | |||
| davinci-002text-completion-openai/davinci-002 | 16.384K | $2 | $2 | — | |||
| mistral-7B-Instruct-v0.1ollama/mistral-7b-instruct-v0.1 | 8.192K | — | — | — | |||
| mistral-7B-Instruct-v0.2ollama/mistral-7b-instruct-v0.2 | 32.768K | — | — | — | |||
| us/gpt-5.6-solazure/us/gpt-5.6-sol | 1.05M | $5.5 | $33 | — | |||
| us/gpt-5.6-terraazure/us/gpt-5.6-terra | 1.05M | $2.75 | $16.5 | — | |||
| mistral-large-instruct-2407ollama/mistral-large-instruct-2407 | 65.536K | — | — | — | |||
| mixtral-8x22B-Instruct-v0.1ollama/mixtral-8x22b-instruct-v0.1 | 65.536K | — | — | — | |||
| MiniMax: MiniMax M3 (batch)minimax/minimax-m3:batch | 524.288K | $0.3 | $1.2 | — | |||
| us/gpt-5.6-lunaazure/us/gpt-5.6-luna | 1.05M | $1.1 | $6.6 | — | |||
| eu/gpt-5.6azure/eu/gpt-5.6 | 1.05M | $5.5 | $33 | — | |||
| eu/gpt-5.6-solazure/eu/gpt-5.6-sol | 1.05M | $5.5 | $33 | — | |||
| pplx-7b-chatperplexity/pplx-7b-chat | 8.192K | $0.07 | $0.28 | — | |||
| eu/gpt-5.6-terraazure/eu/gpt-5.6-terra | 1.05M | $2.75 | $16.5 | — | |||
| eu/gpt-5.6-lunaazure/eu/gpt-5.6-luna | 1.05M | $1.1 | $6.6 | — | |||
| us.anthropic.claude-3-5-haiku-20241022-v1:0bedrock/us.anthropic.claude-3-5-haiku-20241022-v1:0 | 200K | $0.8 | $4 | — | |||
| sonar-medium-chatperplexity/sonar-medium-chat | 16.384K | $0.6 | $1.8 | — | |||
| sonar-medium-onlineperplexity/sonar-medium-online | 12K | — | $1.8 | — | |||
| gpt-5.5azure/gpt-5.5 | 1.05M | $5 | $30 | — | |||
| us/gpt-5.5azure/us/gpt-5.5 | 1.05M | $5.5 | $33 | — | |||
| eu/gpt-5.5azure/eu/gpt-5.5 | 1.05M | $5.5 | $33 | — | |||
| gpt-5.5-2026-04-23azure/gpt-5.5-2026-04-23 | 1.05M | $5 | $30 | — | |||
| llama-3.3-70bcerebras/llama-3.3-70b | 128K | $0.85 | $1.2 | — | |||
| Z.ai: GLM 4.6z-ai/glm-4.6 | 198K | $0.43 | $1.75 | — | |||
| us/gpt-5.5-2026-04-23azure/us/gpt-5.5-2026-04-23 | 1.05M | $5.5 | $33 | — | |||
| eu/gpt-5.5-2026-04-23azure/eu/gpt-5.5-2026-04-23 | 1.05M | $5.5 | $33 | — | |||
| Inception: Mercury 2.5 Previewinception/mercury-2.5-preview | 260K | $0.2 | $0.75 | — | |||
| gpt-5.4-miniazure/gpt-5.4-mini | 272K | $0.75 | $4.5 | — | |||
| llama3.1-70bcerebras/llama3.1-70b | 128K | $0.6 | $0.6 | — | |||
| gpt-5.4-mini-2026-03-17azure/gpt-5.4-mini-2026-03-17 | 272K | $0.75 | $4.5 | — | |||
| Tencent: Hy-MT2-7Btencent/hy-mt2-7b | 8.192K | $0.074 | $0.295 | — | |||
| DeepSeek: DeepSeek V4 Pro 0813 (batch)deepseek/deepseek-v4-pro-0813:batch | 1.04858M | $0.66 | $1.98 | — | |||
| gpt-5.4-nanoazure/gpt-5.4-nano | 272K | $0.2 | $1.25 | — | |||
| Qwen: Qwen3.8 Max (0902)qwen/qwen3.8-max-0902 | 1M | $2 | $6 | — | |||
| gpt-5.4-nano-2026-03-17azure/gpt-5.4-nano-2026-03-17 | 272K | $0.2 | $1.25 | — |