No provider description is available for this model yet.
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Qwen3.8 27B is an open-weight dense vision-language model from Qwen. It is suited for coding, professional workflows, research, multimodal interaction, and long-running agent tasks, with flexible thinking that can be...
Qwen3 Coder Plus is Alibaba's proprietary version of the Open Source Qwen3 Coder 480B A35B. It is a powerful coding agent model specializing in autonomous programming via tool calling and...
No provider description is available for this model yet.
The preview GPT-4 model with improved instruction following, JSON mode, reproducible outputs, parallel function calling, and more. Training data: up to Dec 2023. **Note:** heavily rate limited by OpenAI while...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
GLM 5.2 is a large-scale reasoning model from Z.ai. It supports text input and output with a 1M-token context window, and is suited for long-horizon agent workflows, project-level software engineering,...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Qwen Plus 0728, based on the Qwen3 foundation model, is a 1 million context hybrid reasoning model with a balanced performance, speed, and cost combination.
The Meta Llama 3.3 multilingual large language model (LLM) is a pretrained and instruction tuned generative model in 70B (text in/text out). The Llama 3.3 instruction tuned text only model...
No provider description is available for this model yet.
No provider description is available for this model yet.
MiniMax-M3 is a multimodal foundation model from MiniMax. It supports text, image, and video inputs with text output, a 1M-token context window, and is suited for long-horizon agentic work, coding,...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Qwen 3.6 Plus builds on a hybrid architecture that combines efficient linear attention with sparse mixture-of-experts routing, enabling strong scalability and high-performance inference. Compared to the 3.5 series, it delivers...
No provider description is available for this model yet.
No provider description is available for this model yet.
Gemini 2.5 Pro is Google’s state-of-the-art AI model designed for advanced reasoning, coding, mathematics, and scientific tasks. It employs “thinking” capabilities, enabling it to reason through responses with enhanced accuracy...
No provider description is available for this model yet.
Qwen3 Coder Flash is Alibaba's fast and cost efficient version of their proprietary Qwen3 Coder Plus. It is a powerful coding agent model specializing in autonomous programming via tool calling...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Qwen3-VL-30B-A3B-Instruct is a multimodal model that unifies strong text generation with visual understanding for images and videos. Its Instruct variant optimizes instruction-following for general multimodal tasks. It excels in perception...
Fast-mode variant of [Opus 4.7](/anthropic/claude-opus-4.7) - identical capabilities with higher output speed at premium 6x pricing. Learn more in Anthropic's docs: https://platform.claude.com/docs/en/build-with-claude/fast-mode
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| o1-2024-12-17azure/o1-2024-12-17 | 200K | $15 | $60 | — | |||
| o1-miniazure/o1-mini | 128K | $1.21 | $4.84 | — | |||
| anthropic/claude-opus-5openrouter/anthropic/claude-opus-5 | 1M | $5 | $25 | — | |||
| o1-previewazure/o1-preview | 128K | $15 | $60 | — | |||
| o1-preview-2024-09-12azure/o1-preview-2024-09-12 | 128K | $15 | $60 | — | |||
| llama3.1-8bcerebras/llama3.1-8b | 128K | $0.1 | $0.1 | — | |||
| Qwen: Qwen3.8 27Bqwen/qwen3.8-27b | 262.144K | $0.214 | $2.55 | — | |||
| Qwen: Qwen3 Coder Plusqwen/qwen3-coder-plus | 1M | $0.65 | $3.25 | — | |||
| o3azure/o3 | 200K | $2 | $8 | — | |||
| OpenAI: GPT-4 Turbo Previewopenai/gpt-4-turbo-preview | 128K | $10 | $30 | — | |||
| o3-2025-04-16azure/o3-2025-04-16 | 200K | $2 | $8 | — | |||
| o3-miniazure/o3-mini | 200K | $1.1 | $4.4 | — | |||
| o3-mini-2025-01-31azure/o3-mini-2025-01-31 | 200K | $1.1 | $4.4 | — | |||
| gpt-oss-120bcerebras/gpt-oss-120b | 131.072K | $0.35 | $0.75 | — | |||
| Z.ai: GLM 5.2 (free)z-ai/glm-5.2:free | 256K | Free | Free | — | |||
| o4-miniazure/o4-mini | 200K | $1.1 | $4.4 | — | |||
| o4-mini-2025-04-16azure/o4-mini-2025-04-16 | 200K | $1.1 | $4.4 | — | |||
| us/gpt-4.1-2025-04-14azure/us/gpt-4.1-2025-04-14 | 1.04758M | $2.2 | $8.8 | — | |||
| us/gpt-4.1-mini-2025-04-14azure/us/gpt-4.1-mini-2025-04-14 | 1.04758M | $0.44 | $1.76 | — | |||
| us/gpt-4.1-nano-2025-04-14azure/us/gpt-4.1-nano-2025-04-14 | 1.04758M | $0.11 | $0.44 | — | |||
| us/gpt-4o-2024-08-06azure/us/gpt-4o-2024-08-06 | 128K | $2.75 | $11 | — | |||
| us/gpt-4o-2024-11-20azure/us/gpt-4o-2024-11-20 | 128K | $2.75 | $11 | — | |||
| glm-5-2mistral/glm-5-2 | 1.04858M | $1.4 | $4.4 | — | |||
| us/gpt-5-2025-08-07azure/us/gpt-5-2025-08-07 | 272K | $1.375 | $11 | — | |||
| @cf/meta/llama-3.3-70b-instruct-fp8-fastcloudflare/@cf/meta/llama-3.3-70b-instruct-fp8-fast | 24K | $0.293 | $2.253 | — | |||
| Qwen: Qwen Plus 0728qwen/qwen-plus-2025-07-28 | 1M | $0.26 | $0.78 | — | |||
| Meta: Llama 3.3 70B Instruct (free)meta-llama/llama-3.3-70b-instruct:free | 65.536K | Free | Free | — | |||
| us/gpt-5-nano-2025-08-07azure/us/gpt-5-nano-2025-08-07 | 272K | $0.055 | $0.44 | — | |||
| global.anthropic.claude-opus-4-8bedrock_converse/global.anthropic.claude-opus-4-8 | 1M | $5 | $25 | — | |||
| MiniMax: MiniMax M3minimax/minimax-m3 | 524.288K | $0.3 | $1.2 | — | |||
| @cf/ibm-granite/granite-4.0-h-microcloudflare/@cf/ibm-granite/granite-4.0-h-micro | 131K | $0.017 | $0.112 | — | |||
| us/gpt-5.1-chatazure/us/gpt-5.1-chat | 128K | $1.38 | $11 | — | |||
| us/o1-2024-12-17azure/us/o1-2024-12-17 | 200K | $16.5 | $66 | — | |||
| us/o1-mini-2024-09-12azure/us/o1-mini-2024-09-12 | 128K | $1.21 | $4.84 | — | |||
| @cf/qwen/qwen2.5-coder-32b-instructcloudflare/@cf/qwen/qwen2.5-coder-32b-instruct | 32.768K | $0.66 | $1 | — | |||
| Qwen: Qwen3.6 Plusqwen/qwen3.6-plus | 1M | $0.325 | $1.95 | — | |||
| us/o1-preview-2024-09-12azure/us/o1-preview-2024-09-12 | 128K | $16.5 | $66 | — | |||
| qwen-3-32bcerebras/qwen-3-32b | 128K | $0.4 | $0.8 | — | |||
| Google: Gemini 2.5 Pro (batch)google/gemini-2.5-pro:batch | 1.04858M | $0.625 | $5 | — | |||
| zai-glm-5-2mistral/zai-glm-5-2 | 1.04858M | $1.4 | $4.4 | — | |||
| Qwen: Qwen3 Coder Flashqwen/qwen3-coder-flash | 1M | $0.195 | $0.975 | — | |||
| us/o3-mini-2025-01-31azure/us/o3-mini-2025-01-31 | 200K | $1.21 | $4.84 | — | |||
| us/o4-mini-2025-04-16azure/us/o4-mini-2025-04-16 | 200K | $1.21 | $4.84 | — | |||
| Llama-3.2-11B-Vision-Instructazure_ai/llama-3.2-11b-vision-instruct | 128K | $0.37 | $0.37 | — | |||
| Llama-3.2-90B-Vision-Instructazure_ai/llama-3.2-90b-vision-instruct | 128K | $2.04 | $2.04 | — | |||
| Llama-3.3-70B-Instructazure_ai/llama-3.3-70b-instruct | 128K | $0.71 | $0.71 | — | |||
| Llama-4-Maverick-17B-128E-Instruct-FP8azure_ai/llama-4-maverick-17b-128e-instruct-fp8 | 1M | $1.41 | $0.35 | — | |||
| Llama-4-Scout-17B-16E-Instructazure_ai/llama-4-scout-17b-16e-instruct | 10M | $0.2 | $0.78 | — | |||
| Qwen: Qwen3 VL 30B A3B Instructqwen/qwen3-vl-30b-a3b-instruct | 262.144K | $0.15 | $0.6 | — | |||
| Anthropic: Claude Opus 4.7 (Fast)anthropic/claude-opus-4.7-fast | 1M | $30 | $150 | — |