No provider description is available for this model yet.
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
gpt-oss-20b is an open-weight 21B parameter model released by OpenAI under the Apache 2.0 license. It uses a Mixture-of-Experts (MoE) architecture with 3.6B active parameters per forward pass, optimized for...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Claude Opus 4.1 is an updated version of Anthropic’s flagship model, offering improved performance in coding, reasoning, and agentic tasks. It achieves 74.5% on SWE-bench Verified and shows notable gains...
Inkling Small is an open-weight multimodal mixture-of-experts model from Thinking Machines Lab, with 12B active parameters out of 276B total. It is positioned as the smaller, more efficient member of...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Inkling Small is an open-weight multimodal mixture-of-experts model from Thinking Machines Lab, with 12B active parameters out of 276B total. It is positioned as the smaller, more efficient member of...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
GPT-4.1 is a flagship large language model optimized for advanced instruction following, real-world software engineering, and long-context reasoning. It supports a 1 million token context window and outperforms GPT-4o and...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
For tasks that demand low latency, GPT‑4.1 nano is the fastest and cheapest model in the GPT-4.1 series. It delivers exceptional performance at a small size with its 1 million...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| o4-miniazure/o4-mini | 200K | $1.1 | $4.4 | — | |||
| o4-mini-2025-04-16azure/o4-mini-2025-04-16 | 200K | $1.1 | $4.4 | — | |||
| us/gpt-4.1-2025-04-14azure/us/gpt-4.1-2025-04-14 | 1.04758M | $2.2 | $8.8 | — | |||
| us/gpt-4.1-mini-2025-04-14azure/us/gpt-4.1-mini-2025-04-14 | 1.04758M | $0.44 | $1.76 | — | |||
| us/gpt-4.1-nano-2025-04-14azure/us/gpt-4.1-nano-2025-04-14 | 1.04758M | $0.11 | $0.44 | — | |||
| us/gpt-4o-2024-08-06azure/us/gpt-4o-2024-08-06 | 128K | $2.75 | $11 | — | |||
| us/gpt-4o-2024-11-20azure/us/gpt-4o-2024-11-20 | 128K | $2.75 | $11 | — | |||
| us/gpt-4o-mini-2024-07-18azure/us/gpt-4o-mini-2024-07-18 | 128K | $0.165 | $0.66 | — | |||
| us/gpt-5-2025-08-07azure/us/gpt-5-2025-08-07 | 272K | $1.375 | $11 | — | |||
| @cf/meta/llama-3.3-70b-instruct-fp8-fastcloudflare/@cf/meta/llama-3.3-70b-instruct-fp8-fast | 24K | $0.293 | $2.253 | — | |||
| us/gpt-5-mini-2025-08-07azure/us/gpt-5-mini-2025-08-07 | 272K | $0.275 | $2.2 | — | |||
| upstage/Llama-2-70b-instructupstage/Llama-2-70b-instruct | Not documented | — | — | — | |||
| us/gpt-5-nano-2025-08-07azure/us/gpt-5-nano-2025-08-07 | 272K | $0.055 | $0.44 | — | |||
| us/gpt-5.1azure/us/gpt-5.1 | 272K | $1.38 | $11 | — | |||
| @cf/ibm-granite/granite-4.0-h-microcloudflare/@cf/ibm-granite/granite-4.0-h-micro | 131K | $0.017 | $0.112 | — | |||
| us/gpt-5.1-chatazure/us/gpt-5.1-chat | 128K | $1.38 | $11 | — | |||
| OpenAI: gpt-oss-20b (batch)openai/gpt-oss-20b:batch | 131.072K | $0.05 | $0.2 | — | |||
| us/o1-2024-12-17azure/us/o1-2024-12-17 | 200K | $16.5 | $66 | — | |||
| us/o1-mini-2024-09-12azure/us/o1-mini-2024-09-12 | 128K | $1.21 | $4.84 | — | |||
| @cf/qwen/qwen2.5-coder-32b-instructcloudflare/@cf/qwen/qwen2.5-coder-32b-instruct | 32.768K | $0.66 | $1 | — | |||
| Anthropic: Claude Opus 4.1anthropic/claude-opus-4.1 | 200K | $15 | $75 | — | |||
| Thinking Machines: Inkling Small (free)thinkingmachines/inkling-small:free | 1.04858M | Free | Free | — | |||
| upstage/llama-65b-instructupstage/llama-65b-instruct | Not documented | — | — | — | |||
| us/o1-preview-2024-09-12azure/us/o1-preview-2024-09-12 | 128K | $16.5 | $66 | — | |||
| qwen-3-32bcerebras/qwen-3-32b | 128K | $0.4 | $0.8 | — | |||
| Thinking Machines: Inkling Small (batch)thinkingmachines/inkling-small:batch | 524.288K | $0.5 | $1.2 | — | |||
| us/o3-2025-04-16azure/us/o3-2025-04-16 | 200K | $2.2 | $8.8 | — | |||
| us/o3-mini-2025-01-31azure/us/o3-mini-2025-01-31 | 200K | $1.21 | $4.84 | — | |||
| us/o4-mini-2025-04-16azure/us/o4-mini-2025-04-16 | 200K | $1.21 | $4.84 | — | |||
| Llama-3.2-11B-Vision-Instructazure_ai/llama-3.2-11b-vision-instruct | 128K | $0.37 | $0.37 | — | |||
| OpenAI: GPT-4.1 (batch)openai/gpt-4.1:batch | 1.04758M | $1 | $4 | — | |||
| gpt-5github_copilot/gpt-5 | 128K | — | — | — | |||
| upstage/llama-30b-instruct-2048upstage/llama-30b-instruct-2048 | Not documented | — | — | — | |||
| stabilityai/codellama13b_instruct_260k_synthesisstabilityai/codellama13b_instruct_260k_synthesis | Not documented | — | — | — | |||
| Llama-3.2-90B-Vision-Instructazure_ai/llama-3.2-90b-vision-instruct | 128K | $2.04 | $2.04 | — | |||
| Llama-3.3-70B-Instructazure_ai/llama-3.3-70b-instruct | 128K | $0.71 | $0.71 | — | |||
| Llama-4-Maverick-17B-128E-Instruct-FP8azure_ai/llama-4-maverick-17b-128e-instruct-fp8 | 1M | $1.41 | $0.35 | — | |||
| Llama-4-Scout-17B-16E-Instructazure_ai/llama-4-scout-17b-16e-instruct | 10M | $0.2 | $0.78 | — | |||
| Meta-Llama-3-70B-Instructazure_ai/meta-llama-3-70b-instruct | 8.192K | $1.1 | $0.37 | — | |||
| OpenAI: GPT-4.1 Nano (batch)openai/gpt-4.1-nano:batch | 1.04758M | $0.05 | $0.2 | — | |||
| Meta-Llama-3.1-405B-Instructazure_ai/meta-llama-3.1-405b-instruct | 128K | $5.33 | $16 | — | |||
| Meta-Llama-3.1-70B-Instructazure_ai/meta-llama-3.1-70b-instruct | 128K | $2.68 | $3.54 | — | |||
| Meta-Llama-3.1-8B-Instructazure_ai/meta-llama-3.1-8b-instruct | 128K | $0.3 | $0.61 | — | |||
| Phi-3-medium-128k-instructazure_ai/phi-3-medium-128k-instruct | 128K | $0.17 | $0.68 | — | |||
| Phi-3-medium-4k-instructazure_ai/phi-3-medium-4k-instruct | 4.096K | $0.17 | $0.68 | — | |||
| upstage/llama-30b-instructupstage/llama-30b-instruct | Not documented | — | — | — | |||
| Phi-3-mini-128k-instructazure_ai/phi-3-mini-128k-instruct | 128K | $0.13 | $0.52 | — | |||
| Phi-3-mini-4k-instructazure_ai/phi-3-mini-4k-instruct | 4.096K | $0.13 | $0.52 | — | |||
| Phi-3-small-128k-instructazure_ai/phi-3-small-128k-instruct | 128K | $0.15 | $0.6 | — | |||
| gpt-4o-mini-2024-07-18github_copilot/gpt-4o-mini-2024-07-18 | 64K | — | — | — |