No provider description is available for this model yet.
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Claude Opus 4.1 is an updated version of Anthropic’s flagship model, offering improved performance in coding, reasoning, and agentic tasks. It achieves 74.5% on SWE-bench Verified and shows notable gains...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
OpenAI o4-mini-high is the same model as [o4-mini](/openai/o4-mini) with reasoning_effort set to high. OpenAI o4-mini is a compact reasoning model in the o-series, optimized for fast, cost-efficient performance while retaining...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| XiaomiMiMo/MiMo-V2.5deepinfra/xiaomimimo/mimo-v2.5 | 262.144K | $0.4 | $2 | — | |||
| Qwen/Qwen3-Maxdeepinfra/qwen/qwen3-max | 256K | $1.2 | $6 | — | |||
| google/gemma-4-31B-it-turbodeepinfra/google/gemma-4-31b-it-turbo | 262.144K | $0.09 | $0.34 | — | |||
| databricks-glm-5-3-flashdatabricks/databricks-glm-5-3-flash | 1.04858M | — | — | — | |||
| gemma-4-26b-a4b-itgemini/gemma-4-26b-a4b-it | 262.144K | — | — | — | |||
| gemma-4-31b-itgemini/gemma-4-31b-it | 262.144K | — | — | — | |||
| kimi-k2.7-codemoonshot/kimi-k2.7-code | 262.144K | $0.95 | $4 | — | |||
| thinkingmachines/Inkling-Smalldeepinfra/thinkingmachines/inkling-small | 524.288K | $0.45 | $1.2 | — | |||
| zai-org/GLM-5.3together_ai/zai-org/glm-5.3 | 1.04858M | $1.4 | $4.4 | — | |||
| meta-models/Muse-Glimmer-30Bdeepinfra/meta-models/muse-glimmer-30b | 131.072K | $0.3 | $1.2 | — | |||
| glm-5.3-flashzai/glm-5.3-flash | 1.04858M | $0.15 | $0.5 | — | |||
| gemini-omni-1.1-flashgemini/gemini-omni-1.1-flash | 131.072K | $1.5 | $9 | — | |||
| grok-4.20xai/grok-4.20 | 1M | $1.25 | $2.5 | — | |||
| Qwen/Qwen3-Max-Thinkingdeepinfra/qwen/qwen3-max-thinking | 256K | $1.2 | $6 | — | |||
| grok-4.20-reasoningxai/grok-4.20-reasoning | 1M | $1.25 | $2.5 | — | |||
| Qwen/Qwen3-VL-235B-A22B-Instructdeepinfra/qwen/qwen3-vl-235b-a22b-instruct | 262.144K | $0.2 | $0.88 | — | |||
| Qwen/Qwen3-VL-30B-A3B-Instructdeepinfra/qwen/qwen3-vl-30b-a3b-instruct | 262.144K | $0.15 | $0.6 | — | |||
| grok-4.20-reasoning-latestxai/grok-4.20-reasoning-latest | 1M | $1.25 | $2.5 | — | |||
| grok-4.20-non-reasoningxai/grok-4.20-non-reasoning | 1M | $1.25 | $2.5 | — | |||
| Qwen/Qwen3.5-27Bdeepinfra/qwen/qwen3.5-27b | 262.144K | $0.26 | $2.6 | — | |||
| Anthropic: Claude Opus 4.1anthropic/claude-opus-4.1 | 200K | $15 | $75 | — | |||
| Qwen/Qwen3.6-35B-A3Bdeepinfra/qwen/qwen3.6-35b-a3b | 262.144K | $0.1 | $0.95 | — | |||
| grok-4.20-non-reasoning-latestxai/grok-4.20-non-reasoning-latest | 1M | $1.25 | $2.5 | — | |||
| nvidia/Nemotron-Content-Safety-3.5deepinfra/nvidia/nemotron-content-safety-3.5 | 131.072K | $0.2 | $0.2 | — | |||
| qwen/qwen3.8-27bgroq/qwen/qwen3.8-27b | 131.042K | $0.8 | $4 | — | |||
| mistral-medium-3.5mistral/mistral-medium-3.5 | 262.144K | $1.5 | $7.5 | — | |||
| mistral-vibe-cli-latestmistral/mistral-vibe-cli-latest | 262.144K | $1.5 | $7.5 | — | |||
| anthropic/claude-opus-5deepinfra/anthropic/claude-opus-5 | 1M | $5 | $25 | — | |||
| thinkingmachines/Inklingdeepinfra/thinkingmachines/inkling | 524.288K | $0.95 | $4.05 | — | |||
| moonshotai/Kimi-K2.6deepinfra/moonshotai/kimi-k2.6 | 262.144K | $0.75 | $3.5 | — | |||
| deepseek-ai/DeepSeek-V4-Pro-0813deepinfra/deepseek-ai/deepseek-v4-pro-0813 | 1.04858M | $1.3 | $2.6 | — | |||
| Qwen/Qwen3.7-Maxdeepinfra/qwen/qwen3.7-max | 256K | $2.5 | $7.5 | — | |||
| ByteDance/Seed-2.0-minideepinfra/bytedance/seed-2.0-mini | 256K | $0.1 | $0.4 | — | |||
| mistral-vibe-cli-with-toolsmistral/mistral-vibe-cli-with-tools | 262.144K | $1.5 | $7.5 | — | |||
| Qwen/Qwen3.8-2.4T-A95Bdeepinfra/qwen/qwen3.8-2.4t-a95b | 262.144K | $2 | $6 | — | |||
| MiniMaxAI/MiniMax-M3deepinfra/minimaxai/minimax-m3 | 524.288K | $0.28 | $1.1 | — | |||
| google/gemini-3.1-flash-litedeepinfra/google/gemini-3.1-flash-lite | 1M | $0.25 | $1.5 | — | |||
| google/gemini-3.7-flashdeepinfra/google/gemini-3.7-flash | 1M | $0.75 | $3.75 | — | |||
| inclusionAI/Ling-3.0-flashdeepinfra/inclusionai/ling-3.0-flash | 131.072K | $0.06 | $0.18 | — | |||
| stepfun-ai/Step-3.7-Flashdeepinfra/stepfun-ai/step-3.7-flash | 262.144K | $0.2 | $1.15 | — | |||
| Qwen/Qwen3.5-35B-A3Bdeepinfra/qwen/qwen3.5-35b-a3b | 262.144K | $0.14 | $1 | — | |||
| ByteDance/Seed-1.8deepinfra/bytedance/seed-1.8 | 256K | $0.25 | $2 | — | |||
| OpenAI: o4 Mini High (batch)openai/o4-mini-high:batch | 200K | $0.55 | $2.2 | — | |||
| tencent/Hy3deepinfra/tencent/hy3 | 262.144K | $0.14 | $0.58 | — | |||
| ByteDance/Seed-2.0-codedeepinfra/bytedance/seed-2.0-code | 256K | $0.5 | $3 | — | |||
| ByteDance/Seed-2.0-prodeepinfra/bytedance/seed-2.0-pro | 256K | $0.5 | $3 | — | |||
| zai-org/GLM-5deepinfra/zai-org/glm-5 | 202.752K | $0.6 | $2.08 | — | |||
| nvidia/Nemotron-3-Nano-30B-A3Bdeepinfra/nvidia/nemotron-3-nano-30b-a3b | 262.144K | $0.05 | $0.2 | — | |||
| moonshotai/Kimi-K2.7-Codedeepinfra/moonshotai/kimi-k2.7-code | 262.144K | $0.68 | $3.4 | — | |||
| anthropic/claude-sonnet-5deepinfra/anthropic/claude-sonnet-5 | 1M | $2 | $10 | — |