No provider description is available for this model yet.
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Sonnet 4.6 is Anthropic's most capable Sonnet-class model yet, with frontier performance across coding, agents, and professional work. It excels at iterative development, complex codebase navigation, end-to-end project management with...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
MiniMax-M3 is a multimodal foundation model from MiniMax. It supports text, image, and video inputs with text output, a 1M-token context window, and is suited for long-horizon agentic work, coding,...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
OpenAI o4-mini-high is the same model as [o4-mini](/openai/o4-mini) with reasoning_effort set to high. OpenAI o4-mini is a compact reasoning model in the o-series, optimized for fast, cost-efficient performance while retaining...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| gemma-4-31b-itgemini/gemma-4-31b-it | 262.144K | — | — | — | |||
| kimi-k2.7-codemoonshot/kimi-k2.7-code | 262.144K | $0.95 | $4 | — | |||
| thinkingmachines/Inkling-Smalldeepinfra/thinkingmachines/inkling-small | 524.288K | $0.45 | $1.2 | — | |||
| zai-org/GLM-5.3together_ai/zai-org/glm-5.3 | 1.04858M | $1.4 | $4.4 | — | |||
| meta-models/Muse-Glimmer-30Bdeepinfra/meta-models/muse-glimmer-30b | 131.072K | $0.3 | $1.2 | — | |||
| glm-5.3-flashzai/glm-5.3-flash | 1.04858M | $0.15 | $0.5 | — | |||
| gemini-omni-1.1-flashgemini/gemini-omni-1.1-flash | 131.072K | $1.5 | $9 | — | |||
| grok-4.20xai/grok-4.20 | 1M | $1.25 | $2.5 | — | |||
| Qwen/Qwen3-Max-Thinkingdeepinfra/qwen/qwen3-max-thinking | 256K | $1.2 | $6 | — | |||
| grok-4.20-reasoningxai/grok-4.20-reasoning | 1M | $1.25 | $2.5 | — | |||
| Qwen/Qwen3-VL-235B-A22B-Instructdeepinfra/qwen/qwen3-vl-235b-a22b-instruct | 262.144K | $0.2 | $0.88 | — | |||
| Qwen/Qwen3-VL-30B-A3B-Instructdeepinfra/qwen/qwen3-vl-30b-a3b-instruct | 262.144K | $0.15 | $0.6 | — | |||
| grok-4.20-reasoning-latestxai/grok-4.20-reasoning-latest | 1M | $1.25 | $2.5 | — | |||
| grok-4.20-non-reasoningxai/grok-4.20-non-reasoning | 1M | $1.25 | $2.5 | — | |||
| Qwen/Qwen3.5-27Bdeepinfra/qwen/qwen3.5-27b | 262.144K | $0.26 | $2.6 | — | |||
| Qwen/Qwen3.6-35B-A3Bdeepinfra/qwen/qwen3.6-35b-a3b | 262.144K | $0.1 | $0.95 | — | |||
| Anthropic: Claude Sonnet 4.6anthropic/claude-sonnet-4.6 | 1M | $3 | $15 | — | |||
| grok-4.20-non-reasoning-latestxai/grok-4.20-non-reasoning-latest | 1M | $1.25 | $2.5 | — | |||
| nvidia/Nemotron-Content-Safety-3.5deepinfra/nvidia/nemotron-content-safety-3.5 | 131.072K | $0.2 | $0.2 | — | |||
| qwen/qwen3.8-27bgroq/qwen/qwen3.8-27b | 131.042K | $0.8 | $4 | — | |||
| mistral-medium-3.5mistral/mistral-medium-3.5 | 262.144K | $1.5 | $7.5 | — | |||
| mistral-vibe-cli-latestmistral/mistral-vibe-cli-latest | 262.144K | $1.5 | $7.5 | — | |||
| anthropic/claude-opus-5deepinfra/anthropic/claude-opus-5 | 1M | $5 | $25 | — | |||
| MiniMax: MiniMax M3 (batch)minimax/minimax-m3:batch | 524.288K | $0.3 | $1.2 | — | |||
| thinkingmachines/Inklingdeepinfra/thinkingmachines/inkling | 524.288K | $0.95 | $4.05 | — | |||
| moonshotai/Kimi-K2.6deepinfra/moonshotai/kimi-k2.6 | 262.144K | $0.75 | $3.5 | — | |||
| deepseek-ai/DeepSeek-V4-Pro-0813deepinfra/deepseek-ai/deepseek-v4-pro-0813 | 1.04858M | $1.3 | $2.6 | — | |||
| Qwen/Qwen3.7-Maxdeepinfra/qwen/qwen3.7-max | 256K | $2.5 | $7.5 | — | |||
| ByteDance/Seed-2.0-minideepinfra/bytedance/seed-2.0-mini | 256K | $0.1 | $0.4 | — | |||
| mistral-vibe-cli-with-toolsmistral/mistral-vibe-cli-with-tools | 262.144K | $1.5 | $7.5 | — | |||
| Qwen/Qwen3.8-2.4T-A95Bdeepinfra/qwen/qwen3.8-2.4t-a95b | 262.144K | $2 | $6 | — | |||
| MiniMaxAI/MiniMax-M3deepinfra/minimaxai/minimax-m3 | 524.288K | $0.28 | $1.1 | — | |||
| google/gemini-3.1-flash-litedeepinfra/google/gemini-3.1-flash-lite | 1M | $0.25 | $1.5 | — | |||
| google/gemini-3.7-flashdeepinfra/google/gemini-3.7-flash | 1M | $0.75 | $3.75 | — | |||
| inclusionAI/Ling-3.0-flashdeepinfra/inclusionai/ling-3.0-flash | 131.072K | $0.06 | $0.18 | — | |||
| stepfun-ai/Step-3.7-Flashdeepinfra/stepfun-ai/step-3.7-flash | 262.144K | $0.2 | $1.15 | — | |||
| Qwen/Qwen3.5-35B-A3Bdeepinfra/qwen/qwen3.5-35b-a3b | 262.144K | $0.14 | $1 | — | |||
| ByteDance/Seed-1.8deepinfra/bytedance/seed-1.8 | 256K | $0.25 | $2 | — | |||
| OpenAI: o4 Mini High (batch)openai/o4-mini-high:batch | 200K | $0.55 | $2.2 | — | |||
| tencent/Hy3deepinfra/tencent/hy3 | 262.144K | $0.14 | $0.58 | — | |||
| ByteDance/Seed-2.0-codedeepinfra/bytedance/seed-2.0-code | 256K | $0.5 | $3 | — | |||
| ByteDance/Seed-2.0-prodeepinfra/bytedance/seed-2.0-pro | 256K | $0.5 | $3 | — | |||
| zai-org/GLM-5deepinfra/zai-org/glm-5 | 202.752K | $0.6 | $2.08 | — | |||
| nvidia/Nemotron-3-Nano-30B-A3Bdeepinfra/nvidia/nemotron-3-nano-30b-a3b | 262.144K | $0.05 | $0.2 | — | |||
| moonshotai/Kimi-K2.7-Codedeepinfra/moonshotai/kimi-k2.7-code | 262.144K | $0.68 | $3.4 | — | |||
| anthropic/claude-sonnet-5deepinfra/anthropic/claude-sonnet-5 | 1M | $2 | $10 | — | |||
| mistral-vibe-cli-fastmistral/mistral-vibe-cli-fast | 262.144K | $0.15 | $0.6 | — | |||
| Qwen/Qwen3.5-397B-A17Bdeepinfra/qwen/qwen3.5-397b-a17b | 262.144K | $0.45 | $3 | — | |||
| mistral-code-latestmistral/mistral-code-latest | 128K | $0.3 | $0.9 | — | |||
| mistral-code-fim-latestmistral/mistral-code-fim-latest | 128K | $0.3 | $0.9 | — |