No provider description is available for this model yet.
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
OpenAI o4-mini-high is the same model as [o4-mini](/openai/o4-mini) with reasoning_effort set to high. OpenAI o4-mini is a compact reasoning model in the o-series, optimized for fast, cost-efficient performance while retaining...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Claude Sonnet 4 significantly enhances the capabilities of its predecessor, Sonnet 3.7, excelling in both coding and reasoning tasks with improved precision and controllability. Achieving state-of-the-art performance on SWE-bench (72.7%),...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| gemini-3-flash-previewvertex_ai-language-models/gemini-3-flash-preview | 1.04858M | $0.5 | $3 | — | |||
| gemini-2.0-flash-001gemini/gemini-2.0-flash-001 | 1.04858M | $0.1 | $0.4 | — | |||
| gemini-3.1-pro-preview-customtoolsvertex_ai-language-models/gemini-3.1-pro-preview-customtools | 1.04858M | $2 | $12 | — | |||
| gemini-3-pro-previewvertex_ai/gemini-3-pro-preview | 1.04858M | $2 | $12 | — | |||
| jamba-1.5-large@001vertex_ai-ai21_models/jamba-1.5-large@001 | 256K | $2 | $8 | — | |||
| jamba-1.5-largevertex_ai-ai21_models/jamba-1.5-large | 256K | $2 | $8 | — | |||
| google/gemini-2.0-flash-001openrouter/google/gemini-2.0-flash-001 | 1.04858M | $0.1 | $0.4 | — | |||
| jamba-1.5vertex_ai-ai21_models/jamba-1.5 | 256K | $0.2 | $0.4 | — | |||
| claude-sonnet-4@20250514vertex_ai-anthropic_models/claude-sonnet-4@20250514 | 1M | $3 | $15 | — | |||
| anthropic/claude-opus-4.7openrouter/anthropic/claude-opus-4.7 | 1M | $5 | $25 | — | |||
| nova-lite-v1amazon_nova/nova-lite-v1 | 300K | $0.06 | $0.24 | — | |||
| claude-sonnet-4vertex_ai-anthropic_models/claude-sonnet-4 | 1M | $3 | $15 | — | |||
| claude-opus-4@20250514vertex_ai-anthropic_models/claude-opus-4@20250514 | 200K | $15 | $75 | — | |||
| anthropic/claude-haiku-4.5openrouter/anthropic/claude-haiku-4.5 | 200K | $1 | $5 | — | |||
| claude-sonnet-4-5@20250929vertex_ai-anthropic_models/claude-sonnet-4-5@20250929 | 200K | $3 | $15 | — | |||
| google/gemini-3.1-flash-litedeepinfra/google/gemini-3.1-flash-lite | 1M | $0.25 | $1.5 | — | |||
| claude-sonnet-4-6vertex_ai-anthropic_models/claude-sonnet-4-6 | 1M | $3 | $15 | — | |||
| qwen/qwen3-coder-plusopenrouter/qwen/qwen3-coder-plus | 997.952K | $1 | $5 | — | |||
| Qwen3-4B-Instruct-2507-GGUFlemonade/qwen3-4b-instruct-2507-gguf | 262.144K | — | — | — | |||
| gemini-3.1-pro-preview-customtoolsgemini/gemini-3.1-pro-preview-customtools | 1.04858M | $2 | $12 | — | |||
| MiniMaxAI/MiniMax-M3deepinfra/minimaxai/minimax-m3 | 524.288K | $0.28 | $1.1 | — | |||
| Qwen/Qwen3.8-2.4T-A95Bdeepinfra/qwen/qwen3.8-2.4t-a95b | 262.144K | $2 | $6 | — | |||
| claude-sonnet-5vertex_ai-anthropic_models/claude-sonnet-5 | 1M | $2 | $10 | — | |||
| google/gemini-3.7-flashdeepinfra/google/gemini-3.7-flash | 1M | $0.75 | $3.75 | — | |||
| stepfun-ai/Step-3.7-Flashdeepinfra/stepfun-ai/step-3.7-flash | 262.144K | $0.2 | $1.15 | — | |||
| Qwen/Qwen3.5-35B-A3Bdeepinfra/qwen/qwen3.5-35b-a3b | 262.144K | $0.14 | $1 | — | |||
| ByteDance/Seed-1.8deepinfra/bytedance/seed-1.8 | 256K | $0.25 | $2 | — | |||
| OpenAI: o4 Mini High (batch)openai/o4-mini-high:batch | 200K | $0.55 | $2.2 | — | |||
| tencent/Hy3deepinfra/tencent/hy3 | 262.144K | $0.14 | $0.58 | — | |||
| ByteDance/Seed-2.0-codedeepinfra/bytedance/seed-2.0-code | 256K | $0.5 | $3 | — | |||
| ByteDance/Seed-2.0-prodeepinfra/bytedance/seed-2.0-pro | 256K | $0.5 | $3 | — | |||
| zai-org/GLM-5deepinfra/zai-org/glm-5 | 202.752K | $0.6 | $2.08 | — | |||
| nvidia/Nemotron-3-Nano-30B-A3Bdeepinfra/nvidia/nemotron-3-nano-30b-a3b | 262.144K | $0.05 | $0.2 | — | |||
| moonshotai/Kimi-K2.7-Codedeepinfra/moonshotai/kimi-k2.7-code | 262.144K | $0.68 | $3.4 | — | |||
| anthropic/claude-sonnet-5deepinfra/anthropic/claude-sonnet-5 | 1M | $2 | $10 | — | |||
| mistral-vibe-cli-fastmistral/mistral-vibe-cli-fast | 262.144K | $0.15 | $0.6 | — | |||
| Qwen/Qwen3.5-397B-A17Bdeepinfra/qwen/qwen3.5-397b-a17b | 262.144K | $0.45 | $3 | — | |||
| mistral-code-agent-latestmistral/mistral-code-agent-latest | 256K | $0.4 | $2 | — | |||
| labs-leanstral-1-5-1mistral/labs-leanstral-1-5-1 | 262.144K | — | — | — | |||
| deepseek-ai/DeepSeek-V4-Flash-0731deepinfra/deepseek-ai/deepseek-v4-flash-0731 | 1.04858M | $0.08 | $0.18 | — | |||
| Qwen/Qwen3.8-Maxdeepinfra/qwen/qwen3.8-max | 256K | $1.65 | $4.951 | — | |||
| anthropic/claude-fable-5deepinfra/anthropic/claude-fable-5 | 1M | $10 | $50 | — | |||
| Anthropic: Claude Sonnet 4anthropic/claude-sonnet-4 | 200K | $3 | $15 | — | |||
| mistral-vibe-cli-with-toolsmistral/mistral-vibe-cli-with-tools | 262.144K | $1.5 | $7.5 | — | |||
| nvidia/NVIDIA-Nemotron-3-Ultra-550B-A55Bdeepinfra/nvidia/nvidia-nemotron-3-ultra-550b-a55b | 262.144K | $0.5 | $2.2 | — | |||
| Qwen/Qwen3.5-122B-A10Bdeepinfra/qwen/qwen3.5-122b-a10b | 262.144K | $0.29 | $2.4 | — | |||
| zai-org/GLM-5.1deepinfra/zai-org/glm-5.1 | 202.752K | $1.05 | $3.5 | — | |||
| accounts/fireworks/models/glm-5p3fireworks_ai/accounts/fireworks/models/glm-5p3 | 1.04858M | $1.4 | $4.4 | — | |||
| deepseek-ai/DeepSeek-V4-Prodeepinfra/deepseek-ai/deepseek-v4-pro | 1.04858M | $1.3 | $2.6 | — | |||
| ByteDance/Seed-2.0-minideepinfra/bytedance/seed-2.0-mini | 256K | $0.1 | $0.4 | — |