No provider description is available for this model yet.
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
GPT-5.4 mini brings the core capabilities of GPT-5.4 to a faster, more efficient model optimized for high-throughput workloads. It supports text and image inputs with strong performance across reasoning, coding,...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Ling 3.0 Flash Sante is a health and medicine-focused mixture-of-experts model from InclusionAI, built on Ling 3.0 Flash with 5.1B active parameters out of 124B total. It is designed for...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
GPT-5-Codex is a specialized version of GPT-5 optimized for software engineering and coding workflows. It is designed for both interactive development sessions and long, independent execution of complex engineering tasks....
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
GPT-5.1 Chat (AKA Instant is the fast, lightweight member of the 5.1 family, optimized for low-latency chat while retaining strong general intelligence. It uses adaptive reasoning to selectively “think” on...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| MiniMaxAI/MiniMax-M3wandb/minimaxai/minimax-m3 | 262.144K | $0.23 | $0.96 | — | |||
| moonshotai/Kimi-K2.7-Codewandb/moonshotai/kimi-k2.7-code | 262.144K | $0.71 | $3.5 | — | |||
| qwen/qwen3-next-80b-a3b-thinkingnovita/qwen/qwen3-next-80b-a3b-thinking | 131.072K | $0.15 | $1.5 | — | |||
| moonshotai/Kimi-K2.6wandb/moonshotai/kimi-k2.6 | 262.144K | $0.65 | $3.41 | — | |||
| gemini-2.5-pro-preview-ttsvertex_ai-language-models/gemini-2.5-pro-preview-tts | 1.04858M | $1.25 | $10 | — | |||
| nvidia/NVIDIA-Nemotron-3-Ultra-550B-A55Bwandb/nvidia/nvidia-nemotron-3-ultra-550b-a55b | 262.144K | $0.75 | $2.75 | — | |||
| gpt-audio-2025-08-28azure/gpt-audio-2025-08-28 | 128K | $2.5 | $10 | — | |||
| OpenPipe/Qwen3-14B-Instructwandb/openpipe/qwen3-14b-instruct | 32.768K | $0.05 | $0.22 | — | |||
| OpenAI: GPT-5.4 Mini (batch)openai/gpt-5.4-mini:batch | 400K | $0.375 | $2.25 | — | |||
| Qwen/Qwen3.8-27Bwandb/qwen/qwen3.8-27b | 262.144K | $0.4 | $3 | — | |||
| Qwen/Qwen3.6-35B-A3Bwandb/qwen/qwen3.6-35b-a3b | 262.144K | $0.25 | $1.25 | — | |||
| Qwen/Qwen3.6-27Bwandb/qwen/qwen3.6-27b | 262.144K | $0.6 | $3.6 | — | |||
| qwen/qwen3-next-80b-a3b-instructnovita/qwen/qwen3-next-80b-a3b-instruct | 131.072K | $0.15 | $1.5 | — | |||
| kwaipilot/kat-coder-pronovita/kwaipilot/kat-coder-pro | 256K | $0.3 | $1.2 | — | |||
| inclusionAI: Ling 3.0 Flash Sante (free)inclusionai/ling-3.0-flash-sante:free | 262.144K | Free | Free | — | |||
| Qwen/Qwen3.5-35B-A3Bwandb/qwen/qwen3.5-35b-a3b | 262.144K | $0.25 | $1.25 | — | |||
| Qwen/Qwen3-30B-A3B-Instruct-2507wandb/qwen/qwen3-30b-a3b-instruct-2507 | 262.144K | $0.1 | $0.3 | — | |||
| zai-org/GLM-5.2wandb/zai-org/glm-5.2 | 262.144K | $0.76 | $2.42 | — | |||
| openai/gpt-oss-120b-Turbodeepinfra/openai/gpt-oss-120b-turbo | 131.072K | $0.15 | $0.6 | — | |||
| MiniMaxAI/MiniMax-M2.7deepinfra/minimaxai/minimax-m2.7 | 196.608K | $0.25 | $1 | — | |||
| gpt-4github_copilot/gpt-4 | 32.768K | — | — | — | |||
| OpenAI: GPT-5 Codex (batch)openai/gpt-5-codex:batch | 400K | $0.625 | $5 | — | |||
| google/gemma-4-31B-it-Ultradeepinfra/google/gemma-4-31b-it-ultra | 131.072K | $0.27 | $0.76 | — | |||
| moonshotai/Kimi-K2.5deepinfra/moonshotai/kimi-k2.5 | 262.144K | $0.45 | $2.25 | — | |||
| zai-org/GLM-4.7-Flashdeepinfra/zai-org/glm-4.7-flash | 202.752K | $0.06 | $0.4 | — | |||
| zai-org/GLM-4.6deepinfra/zai-org/glm-4.6 | 202.752K | $0.5 | $2 | — | |||
| anthropic/claude-opus-4-8deepinfra/anthropic/claude-opus-4-8 | 1M | $5 | $25 | — | |||
| anthropic/claude-sonnet-4-6deepinfra/anthropic/claude-sonnet-4-6 | 1M | $3 | $15 | — | |||
| google/gemini-3.5-flashdeepinfra/google/gemini-3.5-flash | 1M | $1.5 | $9 | — | |||
| OpenAI: GPT-5.1 Chatopenai/gpt-5.1-chat | 128K | $1.25 | $10 | — | |||
| zai-org/glm-4.6novita/zai-org/glm-4.6 | 204.8K | $0.55 | $2.2 | — | |||
| XiaomiMiMo/MiMo-V2.5deepinfra/xiaomimimo/mimo-v2.5 | 262.144K | $0.4 | $2 | — | |||
| Qwen/Qwen3-Maxdeepinfra/qwen/qwen3-max | 256K | $1.2 | $6 | — | |||
| google/gemma-4-31B-it-turbodeepinfra/google/gemma-4-31b-it-turbo | 262.144K | $0.09 | $0.34 | — | |||
| zai-org/glm-4.6vnovita/zai-org/glm-4.6v | 131.072K | $0.3 | $0.9 | — | |||
| databricks-glm-5-3-flashdatabricks/databricks-glm-5-3-flash | 1.04858M | — | — | — | |||
| gemma-4-26b-a4b-itgemini/gemma-4-26b-a4b-it | 262.144K | — | — | — | |||
| gemma-4-31b-itgemini/gemma-4-31b-it | 262.144K | — | — | — | |||
| kimi-k2.7-codemoonshot/kimi-k2.7-code | 262.144K | $0.95 | $4 | — | |||
| thinkingmachines/Inkling-Smalldeepinfra/thinkingmachines/inkling-small | 524.288K | $0.45 | $1.2 | — | |||
| gemini-3.1-pro-preview-customtoolsvertex_ai/gemini-3.1-pro-preview-customtools | 1.04858M | $2 | $12 | — | |||
| meta-models/Muse-Glimmer-30Bdeepinfra/meta-models/muse-glimmer-30b | 131.072K | $0.3 | $1.2 | — | |||
| glm-5.3-flashzai/glm-5.3-flash | 1.04858M | $0.15 | $0.5 | — | |||
| gemini-omni-1.1-flashgemini/gemini-omni-1.1-flash | 131.072K | $1.5 | $9 | — | |||
| Prism-ML/Ternary-Bonsai-27Btogether_ai/prism-ml/ternary-bonsai-27b | 262.144K | — | — | — | |||
| Qwen/Qwen3-Max-Thinkingdeepinfra/qwen/qwen3-max-thinking | 256K | $1.2 | $6 | — | |||
| grok-4.20-reasoningxai/grok-4.20-reasoning | 1M | $1.25 | $2.5 | — | |||
| Qwen/Qwen3-VL-235B-A22B-Instructdeepinfra/qwen/qwen3-vl-235b-a22b-instruct | 262.144K | $0.2 | $0.88 | — | |||
| Qwen/Qwen3-VL-30B-A3B-Instructdeepinfra/qwen/qwen3-vl-30b-a3b-instruct | 262.144K | $0.15 | $0.6 | — | |||
| qwen/qwen3-vl-235b-a22b-thinkingnovita/qwen/qwen3-vl-235b-a22b-thinking | 131.072K | $0.98 | $3.95 | — |