No provider description is available for this model yet.
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
GPT-5-Codex is a specialized version of GPT-5 optimized for software engineering and coding workflows. It is designed for both interactive development sessions and long, independent execution of complex engineering tasks....
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
GPT-5.1 Chat (AKA Instant is the fast, lightweight member of the 5.1 family, optimized for low-latency chat while retaining strong general intelligence. It uses adaptive reasoning to selectively “think” on...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Inkling is an open-weight multimodal mixture-of-experts model from Thinking Machines Lab, with 41B active parameters out of 975B total. It is designed for general-purpose reasoning, coding, agentic and tool-use systems,...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Kimi K3 is a 2.8T parameter open-weight multimodal reasoning model from Moonshot AI. It is suited for complex coding, knowledge work, and long-horizon agentic workflows, and is particularly strong at...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Mistral Small 4 is the next major release in the Mistral Small family, unifying the capabilities of several flagship Mistral models into a single system. It combines strong reasoning from...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| Qwen/Qwen3-30B-A3B-Instruct-2507wandb/qwen/qwen3-30b-a3b-instruct-2507 | 262.144K | $0.1 | $0.3 | — | |||
| zai-org/GLM-5.2wandb/zai-org/glm-5.2 | 262.144K | $0.76 | $2.42 | — | |||
| openai/gpt-oss-120b-Turbodeepinfra/openai/gpt-oss-120b-turbo | 131.072K | $0.15 | $0.6 | — | |||
| MiniMaxAI/MiniMax-M2.7deepinfra/minimaxai/minimax-m2.7 | 196.608K | $0.25 | $1 | — | |||
| Qwen/Qwen3.8-27Bdeepinfra/qwen/qwen3.8-27b | 262.144K | $0.4 | $3 | — | |||
| OpenAI: GPT-5 Codex (batch)openai/gpt-5-codex:batch | 400K | $0.625 | $5 | — | |||
| google/gemma-4-31B-it-Ultradeepinfra/google/gemma-4-31b-it-ultra | 131.072K | $0.27 | $0.76 | — | |||
| moonshotai/Kimi-K2.5deepinfra/moonshotai/kimi-k2.5 | 262.144K | $0.45 | $2.25 | — | |||
| zai-org/GLM-4.7-Flashdeepinfra/zai-org/glm-4.7-flash | 202.752K | $0.06 | $0.4 | — | |||
| zai-org/GLM-4.6deepinfra/zai-org/glm-4.6 | 202.752K | $0.5 | $2 | — | |||
| anthropic/claude-opus-4-8deepinfra/anthropic/claude-opus-4-8 | 1M | $5 | $25 | — | |||
| anthropic/claude-sonnet-4-6deepinfra/anthropic/claude-sonnet-4-6 | 1M | $3 | $15 | — | |||
| google/gemini-3.5-flashdeepinfra/google/gemini-3.5-flash | 1M | $1.5 | $9 | — | |||
| OpenAI: GPT-5.1 Chatopenai/gpt-5.1-chat | 128K | $1.25 | $10 | — | |||
| XiaomiMiMo/MiMo-V2.5deepinfra/xiaomimimo/mimo-v2.5 | 262.144K | $0.4 | $2 | — | |||
| Qwen/Qwen3-Maxdeepinfra/qwen/qwen3-max | 256K | $1.2 | $6 | — | |||
| google/gemma-4-31B-it-turbodeepinfra/google/gemma-4-31b-it-turbo | 262.144K | $0.09 | $0.34 | — | |||
| databricks-glm-5-3-flashdatabricks/databricks-glm-5-3-flash | 1.04858M | — | — | — | |||
| gemma-4-26b-a4b-itgemini/gemma-4-26b-a4b-it | 262.144K | — | — | — | |||
| gemma-4-31b-itgemini/gemma-4-31b-it | 262.144K | — | — | — | |||
| kimi-k2.7-codemoonshot/kimi-k2.7-code | 262.144K | $0.95 | $4 | — | |||
| thinkingmachines/Inkling-Smalldeepinfra/thinkingmachines/inkling-small | 524.288K | $0.45 | $1.2 | — | |||
| zai-org/GLM-5.3together_ai/zai-org/glm-5.3 | 1.04858M | $1.4 | $4.4 | — | |||
| meta-models/Muse-Glimmer-30Bdeepinfra/meta-models/muse-glimmer-30b | 131.072K | $0.3 | $1.2 | — | |||
| glm-5.3-flashzai/glm-5.3-flash | 1.04858M | $0.15 | $0.5 | — | |||
| gemini-omni-1.1-flashgemini/gemini-omni-1.1-flash | 131.072K | $1.5 | $9 | — | |||
| grok-4.20xai/grok-4.20 | 1M | $1.25 | $2.5 | — | |||
| Qwen/Qwen3-Max-Thinkingdeepinfra/qwen/qwen3-max-thinking | 256K | $1.2 | $6 | — | |||
| grok-4.20-reasoningxai/grok-4.20-reasoning | 1M | $1.25 | $2.5 | — | |||
| Qwen/Qwen3-VL-235B-A22B-Instructdeepinfra/qwen/qwen3-vl-235b-a22b-instruct | 262.144K | $0.2 | $0.88 | — | |||
| Qwen/Qwen3-VL-30B-A3B-Instructdeepinfra/qwen/qwen3-vl-30b-a3b-instruct | 262.144K | $0.15 | $0.6 | — | |||
| grok-4.20-reasoning-latestxai/grok-4.20-reasoning-latest | 1M | $1.25 | $2.5 | — | |||
| grok-4.20-non-reasoningxai/grok-4.20-non-reasoning | 1M | $1.25 | $2.5 | — | |||
| Qwen/Qwen3.5-27Bdeepinfra/qwen/qwen3.5-27b | 262.144K | $0.26 | $2.6 | — | |||
| Thinking Machines: Inkling (free)thinkingmachines/inkling:free | 1.04858M | Free | Free | — | |||
| Qwen/Qwen3.6-35B-A3Bdeepinfra/qwen/qwen3.6-35b-a3b | 262.144K | $0.1 | $0.95 | — | |||
| grok-4.20-non-reasoning-latestxai/grok-4.20-non-reasoning-latest | 1M | $1.25 | $2.5 | — | |||
| nvidia/Nemotron-Content-Safety-3.5deepinfra/nvidia/nemotron-content-safety-3.5 | 131.072K | $0.2 | $0.2 | — | |||
| MoonshotAI: Kimi K3 (batch)moonshotai/kimi-k3:batch | 1.04858M | $3 | $15 | — | |||
| qwen/qwen3.8-27bgroq/qwen/qwen3.8-27b | 131.042K | $0.8 | $4 | — | |||
| mistral-medium-3.5mistral/mistral-medium-3.5 | 262.144K | $1.5 | $7.5 | — | |||
| mistral-vibe-cli-latestmistral/mistral-vibe-cli-latest | 262.144K | $1.5 | $7.5 | — | |||
| anthropic/claude-opus-5deepinfra/anthropic/claude-opus-5 | 1M | $5 | $25 | — | |||
| Mistral: Mistral Small 4 (batch)mistralai/mistral-small-2603:batch | 262.144K | $0.075 | $0.3 | — | |||
| thinkingmachines/Inklingdeepinfra/thinkingmachines/inkling | 524.288K | $0.95 | $4.05 | — | |||
| moonshotai/Kimi-K2.6deepinfra/moonshotai/kimi-k2.6 | 262.144K | $0.75 | $3.5 | — | |||
| deepseek-ai/DeepSeek-V4-Pro-0813deepinfra/deepseek-ai/deepseek-v4-pro-0813 | 1.04858M | $1.3 | $2.6 | — | |||
| Qwen/Qwen3.7-Maxdeepinfra/qwen/qwen3.7-max | 256K | $2.5 | $7.5 | — | |||
| ByteDance/Seed-2.0-minideepinfra/bytedance/seed-2.0-mini | 256K | $0.1 | $0.4 | — | |||
| mistral-vibe-cli-with-toolsmistral/mistral-vibe-cli-with-tools | 262.144K | $1.5 | $7.5 | — |