No provider description is available for this model yet.
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Ling 3.0 Flash Sante is a health and medicine-focused mixture-of-experts model from InclusionAI, built on Ling 3.0 Flash with 5.1B active parameters out of 124B total. It is designed for...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
GPT-5-Codex is a specialized version of GPT-5 optimized for software engineering and coding workflows. It is designed for both interactive development sessions and long, independent execution of complex engineering tasks....
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
GPT-5.1 Chat (AKA Instant is the fast, lightweight member of the 5.1 family, optimized for low-latency chat while retaining strong general intelligence. It uses adaptive reasoning to selectively “think” on...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
OpenAI o3-mini is a cost-efficient language model optimized for STEM reasoning tasks, particularly excelling in science, mathematics, and coding. This model supports the `reasoning_effort` parameter, which can be set to...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| databricks-gpt-5-3-codexdatabricks/databricks-gpt-5-3-codex | 272K | $1.75 | $14 | — | |||
| databricks-gpt-5-4databricks/databricks-gpt-5-4 | 272K | $2.5 | $15 | — | |||
| databricks-gpt-5-4-minidatabricks/databricks-gpt-5-4-mini | 272K | $0.75 | $4.5 | — | |||
| databricks-gpt-5-4-nanodatabricks/databricks-gpt-5-4-nano | 272K | $0.2 | $1.25 | — | |||
| deepseek/deepseek-v3/communitynovita/deepseek/deepseek-v3/community | 64K | $0.89 | $0.89 | — | |||
| deepseek/deepseek-r1/communitynovita/deepseek/deepseek-r1/community | 64K | $4 | $4 | — | |||
| thudm/glm-4-32b-0414novita/thudm/glm-4-32b-0414 | 32K | $0.55 | $1.66 | — | |||
| meta-llama/llama-3.2-1b-instructnovita/meta-llama/llama-3.2-1b-instruct | 131K | $0.02 | $0.02 | — | |||
| deepseek-ai/DeepSeek-V4-Flashwandb/deepseek-ai/deepseek-v4-flash | 1.04858M | $0.14 | $0.28 | — | |||
| deepseek-ai/DeepSeek-V4-Flash-0731wandb/deepseek-ai/deepseek-v4-flash-0731 | 262.144K | $0.13 | $0.28 | — | |||
| deepseek-ai/DeepSeek-V4-Prowandb/deepseek-ai/deepseek-v4-pro | 1.04858M | $1.15 | $2.55 | — | |||
| google/gemma-4-31B-itwandb/google/gemma-4-31b-it | 262.144K | $0.1 | $0.34 | — | |||
| ibm-granite/granite-4.1-8bwandb/ibm-granite/granite-4.1-8b | 131.072K | $0.05 | $0.1 | — | |||
| JetBrains/Mellum2-12B-A2.5B-Instructwandb/jetbrains/mellum2-12b-a2.5b-instruct | 131.072K | $0.05 | $0.1 | — | |||
| meta-llama/Llama-3.1-70B-Instructwandb/meta-llama/llama-3.1-70b-instruct | 128K | $0.8 | $0.8 | — | |||
| MiniMaxAI/MiniMax-M3wandb/minimaxai/minimax-m3 | 262.144K | $0.23 | $0.96 | — | |||
| moonshotai/Kimi-K2.7-Codewandb/moonshotai/kimi-k2.7-code | 262.144K | $0.71 | $3.5 | — | |||
| moonshotai/Kimi-K2.6wandb/moonshotai/kimi-k2.6 | 262.144K | $0.65 | $3.41 | — | |||
| nvidia/NVIDIA-Nemotron-3.5-Lightning-30B-A3Bwandb/nvidia/nvidia-nemotron-3.5-lightning-30b-a3b | 262.144K | $0.1 | $0.25 | — | |||
| nvidia/NVIDIA-Nemotron-3-Ultra-550B-A55Bwandb/nvidia/nvidia-nemotron-3-ultra-550b-a55b | 262.144K | $0.75 | $2.75 | — | |||
| OpenPipe/Qwen3-14B-Instructwandb/openpipe/qwen3-14b-instruct | 32.768K | $0.05 | $0.22 | — | |||
| Qwen/Qwen3.8-27Bwandb/qwen/qwen3.8-27b | 262.144K | $0.4 | $3 | — | |||
| Qwen/Qwen3.6-35B-A3Bwandb/qwen/qwen3.6-35b-a3b | 262.144K | $0.25 | $1.25 | — | |||
| Qwen/Qwen3.6-27Bwandb/qwen/qwen3.6-27b | 262.144K | $0.6 | $3.6 | — | |||
| Qwen/Qwen3.5-35B-A3Bwandb/qwen/qwen3.5-35b-a3b | 262.144K | $0.25 | $1.25 | — | |||
| inclusionAI: Ling 3.0 Flash Sante (free)inclusionai/ling-3.0-flash-sante:free | 262.144K | Free | Free | — | |||
| Qwen/Qwen3-30B-A3B-Instruct-2507wandb/qwen/qwen3-30b-a3b-instruct-2507 | 262.144K | $0.1 | $0.3 | — | |||
| zai-org/GLM-5.2wandb/zai-org/glm-5.2 | 262.144K | $0.76 | $2.42 | — | |||
| openai/gpt-oss-120b-Turbodeepinfra/openai/gpt-oss-120b-turbo | 131.072K | $0.15 | $0.6 | — | |||
| MiniMaxAI/MiniMax-M2.7deepinfra/minimaxai/minimax-m2.7 | 196.608K | $0.25 | $1 | — | |||
| Qwen/Qwen3.8-27Bdeepinfra/qwen/qwen3.8-27b | 262.144K | $0.4 | $3 | — | |||
| OpenAI: GPT-5 Codex (batch)openai/gpt-5-codex:batch | 400K | $0.625 | $5 | — | |||
| google/gemma-4-31B-it-Ultradeepinfra/google/gemma-4-31b-it-ultra | 131.072K | $0.27 | $0.76 | — | |||
| moonshotai/Kimi-K2.5deepinfra/moonshotai/kimi-k2.5 | 262.144K | $0.45 | $2.25 | — | |||
| zai-org/GLM-4.7-Flashdeepinfra/zai-org/glm-4.7-flash | 202.752K | $0.06 | $0.4 | — | |||
| zai-org/GLM-4.6deepinfra/zai-org/glm-4.6 | 202.752K | $0.5 | $2 | — | |||
| anthropic/claude-opus-4-8deepinfra/anthropic/claude-opus-4-8 | 1M | $5 | $25 | — | |||
| anthropic/claude-sonnet-4-6deepinfra/anthropic/claude-sonnet-4-6 | 1M | $3 | $15 | — | |||
| google/gemini-3.5-flashdeepinfra/google/gemini-3.5-flash | 1M | $1.5 | $9 | — | |||
| OpenAI: GPT-5.1 Chatopenai/gpt-5.1-chat | 128K | $1.25 | $10 | — | |||
| XiaomiMiMo/MiMo-V2.5deepinfra/xiaomimimo/mimo-v2.5 | 262.144K | $0.4 | $2 | — | |||
| Qwen/Qwen3-Maxdeepinfra/qwen/qwen3-max | 256K | $1.2 | $6 | — | |||
| google/gemma-4-31B-it-turbodeepinfra/google/gemma-4-31b-it-turbo | 262.144K | $0.09 | $0.34 | — | |||
| OpenAI: o3 Mini (batch)openai/o3-mini:batch | 200K | $0.55 | $2.2 | — | |||
| databricks-glm-5-3-flashdatabricks/databricks-glm-5-3-flash | 1.04858M | — | — | — | |||
| gemma-4-26b-a4b-itgemini/gemma-4-26b-a4b-it | 262.144K | — | — | — | |||
| gemma-4-31b-itgemini/gemma-4-31b-it | 262.144K | — | — | — | |||
| kimi-k2.7-codemoonshot/kimi-k2.7-code | 262.144K | $0.95 | $4 | — | |||
| thinkingmachines/Inkling-Smalldeepinfra/thinkingmachines/inkling-small | 524.288K | $0.45 | $1.2 | — | |||
| zai-org/GLM-5.3together_ai/zai-org/glm-5.3 | 1.04858M | $1.4 | $4.4 | — |