No provider description is available for this model yet.
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Qwen3-235B-A22B is a 235B parameter mixture-of-experts (MoE) model developed by Qwen, activating 22B parameters per forward pass. It supports seamless switching between a "thinking" mode for complex reasoning, math, and...
GPT-5.1 Chat (AKA Instant is the fast, lightweight member of the 5.1 family, optimized for low-latency chat while retaining strong general intelligence. It uses adaptive reasoning to selectively “think” on...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
GPT-5-Codex is a specialized version of GPT-5 optimized for software engineering and coding workflows. It is designed for both interactive development sessions and long, independent execution of complex engineering tasks....
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| databricks-grok-4-6databricks/databricks-grok-4-6 | 500K | $2.5 | $7.5 | — | |||
| databricks-inklingdatabricks/databricks-inkling | 1M | $1 | $4.05 | — | |||
| databricks-qwen35-122b-a10bdatabricks/databricks-qwen35-122b-a10b | 262.144K | $0.22 | $2.2 | — | |||
| google/gemma-4-31B-it-turbodeepinfra/google/gemma-4-31b-it-turbo | 262.144K | $0.09 | $0.34 | — | |||
| databricks-qwen3-next-80b-a3b-instructdatabricks/databricks-qwen3-next-80b-a3b-instruct | Not documented | $0.15 | $1.2 | — | |||
| deepseek-ai/DeepSeek-V4-Flashnebius/deepseek-ai/deepseek-v4-flash | 1.04858M | $0.14 | $0.28 | — | |||
| deepseek-ai/DeepSeek-V4-Flash-0731nebius/deepseek-ai/deepseek-v4-flash-0731 | 1.024M | $0.14 | $0.28 | — | |||
| deepseek-ai/DeepSeek-V4-Pronebius/deepseek-ai/deepseek-v4-pro | 1.04858M | $1.75 | $3.5 | — | |||
| MiniMaxAI/MiniMax-M2.5nebius/minimaxai/minimax-m2.5 | 196.608K | $0.3 | $1.2 | — | |||
| MiniMaxAI/MiniMax-M3nebius/minimaxai/minimax-m3 | 1.04858M | $0.3 | $1.2 | — | |||
| moonshotai/Kimi-K2.6nebius/moonshotai/kimi-k2.6 | 262.144K | $0.95 | $4 | — | |||
| moonshotai/Kimi-K2.7-Codenebius/moonshotai/kimi-k2.7-code | 262.144K | $0.95 | $4 | — | |||
| moonshotai/Kimi-K3nebius/moonshotai/kimi-k3 | 1.024M | $3 | $15 | — | |||
| NousResearch/Hermes-4-405Bnebius/nousresearch/hermes-4-405b | 131.072K | $1 | $3 | — | |||
| NousResearch/Hermes-4-70Bnebius/nousresearch/hermes-4-70b | 131.072K | $0.13 | $0.4 | — | |||
| nvidia/Cosmos3-Super-Reasonernebius/nvidia/cosmos3-super-reasoner | 262.144K | $0.1 | $0.3 | — | |||
| nvidia/Llama-3_1-Nemotron-Ultra-253B-v1nebius/nvidia/llama-3_1-nemotron-ultra-253b-v1 | 131.072K | $0.6 | $1.8 | — | |||
| nvidia/NVIDIA-Nemotron-3-Nano-30B-A3Bnebius/nvidia/nvidia-nemotron-3-nano-30b-a3b | 262.144K | $0.06 | $0.24 | — | |||
| Qwen/Qwen3-Maxdeepinfra/qwen/qwen3-max | 256K | $1.2 | $6 | — | |||
| nvidia/nemotron-3-super-120b-a12bnebius/nvidia/nemotron-3-super-120b-a12b | 262.144K | $0.3 | $0.9 | — | |||
| meta-llama/llama-3-8b-instructnovita/meta-llama/llama-3-8b-instruct | 8.192K | $0.04 | $0.04 | — | |||
| nvidia/Nemotron-3_5-Lightningnebius/nvidia/nemotron-3_5-lightning | 1.04858M | $0.06 | $0.24 | — | |||
| XiaomiMiMo/MiMo-V2.5deepinfra/xiaomimimo/mimo-v2.5 | 262.144K | $0.4 | $2 | — | |||
| deepseek/deepseek-r1-distill-qwen-32bnovita/deepseek/deepseek-r1-distill-qwen-32b | 64K | $0.3 | $0.3 | — | |||
| accounts/fireworks/models/mistral-small-24b-instruct-2501fireworks_ai/accounts/fireworks/models/mistral-small-24b-instruct-2501 | 32.768K | $0.9 | $0.9 | — | |||
| Qwen: Qwen3 235B A22Bqwen/qwen3-235b-a22b | 131.072K | $0.455 | $1.82 | — | |||
| OpenAI: GPT-5.1 Chatopenai/gpt-5.1-chat | 128K | $1.25 | $10 | — | |||
| deepseek/deepseek-r1-0528novita/deepseek/deepseek-r1-0528 | 163.84K | $0.7 | $2.5 | — | |||
| google/gemini-3.5-flashdeepinfra/google/gemini-3.5-flash | 1M | $1.5 | $9 | — | |||
| anthropic/claude-sonnet-4-6deepinfra/anthropic/claude-sonnet-4-6 | 1M | $3 | $15 | — | |||
| minimaxai/minimax-m1-80knovita/minimaxai/minimax-m1-80k | 1M | $0.55 | $2.2 | — | |||
| accounts/fireworks/models/mistral-nemo-instruct-2407fireworks_ai/accounts/fireworks/models/mistral-nemo-instruct-2407 | 128K | $0.2 | $0.2 | — | |||
| accounts/fireworks/models/deepseek-coder-1b-basefireworks_ai/accounts/fireworks/models/deepseek-coder-1b-base | 16.384K | $0.1 | $0.1 | — | |||
| grok-vision-betaxai/grok-vision-beta | 8.192K | $5 | $15 | — | |||
| grok-4xai/grok-4 | 256K | $3 | $15 | — | |||
| grok-3-betaxai/grok-3-beta | 131.072K | $3 | $15 | — | |||
| grok-2-latestxai/grok-2-latest | 131.072K | $2 | $10 | — | |||
| anthropic/claude-opus-4-8deepinfra/anthropic/claude-opus-4-8 | 1M | $5 | $25 | — | |||
| zai-org/GLM-4.6deepinfra/zai-org/glm-4.6 | 202.752K | $0.5 | $2 | — | |||
| mistralai/mistral-nemonovita/mistralai/mistral-nemo | 60.288K | $0.04 | $0.17 | — | |||
| zai-org/GLM-4.7-Flashdeepinfra/zai-org/glm-4.7-flash | 202.752K | $0.06 | $0.4 | — | |||
| moonshotai/Kimi-K2.5deepinfra/moonshotai/kimi-k2.5 | 262.144K | $0.45 | $2.25 | — | |||
| qwen/qwen-2.5-72b-instructnovita/qwen/qwen-2.5-72b-instruct | 32K | $0.38 | $0.4 | — | |||
| accounts/fireworks/models/mistral-nemo-base-2407fireworks_ai/accounts/fireworks/models/mistral-nemo-base-2407 | 128K | $0.2 | $0.2 | — | |||
| google/gemma-4-31B-it-Ultradeepinfra/google/gemma-4-31b-it-ultra | 131.072K | $0.27 | $0.76 | — | |||
| OpenAI: GPT-5 Codex (batch)openai/gpt-5-codex:batch | 400K | $0.625 | $5 | — | |||
| meta-llama/llama-3.3-70b-instructnovita/meta-llama/llama-3.3-70b-instruct | 131.072K | $0.135 | $0.4 | — | |||
| Qwen/Qwen3.8-27Bdeepinfra/qwen/qwen3.8-27b | 262.144K | $0.4 | $3 | — | |||
| MiniMaxAI/MiniMax-M2.7deepinfra/minimaxai/minimax-m2.7 | 196.608K | $0.25 | $1 | — | |||
| deepseek/deepseek-r1-distill-qwen-14bnovita/deepseek/deepseek-r1-distill-qwen-14b | 32.768K | $0.15 | $0.15 | — |