No provider description is available for this model yet.
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
GPT-4.1 Mini is a mid-sized model delivering performance competitive with GPT-4o at substantially lower latency and cost. It retains a 1 million token context window and scores 45.1% on hard...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
GPT-5 is OpenAI’s most advanced model, offering major improvements in reasoning, code quality, and user experience. It is optimized for complex tasks that require step-by-step reasoning, instruction following, and accuracy...
No provider description is available for this model yet.
DeepSeek-V3.2-Exp is an experimental large language model released by DeepSeek as an intermediate step between V3.1 and future architectures. It introduces DeepSeek Sparse Attention (DSA), a fine-grained sparse attention mechanism...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
MiniMax-M3 is a multimodal foundation model from MiniMax. It supports text, image, and video inputs with text output, a 1M-token context window, and is suited for long-horizon agentic work, coding,...
No provider description is available for this model yet.
No provider description is available for this model yet.
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| databricks-gemini-3-6-flashdatabricks/databricks-gemini-3-6-flash | 1.04858M | $1.875 | $9.375 | — | |||
| databricks-gemini-3-5-flashdatabricks/databricks-gemini-3-5-flash | 1.04858M | $1.875 | $11.25 | — | |||
| databricks-gemini-3-5-flash-litedatabricks/databricks-gemini-3-5-flash-lite | 1.04858M | $0.375 | $3.125 | — | |||
| databricks-glm-5-3databricks/databricks-glm-5-3 | 1.04858M | $1.4 | $4.4 | — | |||
| databricks-gpt-5-6-soldatabricks/databricks-gpt-5-6-sol | 922K | $4 | $20 | — | |||
| databricks-gpt-5-6-terradatabricks/databricks-gpt-5-6-terra | 922K | $2.5 | $15 | — | |||
| databricks-gpt-5-6-lunadatabricks/databricks-gpt-5-6-luna | 922K | $1 | $6 | — | |||
| databricks-grok-4-6databricks/databricks-grok-4-6 | 500K | $2.5 | $7.5 | — | |||
| databricks-inklingdatabricks/databricks-inkling | 1M | $1 | $4.05 | — | |||
| databricks-qwen35-122b-a10bdatabricks/databricks-qwen35-122b-a10b | 262.144K | $0.22 | $2.2 | — | |||
| deepseek-ai/DeepSeek-V4-Flashnebius/deepseek-ai/deepseek-v4-flash | 1.04858M | $0.14 | $0.28 | — | |||
| deepseek-ai/DeepSeek-V4-Flash-0731nebius/deepseek-ai/deepseek-v4-flash-0731 | 1.024M | $0.14 | $0.28 | — | |||
| deepseek-ai/DeepSeek-V4-Pronebius/deepseek-ai/deepseek-v4-pro | 1.04858M | $1.75 | $3.5 | — | |||
| MiniMaxAI/MiniMax-M2.5nebius/minimaxai/minimax-m2.5 | 196.608K | $0.3 | $1.2 | — | |||
| MiniMaxAI/MiniMax-M3nebius/minimaxai/minimax-m3 | 1.04858M | $0.3 | $1.2 | — | |||
| moonshotai/Kimi-K2.6nebius/moonshotai/kimi-k2.6 | 262.144K | $0.95 | $4 | — | |||
| moonshotai/Kimi-K2.7-Codenebius/moonshotai/kimi-k2.7-code | 262.144K | $0.95 | $4 | — | |||
| moonshotai/Kimi-K3nebius/moonshotai/kimi-k3 | 1.024M | $3 | $15 | — | |||
| mistral-vibe-cli-fastmistral/mistral-vibe-cli-fast | 262.144K | $0.15 | $0.6 | — | |||
| NousResearch/Hermes-4-70Bnebius/nousresearch/hermes-4-70b | 131.072K | $0.13 | $0.4 | — | |||
| nvidia/Cosmos3-Super-Reasonernebius/nvidia/cosmos3-super-reasoner | 262.144K | $0.1 | $0.3 | — | |||
| nvidia/Llama-3_1-Nemotron-Ultra-253B-v1nebius/nvidia/llama-3_1-nemotron-ultra-253b-v1 | 131.072K | $0.6 | $1.8 | — | |||
| nvidia/NVIDIA-Nemotron-3-Nano-30B-A3Bnebius/nvidia/nvidia-nemotron-3-nano-30b-a3b | 262.144K | $0.06 | $0.24 | — | |||
| nvidia/Nemotron-3-Nano-Omninebius/nvidia/nemotron-3-nano-omni | 262.144K | $0.06 | $0.24 | — | |||
| nvidia/nemotron-3-super-120b-a12bnebius/nvidia/nemotron-3-super-120b-a12b | 262.144K | $0.3 | $0.9 | — | |||
| nvidia/Nemotron-3-Ultra-550b-a55bnebius/nvidia/nemotron-3-ultra-550b-a55b | 1.04858M | $1 | $3 | — | |||
| nvidia/Nemotron-3_5-Lightningnebius/nvidia/nemotron-3_5-lightning | 1.04858M | $0.06 | $0.24 | — | |||
| openai/gpt-oss-120bnebius/openai/gpt-oss-120b | 131.072K | $0.15 | $0.6 | — | |||
| openbmb/MiniCPM-V-4_5nebius/openbmb/minicpm-v-4_5 | 32K | $0.658 | $1.11 | — | |||
| Qwen/Qwen3-235B-A22B-Instruct-2507nebius/qwen/qwen3-235b-a22b-instruct-2507 | 262.144K | $0.2 | $0.6 | — | |||
| Qwen/Qwen3-30B-A3B-Instruct-2507nebius/qwen/qwen3-30b-a3b-instruct-2507 | 262.144K | $0.1 | $0.3 | — | |||
| Qwen/Qwen3-Next-80B-A3B-Thinkingnebius/qwen/qwen3-next-80b-a3b-thinking | 128K | $0.15 | $1.2 | — | |||
| OpenAI: GPT-4.1 Mini (batch)openai/gpt-4.1-mini:batch | 1.04758M | $0.2 | $0.8 | — | |||
| zai-org/GLM-5.1nebius/zai-org/glm-5.1 | 202.752K | $1.4 | $4.4 | — | |||
| zai-org/GLM-5.2nebius/zai-org/glm-5.2 | 1.04858M | $1.4 | $4.4 | — | |||
| zai-org/GLM-5.3-Flashnebius/zai-org/glm-5.3-flash | 1.024M | $0.15 | $0.5 | — | |||
| meta-llama/llama-4-maverick-17b-128e-instruct-fp8watsonx/meta-llama/llama-4-maverick-17b-128e-instruct-fp8 | 131.072K | $0.371 | $1.484 | — | |||
| claude-mythos-5-1anthropic/claude-mythos-5-1 | 1M | $10 | $50 | — | |||
| accounts/fireworks/models/deepseek-v4-flash-vision-expfireworks_ai/accounts/fireworks/models/deepseek-v4-flash-vision-exp | 1.04858M | $0.22 | $0.66 | — | |||
| deepseek-v4-flash-vision-expfireworks_ai/deepseek-v4-flash-vision-exp | 1.04858M | $0.22 | $0.66 | — | |||
| OpenAI: GPT-5 (batch)openai/gpt-5:batch | 400K | $0.625 | $5 | — | |||
| lyria-3.5-clip-previewgemini/lyria-3.5-clip-preview | 131.072K | — | — | — | |||
| DeepSeek: DeepSeek V3.2 Expdeepseek/deepseek-v3.2-exp | 163.84K | $0.27 | $0.41 | — | |||
| anthropic/claude-sonnet-5deepinfra/anthropic/claude-sonnet-5 | 1M | $2 | $10 | — | |||
| anthropic/claude-fable-5openrouter/anthropic/claude-fable-5 | 1M | $10 | $50 | — | |||
| anthropic/claude-fable-5.1openrouter/anthropic/claude-fable-5.1 | 1M | $10 | $50 | — | |||
| google/gemma-4-31B-itwandb/google/gemma-4-31b-it | 262.144K | $0.1 | $0.34 | — | |||
| MiniMax: MiniMax M3 (free)minimax/minimax-m3:free | 1.04858M | Free | Free | — | |||
| anthropic/claude-sonnet-5openrouter/anthropic/claude-sonnet-5 | 1M | $2 | $10 | — | |||
| moonshotai/Kimi-K2.7-Codedeepinfra/moonshotai/kimi-k2.7-code | 262.144K | $0.68 | $3.4 | — |