No provider description is available for this model yet.
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Gemini 3.1 Pro Preview is Google’s frontier reasoning model, delivering enhanced software engineering performance, improved agentic reliability, and more efficient token usage across complex workflows. Building on the multimodal foundation...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Gemini 3.5 Flash is Google's high-efficiency multimodal model, bringing near-Pro level coding and reasoning at Flash-tier cost and speed. It is highly optimized for coding proficiency and parallel agentic execution...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
The latest GPT-4 Turbo model with vision capabilities. Vision requests can now use JSON mode and function calling. Training data: up to December 2023.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Qwen3-Max is an updated release built on the Qwen3 series, offering major improvements in reasoning, instruction following, multilingual support, and long-tail knowledge coverage compared to the January 2025 version. It...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Relace Apply 3 is a specialized code-patching LLM that merges AI-suggested edits straight into your source files. It can apply updates from GPT-4o, Claude, and others into your files at...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Mistral Large 3 2512 is Mistral’s most capable model to date, featuring a sparse mixture-of-experts architecture with 41B active parameters (675B total), and released under the Apache 2.0 license.
No provider description is available for this model yet.
No provider description is available for this model yet.
Mistral Medium 3 is a high-performance enterprise-grade language model designed to deliver frontier-level capabilities at significantly reduced operational cost. It balances state-of-the-art reasoning and multimodal performance with 8× lower cost...
No provider description is available for this model yet.
MiniMax-M3 is a multimodal foundation model from MiniMax. It supports text, image, and video inputs with text output, a 1M-token context window, and is suited for long-horizon agentic work, coding,...
Claude Opus 4.8 is Anthropic's most capable generally available model in the Opus family. It supports text, image, and file inputs with text output, with reasoning support and a 1M-token...
No provider description is available for this model yet.
No provider description is available for this model yet.
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| accounts/fireworks/models/glm-5p3fireworks_ai/accounts/fireworks/models/glm-5p3 | 1.04858M | $1.4 | $4.4 | — | |||
| zai-org/GLM-5.1deepinfra/zai-org/glm-5.1 | 202.752K | $1.05 | $3.5 | — | |||
| moonshotai/Kimi-K2.7-Codewandb/moonshotai/kimi-k2.7-code | 262.144K | $0.71 | $3.5 | — | |||
| Qwen/Qwen3.5-122B-A10Bdeepinfra/qwen/qwen3.5-122b-a10b | 262.144K | $0.29 | $2.4 | — | |||
| MiniMaxAI/MiniMax-M3wandb/minimaxai/minimax-m3 | 262.144K | $0.23 | $0.96 | — | |||
| google/gemma-4-26b-a4b-itnovita/google/gemma-4-26b-a4b-it | 262.144K | $0.13 | $0.4 | — | |||
| qwen/qwen3.8-maxnovita/qwen/qwen3.8-max | 1M | $2 | $6 | — | |||
| Google: Gemini 3.1 Pro Preview (batch)google/gemini-3.1-pro-preview:batch | 1.04858M | $1 | $6 | — | |||
| deepseek/deepseek-v4-pro-0813novita/deepseek/deepseek-v4-pro-0813 | 1.04858M | $1.32 | $3.96 | — | |||
| glm-5.3zai/glm-5.3 | 1M | $1.4 | $4.4 | — | |||
| nvidia/NVIDIA-Nemotron-3-Ultra-550B-A55Bdeepinfra/nvidia/nvidia-nemotron-3-ultra-550b-a55b | 262.144K | $0.5 | $2.2 | — | |||
| Google: Gemini 3.5 Flash (batch)google/gemini-3.5-flash:batch | 1.04858M | $0.75 | $4.5 | — | |||
| meta-llama/Llama-3.1-70B-Instructwandb/meta-llama/llama-3.1-70b-instruct | 128K | $0.8 | $0.8 | — | |||
| anthropic/claude-fable-5deepinfra/anthropic/claude-fable-5 | 1M | $10 | $50 | — | |||
| Qwen/Qwen3.8-Maxdeepinfra/qwen/qwen3.8-max | 256K | $1.65 | $4.951 | — | |||
| JetBrains/Mellum2-12B-A2.5B-Instructwandb/jetbrains/mellum2-12b-a2.5b-instruct | 131.072K | $0.05 | $0.1 | — | |||
| OpenAI: GPT-4 Turbo (batch)openai/gpt-4-turbo:batch | 128K | $5 | $15 | — | |||
| deepseek-ai/DeepSeek-V3.2deepinfra/deepseek-ai/deepseek-v3.2 | 163.84K | $0.26 | $0.38 | — | |||
| google/gemma-4-E4B-itdeepinfra/google/gemma-4-e4b-it | 131.072K | $0.02 | $0.1 | — | |||
| ibm-granite/granite-4.1-8bwandb/ibm-granite/granite-4.1-8b | 131.072K | $0.05 | $0.1 | — | |||
| deepseek-ai/DeepSeek-V4-Flash-0731deepinfra/deepseek-ai/deepseek-v4-flash-0731 | 1.04858M | $0.08 | $0.18 | — | |||
| labs-leanstral-1-5-1mistral/labs-leanstral-1-5-1 | 262.144K | — | — | — | |||
| google/gemma-4-31B-itwandb/google/gemma-4-31b-it | 262.144K | $0.1 | $0.34 | — | |||
| Qwen: Qwen3 Maxqwen/qwen3-max | 262.144K | $0.78 | $3.9 | — | |||
| inclusionai/ling-3.0-flash-fastnovita/inclusionai/ling-3.0-flash-fast | 262.144K | $0.06 | $0.18 | — | |||
| mistral-code-agent-latestmistral/mistral-code-agent-latest | 256K | $0.4 | $2 | — | |||
| mistral-code-fim-latestmistral/mistral-code-fim-latest | 128K | $0.3 | $0.9 | — | |||
| deepseek-ai/DeepSeek-V4-Prowandb/deepseek-ai/deepseek-v4-pro | 1.04858M | $1.15 | $2.55 | — | |||
| mistral-code-latestmistral/mistral-code-latest | 128K | $0.3 | $0.9 | — | |||
| deepseek-ai/DeepSeek-V4-Flash-0731wandb/deepseek-ai/deepseek-v4-flash-0731 | 262.144K | $0.13 | $0.28 | — | |||
| zai-org/glm-5v-turbonovita/zai-org/glm-5v-turbo | 204.8K | $1.2 | $4 | — | |||
| Qwen/Qwen3.5-397B-A17Bdeepinfra/qwen/qwen3.5-397b-a17b | 262.144K | $0.45 | $3 | — | |||
| mistral-vibe-cli-fastmistral/mistral-vibe-cli-fast | 262.144K | $0.15 | $0.6 | — | |||
| deepseek-ai/DeepSeek-V4-Flashwandb/deepseek-ai/deepseek-v4-flash | 1.04858M | $0.14 | $0.28 | — | |||
| anthropic/claude-sonnet-5deepinfra/anthropic/claude-sonnet-5 | 1M | $2 | $10 | — | |||
| moonshotai/Kimi-K2.7-Codedeepinfra/moonshotai/kimi-k2.7-code | 262.144K | $0.68 | $3.4 | — | |||
| meta-llama/llama-3.2-1b-instructnovita/meta-llama/llama-3.2-1b-instruct | 131K | $0.02 | $0.02 | — | |||
| Relace: Relace Apply 3relace/relace-apply-3 | 256K | $0.85 | $1.25 | — | |||
| deepseek/deepseek-v4-pronovita/deepseek/deepseek-v4-pro | 1.04858M | $1.6 | $3.2 | — | |||
| moonshotai/kimi-k2.7-codenovita/moonshotai/kimi-k2.7-code | 262.144K | $0.95 | $4 | — | |||
| nvidia/Nemotron-3-Nano-30B-A3Bdeepinfra/nvidia/nemotron-3-nano-30b-a3b | 262.144K | $0.05 | $0.2 | — | |||
| Mistral: Mistral Large 3 2512 (batch)mistralai/mistral-large-2512:batch | 262.144K | $0.25 | $0.75 | — | |||
| zai-org/GLM-5deepinfra/zai-org/glm-5 | 202.752K | $0.6 | $2.08 | — | |||
| thudm/glm-4-32b-0414novita/thudm/glm-4-32b-0414 | 32K | $0.55 | $1.66 | — | |||
| Mistral: Mistral Medium 3mistralai/mistral-medium-3 | 131.072K | $0.4 | $2 | — | |||
| ByteDance/Seed-2.0-prodeepinfra/bytedance/seed-2.0-pro | 256K | $0.5 | $3 | — | |||
| MiniMax: MiniMax M3 (free)minimax/minimax-m3:free | 1.04858M | Free | Free | — | |||
| Anthropic: Claude Opus 4.8 (batch)anthropic/claude-opus-4.8:batch | 1M | $2.5 | $12.5 | — | |||
| ByteDance/Seed-2.0-codedeepinfra/bytedance/seed-2.0-code | 256K | $0.5 | $3 | — | |||
| deepseek/deepseek-r1/communitynovita/deepseek/deepseek-r1/community | 64K | $4 | $4 | — |