No provider description is available for this model yet.
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
No provider description is available for this model yet.
Gemma 4 31B Instruct is Google DeepMind's 30.7B dense multimodal model supporting text and image input with text output. Features a 256K token context window, configurable thinking/reasoning mode, native function...
No provider description is available for this model yet.
No provider description is available for this model yet.
A balanced model in the Ministral 3 family, Ministral 3 8B is a powerful, efficient tiny language model with vision capabilities.
No provider description is available for this model yet.
No provider description is available for this model yet.
GLM 5.2 is a large-scale reasoning model from Z.ai. It supports text input and output with a 1M-token context window, and is suited for long-horizon agent workflows, project-level software engineering,...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Claude Sonnet 4.5 is Anthropic’s most advanced Sonnet model to date, optimized for real-world agents and coding workflows. It delivers state-of-the-art performance on coding benchmarks such as SWE-bench Verified, with...
MiniMax-M3 is a multimodal foundation model from MiniMax. It supports text, image, and video inputs with text output, a 1M-token context window, and is suited for long-horizon agentic work, coding,...
Nex-N2.5 is an agentic model built to turn goals into working, verified outcomes. Its core strength is agentic coding within a visual feedback loop: it can explore codebases, implement multi-file...
Claude Fable 5.1 improves on Claude Fable 5 across the board, with the biggest gains in agentic coding, long-running agentic workflows, and knowledge work: long code refactors, front-end and visual...
Muse Glimmer 30B is a dense, open-weight multimodal model from Meta Superintelligence Labs, distilled from Muse Spark and optimized for autonomous agents on consumer hardware. It is suited for long-horizon...
Gemini 2.5 Pro is Google’s state-of-the-art AI model designed for advanced reasoning, coding, mathematics, and scientific tasks. It employs “thinking” capabilities, enabling it to reason through responses with enhanced accuracy...
Qwen3-14B is a dense 14.8B parameter causal language model from the Qwen3 series, designed for both complex reasoning and efficient dialogue. It supports seamless switching between a "thinking" mode for...
Gemini 3.5 Flash Lite is a high-efficiency model from Google with upgraded agentic capabilities. It is suited for subagents that execute focused tasks within complex, multi-agent workflows.
Gemini 3 Flash Preview is a high speed, high value thinking model designed for agentic workflows, multi turn chat, and coding assistance. It delivers near Pro level reasoning and tool...
Claude Opus 4.5 is Anthropic’s frontier reasoning model optimized for complex software engineering, agentic workflows, and long-horizon computer use. It offers strong multimodal capabilities, competitive performance across real-world coding and...
DeepSeek V4 Flash 0731 is a sparse mixture-of-experts model from DeepSeek, with 13B active parameters out of 284B total. This re-post-trained revision is suited for coding, reasoning, and agent workflows....
Claude Haiku 4.5 is Anthropic’s fastest and most efficient model, delivering near-frontier intelligence at a fraction of the cost and latency of larger Claude models. Matching Claude Sonnet 4’s performance...
GPT-5 is OpenAI’s most advanced model, offering major improvements in reasoning, code quality, and user experience. It is optimized for complex tasks that require step-by-step reasoning, instruction following, and accuracy...
GPT-5 Mini is a compact version of GPT-5, designed to handle lighter-weight reasoning tasks. It provides the same instruction-following and safety-tuning benefits as GPT-5, but with reduced latency and cost....
GPT-5.2 Pro is OpenAI’s most advanced model, offering major improvements in agentic coding and long context performance over GPT-5 Pro. It is optimized for complex tasks that require step-by-step reasoning,...
Relace Apply 3 is a specialized code-patching LLM that merges AI-suggested edits straight into your source files. It can apply updates from GPT-4o, Claude, and others into your files at...
GPT-5.4 Pro is OpenAI's most advanced model, building on GPT-5.4's unified architecture with enhanced reasoning capabilities for complex, high-stakes tasks. It features a 1M+ token context window (922K input, 128K...
GPT-5.6 Terra is a balanced model in OpenAI's GPT-5.6 series, positioned between the flagship Sol tier and the cost-efficient Luna tier. It is suited for everyday coding, reasoning, and agentic...
NVIDIA Nemotron 3 Ultra is an open frontier-reasoning and orchestration model from NVIDIA, with 55B active parameters out of 550B total (MoE). Built on a hybrid Transformer-Mamba mixture-of-experts architecture, it...
For tasks that demand low latency, GPT‑4.1 nano is the fastest and cheapest model in the GPT-4.1 series. It delivers exceptional performance at a small size with its 1 million...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| qwen3-max-2026-01-23qwen_ai_platform/qwen3-max-2026-01-23 | 258.048K | — | — | — | |||
| qwen3-next-80b-a3b-instructqwen_ai_platform/qwen3-next-80b-a3b-instruct | 262.144K | $0.15 | $1.2 | — | |||
| Google: Gemma 4 31B (batch)google/gemma-4-31b-it:batch | 262.144K | $0.39 | $0.97 | — | |||
| qwen3-next-80b-a3b-thinkingqwen_ai_platform/qwen3-next-80b-a3b-thinking | 262.144K | $0.15 | $1.2 | — | |||
| qwen3-vl-235b-a22b-instructqwen_ai_platform/qwen3-vl-235b-a22b-instruct | 131.072K | $0.4 | $1.6 | — | |||
| Mistral: Ministral 3 8B 2512 (batch)mistralai/ministral-8b-2512:batch | 262.144K | $0.075 | $0.075 | — | |||
| qwen3-vl-235b-a22b-thinkingqwen_ai_platform/qwen3-vl-235b-a22b-thinking | 131.072K | $0.4 | $4 | — | |||
| qwen3-vl-32b-instructqwen_ai_platform/qwen3-vl-32b-instruct | 131.072K | $0.16 | $0.64 | — | |||
| Z.ai: GLM 5.2 (batch)z-ai/glm-5.2:batch | 1.04858M | $0.7 | $2.2 | — | |||
| qwen3-vl-32b-thinkingqwen_ai_platform/qwen3-vl-32b-thinking | 131.072K | $0.16 | $2.87 | — | |||
| qwen3-vl-plusqwen_ai_platform/qwen3-vl-plus | 260.096K | — | — | — | |||
| qwen3.5-plusqwen_ai_platform/qwen3.5-plus | 991.808K | — | — | — | |||
| qwen3.7-maxqwen_ai_platform/qwen3.7-max | 991.808K | $2.5 | $7.5 | — | |||
| qwen3.7-plusqwen_ai_platform/qwen3.7-plus | 991.808K | — | — | — | |||
| qwen3.8-maxqwen_ai_platform/qwen3.8-max | 991.808K | $2 | $6 | — | |||
| qwq-plusqwen_ai_platform/qwq-plus | 98.304K | $0.8 | $2.4 | — | |||
| databricks-deepseek-v4-flash-0731databricks/databricks-deepseek-v4-flash-0731 | 1M | $0.14 | $0.28 | — | |||
| databricks-deepseek-v4-pro-0813databricks/databricks-deepseek-v4-pro-0813 | 1M | $1.32 | $3.96 | — | |||
| zai-org/GLM-5.3-Flashfriendliai/zai-org/glm-5.3-flash | 1.04858M | $0.15 | $0.5 | — | |||
| zai-org/GLM-5.3friendliai/zai-org/glm-5.3 | 1.04858M | $1.26 | $3.96 | — | |||
| GigaChat-2gigachat/gigachat-2 | 128K | — | — | — | |||
| claude-fable-5-1vertex_ai-anthropic_models/claude-fable-5-1 | 1M | $10 | $50 | — | |||
| claude-fable-5-1@defaultvertex_ai-anthropic_models/claude-fable-5-1@default | 1M | $10 | $50 | — | |||
| glm-5.2zai/glm-5.2 | 1M | $1.4 | $4.4 | — | |||
| Qwen/Qwen3.8-Flashtogether_ai/qwen/qwen3.8-flash | 1M | $0.15 | $0.47 | — | |||
| gemma-4-31bcerebras/gemma-4-31b | 131.072K | $0.99 | $1.49 | — | |||
| Anthropic: Claude Sonnet 4.5anthropic/claude-sonnet-4.5 | 1M | $3 | $15 | — | |||
| MiniMax: MiniMax M3 (free)minimax/minimax-m3:free | 1.04858M | Free | Free | — | |||
| Nex AGI: Nex-N2.5-Pro (free)nex-agi/nex-n2.5-pro:free | 262.144K | Free | Free | — | |||
| Anthropic: Claude Fable 5.1 (batch)anthropic/claude-fable-5.1:batch | 1M | $5 | $25 | — | |||
| Meta: Muse Glimmer 30B (batch)meta/muse-glimmer-30b:batch | 131.072K | $0.175 | $0.75 | — | |||
| Google: Gemini 2.5 Pro Preview 05-06google/gemini-2.5-pro-preview-05-06 | 1.04858M | $1.25 | $10 | — | |||
| Qwen: Qwen3 14Bqwen/qwen3-14b | 40.96K | $0.12 | $0.24 | — | |||
| Google: Gemini 3.5 Flash Lite (batch)google/gemini-3.5-flash-lite:batch | 1.04858M | $0.15 | $1.25 | — | |||
| Google: Gemini 3 Flash Preview (batch)google/gemini-3-flash-preview:batch | 1.04858M | $0.25 | $1.5 | — | |||
| Anthropic: Claude Opus 4.5 (batch)anthropic/claude-opus-4.5:batch | 200K | $2.5 | $12.5 | — | |||
| DeepSeek: DeepSeek V4 Flash 0731 (batch)deepseek/deepseek-v4-flash-0731:batch | 1.04858M | $0.11 | $0.33 | — | |||
| Anthropic: Claude Haiku 4.5 (batch)anthropic/claude-haiku-4.5:batch | 200K | $0.5 | $2.5 | — | |||
| OpenAI: GPT-5 (batch)openai/gpt-5:batch | 400K | $0.625 | $5 | — | |||
| OpenAI: GPT-5 Mini (batch)openai/gpt-5-mini:batch | 400K | $0.125 | $1 | — | |||
| OpenAI: GPT-5.2 Pro (batch)openai/gpt-5.2-pro:batch | 400K | $10.5 | $84 | — | |||
| Relace: Relace Apply 3relace/relace-apply-3 | 256K | $0.85 | $1.25 | — | |||
| OpenAI: GPT-5.4 Pro (batch)openai/gpt-5.4-pro:batch | 1.05M | $15 | $90 | — | |||
| OpenAI: GPT-5.6 Terra (batch)openai/gpt-5.6-terra:batch | 1.05M | $1 | $6 | — | |||
| NVIDIA: Nemotron 3 Ultra (free)nvidia/nemotron-3-ultra-550b-a55b:free | 1M | Free | Free | — | |||
| OpenAI: GPT-4.1 Nano (batch)openai/gpt-4.1-nano:batch | 1.04758M | $0.05 | $0.2 | — | |||
| gpt-chat-latestazure_ai/gpt-chat-latest | 272K | $5 | $30 | — | |||
| model-routerazure_ai/model-router | 200K | $0.14 | — | — | |||
| cohere-command-aazure_ai/cohere-command-a | 131.072K | $2.5 | $10 | — | |||
| xai/grok-4.3vertex_ai/xai/grok-4.3 | 200K | $1.25 | $2.5 | — |