Claude Opus 4.5 is Anthropic’s frontier reasoning model optimized for complex software engineering, agentic workflows, and long-horizon computer use. It offers strong multimodal capabilities, competitive performance across real-world coding and...
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
Claude Opus 4.5 is Anthropic’s frontier reasoning model optimized for complex software engineering, agentic workflows, and long-horizon computer use. It offers strong multimodal capabilities, competitive performance across real-world coding and...
Open MiMo model for multimodal coding agents and long-context automation
MiniMax-M3 is a multimodal foundation model from MiniMax. It supports text, image, and video inputs with text output, a 1M-token context window, and is suited for long-horizon agentic work, coding,...
MiniMax-M3 is a multimodal foundation model from MiniMax. It supports text, image, and video inputs with text output, a 1M-token context window, and is suited for long-horizon agentic work, coding,...
MiniMax-M3 is a multimodal foundation model from MiniMax. It supports text, image, and video inputs with text output, a 1M-token context window, and is suited for long-horizon agentic work, coding,...
Grok 4.20 is a reasoning model from SpaceXAI with industry-leading speed and agentic tool calling capabilities. It combines the lowest hallucination rate on the market with strict prompt adherance, delivering...
Qwen 3.6 Plus builds on a hybrid architecture that combines efficient linear attention with sparse mixture-of-experts routing, enabling strong scalability and high-performance inference. Compared to the 3.5 series, it delivers...
Fast DeepSeek V4 lane for economical reasoning, coding, and long-context work
Claude Opus 4.1 is an updated version of Anthropic’s flagship model, offering improved performance in coding, reasoning, and agentic tasks. It achieves 74.5% on SWE-bench Verified and shows notable gains...
Claude Opus 4.1 is an updated version of Anthropic’s flagship model, offering improved performance in coding, reasoning, and agentic tasks. It achieves 74.5% on SWE-bench Verified and shows notable gains...
GLM-5V-Turbo is Z.ai’s first native multimodal agent foundation model, built for vision-based coding and agent-driven tasks. It natively handles image, video, and text inputs, excels at long-horizon planning, complex coding,...
Earlier Kimi frontier model for long-context agents, coding, and multimodal work
MiniMax-M2.5 is a SOTA large language model designed for real-world productivity. Trained in a diverse range of complex real-world digital working environments, M2.5 builds upon the coding expertise of M2.1...
GPT-5.1 is the latest frontier-grade model in the GPT-5 series, offering stronger general-purpose reasoning, improved instruction adherence, and a more natural conversational style compared to GPT-5. It uses adaptive reasoning...
Sharper GPT-5 generation for coding, product work, and tool-assisted tasks
GLM-4.7 is Z.ai’s latest flagship model, featuring upgrades in two key areas: enhanced programming capabilities and more stable multi-step reasoning/execution. It demonstrates significant improvements in executing complex agent tasks while...
Open MoE flagship with million-token context for coding and long agent runs
The Qwen3.5 series 397B-A17B native vision-language model is built on a hybrid architecture that integrates a linear attention mechanism with a sparse mixture-of-experts model, achieving higher inference efficiency. It delivers...
Reliable GPT generation for broad coding, writing, and tool-assisted product work
GPT-5.2 is the latest frontier-grade model in the GPT-5 series, offering stronger agentic and long context perfomance compared to GPT-5.1. It uses adaptive reasoning to allocate computation dynamically, responding quickly...
MiniMax-M2.7 is a next-generation large language model designed for autonomous, real-world productivity and continuous improvement. Built to actively participate in its own evolution, M2.7 integrates advanced agentic capabilities through multi-agent...
MiniMax-M2.1 is a lightweight, state-of-the-art large language model optimized for coding, agentic workflows, and modern application development. With only 10 billion activated parameters, it delivers a major jump in real-world...
Coding-optimized GPT model for repository edits, reviews, and agentic software work
Claude Opus 4 is benchmarked as the world’s best coding model, at time of release, bringing sustained performance on complex, long-running tasks and agent workflows. It sets new benchmarks in...
Claude Sonnet 4.5 is Anthropic’s most advanced Sonnet model to date, optimized for real-world agents and coding workflows. It delivers state-of-the-art performance on coding benchmarks such as SWE-bench Verified, with...
Claude Sonnet 4.5 is Anthropic’s most advanced Sonnet model to date, optimized for real-world agents and coding workflows. It delivers state-of-the-art performance on coding benchmarks such as SWE-bench Verified, with...
Inkling is an open-weight multimodal mixture-of-experts model from Thinking Machines Lab, with 41B active parameters out of 975B total. It is designed for general-purpose reasoning, coding, agentic and tool-use systems,...
Inkling is an open-weight multimodal mixture-of-experts model from Thinking Machines Lab, with 41B active parameters out of 975B total. It is designed for general-purpose reasoning, coding, agentic and tool-use systems,...
Multimodal MoE reasoning model (975B total, 41B active) for text, image, and audio
The Qwen3.5 native vision-language series Plus models are built on a hybrid architecture that integrates linear attention mechanisms with sparse mixture-of-experts models, achieving higher inference efficiency. In a variety of...
MiniMax-M2 is a compact, high-efficiency large language model optimized for end-to-end coding and agentic workflows. With 10 billion activated parameters (230 billion total), it delivers near-frontier intelligence across general reasoning,...
Small GPT-5 for responsive agents, coding help, and everyday automation
GPT-5 Mini is a compact version of GPT-5, designed to handle lighter-weight reasoning tasks. It provides the same instruction-following and safety-tuning benefits as GPT-5, but with reduced latency and cost....
NVIDIA Nemotron 3 Ultra is an open frontier-reasoning and orchestration model from NVIDIA, with 55B active parameters out of 550B total (MoE). Built on a hybrid Transformer-Mamba mixture-of-experts architecture, it...
Grok 4.3 is a reasoning model from SpaceXAI. It accepts text and image inputs with text output, and is suited for agentic workflows, instruction-following tasks, and applications requiring high factual...
Grok 4.3 is a reasoning model from SpaceXAI. It accepts text and image inputs with text output, and is suited for agentic workflows, instruction-following tasks, and applications requiring high factual...
Claude Sonnet 4 significantly enhances the capabilities of its predecessor, Sonnet 3.7, excelling in both coding and reasoning tasks with improved precision and controllability. Achieving state-of-the-art performance on SWE-bench (72.7%),...
NVIDIA Nemotron 3 Ultra is an open frontier-reasoning and orchestration model from NVIDIA, with 55B active parameters out of 550B total (MoE). Built on a hybrid Transformer-Mamba mixture-of-experts architecture, it...
Largest Nemotron 3 model for maximum open-weight reasoning and agent accuracy
Newer StepFun flash model for faster agents, coding, and multimodal prompts
Low-latency Gemini model for high-volume multimodal and agent workloads
Claude Haiku 4.5 is Anthropic’s fastest and most efficient model, delivering near-frontier intelligence at a fraction of the cost and latency of larger Claude models. Matching Claude Sonnet 4’s performance...
Claude Haiku 4.5 is Anthropic’s fastest and most efficient model, delivering near-frontier intelligence at a fraction of the cost and latency of larger Claude models. Matching Claude Sonnet 4’s performance...
Fast Gemini workhorse for multimodal apps where latency and price matter
Gemini 2.5 Flash is Google's state-of-the-art workhorse model, specifically designed for advanced reasoning, coding, mathematics, and scientific tasks. It includes built-in "thinking" capabilities, enabling it to provide responses with greater...
Reasoning-optimized 398B MoE agent model with extended thinking for long-horizon and multi-turn tool use
Qwen3-Max is an updated release built on the Qwen3 series, offering major improvements in reasoning, instruction following, multilingual support, and long-tail knowledge coverage compared to the January 2025 version. It...
Mistral Large 3 2512 is Mistral’s most capable model to date, featuring a sparse mixture-of-experts architecture with 41B active parameters (675B total), and released under the Apache 2.0 license.
Mistral Large 3 2512 is Mistral’s most capable model to date, featuring a sparse mixture-of-experts architecture with 41B active parameters (675B total), and released under the Apache 2.0 license.
| Model | Creator | Score | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|---|
| Anthropic: Claude Opus 4.5 (batch)anthropic/claude-opus-4.5:batch | 1206.0 | 200K | $2.5 | $12.5 | — | |||
| Anthropic: Claude Opus 4.5anthropic/claude-opus-4.5 | 1206.0 | 200K | $5 | $25 | — | |||
| MiMo-V2.5xiaomi/mimo-v2.5 | 1200.0 | 1.04858M | $0.14 | $0.28 | 2026-04-22 | |||
| MiniMax: MiniMax M3minimax/minimax-m3 | 1194.0 | 524.288K | $0.3 | $1.2 | — | |||
| MiniMax: MiniMax M3 (batch)minimax/minimax-m3:batch | 1194.0 | 524.288K | $0.3 | $1.2 | — | |||
| MiniMax: MiniMax M3 (free)minimax/minimax-m3:free | 1194.0 | 1.04858M | Free | Free | — | |||
| SpaceXAI: Grok 4.20x-ai/grok-4.20 | 1190.0 | 2M | $1.25 | $2.5 | — | |||
| Qwen: Qwen3.6 Plusqwen/qwen3.6-plus | 1187.0 | 1M | $0.325 | $1.95 | — | |||
| DeepSeek V4 Flashdeepseek/deepseek-v4-flash | 1181.0 | 1M | $0.15 | $0.6 | 2026-04-24 | |||
| Anthropic: Claude Opus 4.1 (batch)anthropic/claude-opus-4.1:batch | 1178.0 | 200K | $7.5 | $37.5 | — | |||
| Anthropic: Claude Opus 4.1anthropic/claude-opus-4.1 | 1178.0 | 200K | $15 | $75 | — | |||
| Z.ai: GLM 5V Turboz-ai/glm-5v-turbo | 1176.0 | 202.752K | $1.2 | $4 | — | |||
| Kimi K2.5moonshotai/kimi-k2.5 | 1175.0 | 262.144K | $0.3 | $1.9 | 2026-01 | |||
| MiniMax: MiniMax M2.5minimax/minimax-m2.5 | 1175.0 | 204.8K | $0.3 | $1.2 | — | |||
| OpenAI: GPT-5.1 (batch)openai/gpt-5.1:batch | 1174.0 | 400K | $0.625 | $5 | — | |||
| GPT-5.1openai/gpt-5.1 | 1174.0 | 400K | $1.25 | $10 | 2025-11-13 | |||
| Z.ai: GLM 4.7z-ai/glm-4.7 | 1168.0 | 202.752K | $0.4 | $1.75 | — | |||
| DeepSeek V4 Prodeepseek/deepseek-v4-pro | 1166.0 | 1M | $0.435 | $0.87 | 2026-04-24 | |||
| Qwen: Qwen3.5 397B A17Bqwen/qwen3.5-397b-a17b | 1165.0 | 262.144K | $0.55 | $3.5 | — | |||
| GPT-5.2openai/gpt-5.2 | 1164.0 | 400K | $1.75 | $14 | 2025-12-11 | |||
| OpenAI: GPT-5.2 (batch)openai/gpt-5.2:batch | 1164.0 | 400K | $0.875 | $7 | — | |||
| MiniMax: MiniMax M2.7minimax/minimax-m2.7 | 1162.0 | 204.8K | $0.3 | $1.2 | — | |||
| MiniMax: MiniMax M2.1minimax/minimax-m2.1 | 1157.0 | 204.8K | $0.3 | $1.2 | — | |||
| GPT-5.3 Codexopenai/gpt-5.3-codex | 1154.0 | 400K | $1.75 | $14 | 2026-02-05 | |||
| Anthropic: Claude Opus 4anthropic/claude-opus-4 | 1152.0 | 200K | $15 | $75 | — | |||
| Anthropic: Claude Sonnet 4.5 (batch)anthropic/claude-sonnet-4.5:batch | 1140.0 | 1M | $1.5 | $7.5 | — | |||
| Anthropic: Claude Sonnet 4.5anthropic/claude-sonnet-4.5 | 1140.0 | 1M | $3 | $15 | — | |||
| Thinking Machines: Inkling (batch)thinkingmachines/inkling:batch | 1133.0 | 524.288K | $1 | $4.05 | — | |||
| Thinking Machines: Inkling (free)thinkingmachines/inkling:free | 1133.0 | 1.04858M | Free | Free | — | |||
| Inklingthinkingmachines/inkling | 1133.0 | 1.04858M | $1.87 | $4.68 | 2026-07-15 | |||
| Qwen: Qwen3.5 Plus 2026-02-15qwen/qwen3.5-plus-02-15 | 1133.0 | 1M | $0.26 | $1.56 | — | |||
| MiniMax: MiniMax M2minimax/minimax-m2 | 1125.0 | 204.8K | $0.255 | $1.02 | — | |||
| GPT-5 Miniopenai/gpt-5-mini | 1116.0 | 400K | $0.25 | $2 | 2025-08-07 | |||
| OpenAI: GPT-5 Mini (batch)openai/gpt-5-mini:batch | 1116.0 | 400K | $0.125 | $1 | — | |||
| NVIDIA: Nemotron 3 Ultra (batch)nvidia/nemotron-3-ultra-550b-a55b:batch | 1115.0 | 512.288K | $0.6 | $3.6 | — | |||
| SpaceXAI: Grok 4.3 (batch)x-ai/grok-4.3:batch | 1110.0 | 1M | $1 | $2 | — | |||
| SpaceXAI: Grok 4.3x-ai/grok-4.3 | 1110.0 | 1M | $1.25 | $2.5 | — | |||
| Anthropic: Claude Sonnet 4anthropic/claude-sonnet-4 | 1104.0 | 200K | $3 | $15 | — | |||
| NVIDIA: Nemotron 3 Ultra (free)nvidia/nemotron-3-ultra-550b-a55b:free | 1099.0 | 1M | Free | Free | — | |||
| Nemotron 3 Ultra 550B A55Bnvidia/nemotron-3-ultra-550b-a55b | 1099.0 | 1M | $0.5 | $2.5 | 2026-06-04 | |||
| Step 3.7 Flashstepfun/step-3.7-flash | 1098.0 | 256K | $0.185 | $1.11 | 2026-05-29 | |||
| Gemini 3.1 Flash Lite Previewgoogle/gemini-3.1-flash-lite-preview | 1077.0 | 1.04858M | $0.25 | $1.5 | 2026-03-03 | |||
| Anthropic: Claude Haiku 4.5 (batch)anthropic/claude-haiku-4.5:batch | 1051.0 | 200K | $0.5 | $2.5 | — | |||
| Anthropic: Claude Haiku 4.5anthropic/claude-haiku-4.5 | 1051.0 | 200K | $1 | $5 | — | |||
| Gemini 2.5 Flashgoogle/gemini-2.5-flash | 1045.0 | 1.04858M | $0.3 | $2.5 | 2025-06-17 | |||
| Google: Gemini 2.5 Flash (batch)google/gemini-2.5-flash:batch | 1045.0 | 1.04858M | $0.15 | $1.25 | — | |||
| Trinity Large Thinkingarcee-ai/trinity-large-thinking | 1041.0 | 524.288K | $0.25 | $0.9 | 2026-04-01 | |||
| Qwen: Qwen3 Maxqwen/qwen3-max | 1036.0 | 262.144K | $0.78 | $3.9 | — | |||
| Mistral: Mistral Large 3 2512 (batch)mistralai/mistral-large-2512:batch | 1018.0 | 262.144K | $0.25 | $0.75 | — | |||
| Mistral: Mistral Large 3 2512mistralai/mistral-large-2512 | 1018.0 | 262.144K | $0.5 | $1.5 | — |