GPT-5.4 is OpenAI’s latest frontier model, unifying the Codex and GPT lines into a single system. It features a 1M+ token context window (922K input, 128K output) with support for...
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
Agent-ready GPT for coding and computer-use workflows at a lower cost
Claude Opus 4.5 is Anthropic’s frontier reasoning model optimized for complex software engineering, agentic workflows, and long-horizon computer use. It offers strong multimodal capabilities, competitive performance across real-world coding and...
Claude Opus 4.5 is Anthropic’s frontier reasoning model optimized for complex software engineering, agentic workflows, and long-horizon computer use. It offers strong multimodal capabilities, competitive performance across real-world coding and...
DeepSeek V4 Flash 0731 is a sparse mixture-of-experts model from DeepSeek, with 13B active parameters out of 284B total. This re-post-trained revision is suited for coding, reasoning, and agent workflows....
Codex GPT for repository edits, code review, and practical software agents
Qwen 3.6 Plus builds on a hybrid architecture that combines efficient linear attention with sparse mixture-of-experts routing, enabling strong scalability and high-performance inference. Compared to the 3.5 series, it delivers...
GLM-5 is Z.ai’s flagship open-source foundation model engineered for complex systems design and long-horizon agent workflows. Built for expert developers, it delivers production-grade performance on large-scale programming tasks, rivaling leading...
MiniMax-M2.1 is a lightweight, state-of-the-art large language model optimized for coding, agentic workflows, and modern application development. With only 10 billion activated parameters, it delivers a major jump in real-world...
Nex-N2-Pro is an agentic mixture-of-experts model from Nex AGI, with 17B active parameters out of 397B total. Built on the Qwen3.5 architecture, it accepts text and image input and produces...
Hy3 is a 295B-parameter Mixture-of-Experts model from Tencent (21B active, 192 experts with top-8 routing) built for reasoning, agentic workflows, and real-world production use. It supports a configurable reasoning effort:...
MiniMax-M2.7 is a next-generation large language model designed for autonomous, real-world productivity and continuous improvement. Built to actively participate in its own evolution, M2.7 integrates advanced agentic capabilities through multi-agent...
GLM-5V-Turbo is Z.ai’s first native multimodal agent foundation model, built for vision-based coding and agent-driven tasks. It natively handles image, video, and text inputs, excels at long-horizon planning, complex coding,...
MiniMax-M2.7 is a next-generation large language model designed for autonomous, real-world productivity and continuous improvement. Built to actively participate in its own evolution, M2.7 integrates advanced agentic capabilities through multi-agent...
As a 30B-class SOTA model, GLM-4.7-Flash offers a new option that balances performance and efficiency. It is further optimized for agentic coding use cases, strengthening coding capabilities, long-horizon task planning,...
GLM-4.7 is Z.ai’s latest flagship model, featuring upgrades in two key areas: enhanced programming capabilities and more stable multi-step reasoning/execution. It demonstrates significant improvements in executing complex agent tasks while...
Grok 4.20 is a reasoning model from SpaceXAI with industry-leading speed and agentic tool calling capabilities. It combines the lowest hallucination rate on the market with strict prompt adherance, delivering...
Inkling is an open-weight multimodal mixture-of-experts model from Thinking Machines Lab, with 41B active parameters out of 975B total. It is designed for general-purpose reasoning, coding, agentic and tool-use systems,...
Inkling is an open-weight multimodal mixture-of-experts model from Thinking Machines Lab, with 41B active parameters out of 975B total. It is designed for general-purpose reasoning, coding, agentic and tool-use systems,...
Grok 4.3 is a reasoning model from SpaceXAI. It accepts text and image inputs with text output, and is suited for agentic workflows, instruction-following tasks, and applications requiring high factual...
Grok 4.3 is a reasoning model from SpaceXAI. It accepts text and image inputs with text output, and is suited for agentic workflows, instruction-following tasks, and applications requiring high factual...
Reliable GPT generation for broad coding, writing, and tool-assisted product work
GPT-5.2 is the latest frontier-grade model in the GPT-5 series, offering stronger agentic and long context perfomance compared to GPT-5.1. It uses adaptive reasoning to allocate computation dynamically, responding quickly...
DeepSeek-V3.1 Terminus is an update to [DeepSeek V3.1](/deepseek/deepseek-chat-v3.1) that maintains the model's original capabilities while addressing issues reported by users, including language consistency and agent capabilities, further optimizing the model's...
GPT-5 is OpenAI’s most advanced model, offering major improvements in reasoning, code quality, and user experience. It is optimized for complex tasks that require step-by-step reasoning, instruction following, and accuracy...
Original GPT-5 workhorse for reasoning, coding, writing, and tool workflows
The Qwen3.5 native vision-language series Plus models are built on a hybrid architecture that integrates linear attention mechanisms with sparse mixture-of-experts models, achieving higher inference efficiency. In a variety of...
Claude Sonnet 4.5 is Anthropic’s most advanced Sonnet model to date, optimized for real-world agents and coding workflows. It delivers state-of-the-art performance on coding benchmarks such as SWE-bench Verified, with...
Claude Sonnet 4.5 is Anthropic’s most advanced Sonnet model to date, optimized for real-world agents and coding workflows. It delivers state-of-the-art performance on coding benchmarks such as SWE-bench Verified, with...
MiniMax-M2.5 is a SOTA large language model designed for real-world productivity. Trained in a diverse range of complex real-world digital working environments, M2.5 builds upon the coding expertise of M2.1...
DeepSeek-V3.2-Exp is an experimental large language model released by DeepSeek as an intermediate step between V3.1 and future architectures. It introduces DeepSeek Sparse Attention (DSA), a fine-grained sparse attention mechanism...
GPT-5.1 is the latest frontier-grade model in the GPT-5 series, offering stronger general-purpose reasoning, improved instruction adherence, and a more natural conversational style compared to GPT-5. It uses adaptive reasoning...
Sharper GPT-5 generation for coding, product work, and tool-assisted tasks
Claude Opus 4.1 is an updated version of Anthropic’s flagship model, offering improved performance in coding, reasoning, and agentic tasks. It achieves 74.5% on SWE-bench Verified and shows notable gains...
Claude Opus 4.1 is an updated version of Anthropic’s flagship model, offering improved performance in coding, reasoning, and agentic tasks. It achieves 74.5% on SWE-bench Verified and shows notable gains...
The Qwen3.5 series 397B-A17B native vision-language model is built on a hybrid architecture that integrates a linear attention mechanism with a sparse mixture-of-experts model, achieving higher inference efficiency. It delivers...
Claude Opus 4 is benchmarked as the world’s best coding model, at time of release, bringing sustained performance on complex, long-running tasks and agent workflows. It sets new benchmarks in...
Upstage's flagship model, specialized for agentic use
Compared with GLM-4.5, this generation brings several key improvements: Longer context window: The context window has been expanded from 128K to 200K tokens, enabling the model to handle more complex...
NVIDIA Nemotron 3 Ultra is an open frontier-reasoning and orchestration model from NVIDIA, with 55B active parameters out of 550B total (MoE). Built on a hybrid Transformer-Mamba mixture-of-experts architecture, it...
GLM-4.5 is our latest flagship foundation model, purpose-built for agent-based applications. It leverages a Mixture-of-Experts (MoE) architecture and supports a context length of up to 128k tokens. GLM-4.5 delivers significantly...
Qwen3-Coder-480B-A35B-Instruct is a Mixture-of-Experts (MoE) code generation model developed by the Qwen team. It is optimized for agentic coding tasks such as function calling, tool use, and long-context reasoning over...
Google's proven reasoning model for coding, math, and multimodal analysis
Gemini 2.5 Pro is Google’s state-of-the-art AI model designed for advanced reasoning, coding, mathematics, and scientific tasks. It employs “thinking” capabilities, enabling it to reason through responses with enhanced accuracy...
Coding-optimized GPT model for repository edits, reviews, and agentic software work
NVIDIA Nemotron 3 Ultra is an open frontier-reasoning and orchestration model from NVIDIA, with 55B active parameters out of 550B total (MoE). Built on a hybrid Transformer-Mamba mixture-of-experts architecture, it...
MiniMax-M2 is a compact, high-efficiency large language model optimized for end-to-end coding and agentic workflows. With 10 billion activated parameters (230 billion total), it delivers near-frontier intelligence across general reasoning,...
Claude Sonnet 4 significantly enhances the capabilities of its predecessor, Sonnet 3.7, excelling in both coding and reasoning tasks with improved precision and controllability. Achieving state-of-the-art performance on SWE-bench (72.7%),...
GLM-4.5-Air is the lightweight variant of our latest flagship model family, also purpose-built for agent-centric applications. Like GLM-4.5, it adopts the Mixture-of-Experts (MoE) architecture but with a more compact parameter...
Qwen3-Coder-480B-A35B-Instruct is a Mixture-of-Experts (MoE) code generation model developed by the Qwen team. It is optimized for agentic coding tasks such as function calling, tool use, and long-context reasoning over...
| Model | Creator | Score | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|---|
| OpenAI: GPT-5.4 (batch)openai/gpt-5.4:batch | 1254.0 | 1.05M | $1.25 | $7.5 | — | |||
| GPT-5.4openai/gpt-5.4 | 1254.0 | 1.05M | $2.5 | $15 | 2026-03-05 | |||
| Anthropic: Claude Opus 4.5anthropic/claude-opus-4.5 | 1253.0 | 200K | $5 | $25 | — | |||
| Anthropic: Claude Opus 4.5 (batch)anthropic/claude-opus-4.5:batch | 1253.0 | 200K | $2.5 | $12.5 | — | |||
| DeepSeek: DeepSeek V4 Flash 0731 (batch)deepseek/deepseek-v4-flash-0731:batch | 1252.0 | 1.04858M | $0.11 | $0.33 | — | |||
| GPT-5.1 Codexopenai/gpt-5.1-codex | 1252.0 | 400K | $1.07 | $8.5 | 2025-11-13 | |||
| Qwen: Qwen3.6 Plusqwen/qwen3.6-plus | 1248.0 | 1M | $0.325 | $1.95 | — | |||
| Z.ai: GLM 5z-ai/glm-5 | 1245.0 | 198K | $0.6 | $1.92 | — | |||
| MiniMax: MiniMax M2.1minimax/minimax-m2.1 | 1238.0 | 204.8K | $0.3 | $1.2 | — | |||
| Nex AGI: Nex-N2-Pronex-agi/nex-n2-pro | 1237.0 | 262.144K | $0.25 | $1 | — | |||
| Tencent: Hy3 (free)tencent/hy3:free | 1235.0 | 262.144K | Free | Free | — | |||
| MiniMax: MiniMax M2.7 (free)minimax/minimax-m2.7:free | 1232.0 | 196.608K | Free | Free | — | |||
| Z.ai: GLM 5V Turboz-ai/glm-5v-turbo | 1227.0 | 202.752K | $1.2 | $4 | — | |||
| MiniMax: MiniMax M2.7minimax/minimax-m2.7 | 1225.0 | 204.8K | $0.3 | $1.2 | — | |||
| Z.ai: GLM 4.7 Flashz-ai/glm-4.7-flash | 1222.0 | 131.072K | $0.061 | $0.4 | — | |||
| Z.ai: GLM 4.7z-ai/glm-4.7 | 1214.0 | 202.752K | $0.4 | $1.75 | — | |||
| SpaceXAI: Grok 4.20x-ai/grok-4.20 | 1213.0 | 2M | $1.25 | $2.5 | — | |||
| Thinking Machines: Inkling (batch)thinkingmachines/inkling:batch | 1211.0 | 524.288K | $1 | $4.05 | — | |||
| Thinking Machines: Inkling (free)thinkingmachines/inkling:free | 1211.0 | 1.04858M | Free | Free | — | |||
| SpaceXAI: Grok 4.3x-ai/grok-4.3 | 1206.0 | 1M | $1.25 | $2.5 | — | |||
| SpaceXAI: Grok 4.3 (batch)x-ai/grok-4.3:batch | 1206.0 | 1M | $1 | $2 | — | |||
| GPT-5.2openai/gpt-5.2 | 1205.0 | 400K | $1.75 | $14 | 2025-12-11 | |||
| OpenAI: GPT-5.2 (batch)openai/gpt-5.2:batch | 1205.0 | 400K | $0.875 | $7 | — | |||
| DeepSeek: DeepSeek V3.1 Terminusdeepseek/deepseek-v3.1-terminus | 1197.0 | 131.072K | $0.27 | $1 | — | |||
| OpenAI: GPT-5 (batch)openai/gpt-5:batch | 1194.0 | 400K | $0.625 | $5 | — | |||
| GPT-5openai/gpt-5 | 1194.0 | 400K | $1.25 | $10 | 2025-08-07 | |||
| Qwen: Qwen3.5 Plus 2026-02-15qwen/qwen3.5-plus-02-15 | 1192.0 | 1M | $0.26 | $1.56 | — | |||
| Anthropic: Claude Sonnet 4.5anthropic/claude-sonnet-4.5 | 1187.0 | 1M | $3 | $15 | — | |||
| Anthropic: Claude Sonnet 4.5 (batch)anthropic/claude-sonnet-4.5:batch | 1187.0 | 1M | $1.5 | $7.5 | — | |||
| MiniMax: MiniMax M2.5minimax/minimax-m2.5 | 1187.0 | 200K | $0.27 | $1.08 | — | |||
| DeepSeek: DeepSeek V3.2 Expdeepseek/deepseek-v3.2-exp | 1181.0 | 163.84K | $0.27 | $0.41 | — | |||
| OpenAI: GPT-5.1 (batch)openai/gpt-5.1:batch | 1181.0 | 400K | $0.625 | $5 | — | |||
| GPT-5.1openai/gpt-5.1 | 1181.0 | 400K | $1.25 | $10 | 2025-11-13 | |||
| Anthropic: Claude Opus 4.1anthropic/claude-opus-4.1 | 1179.0 | 200K | $15 | $75 | — | |||
| Anthropic: Claude Opus 4.1 (batch)anthropic/claude-opus-4.1:batch | 1179.0 | 200K | $7.5 | $37.5 | — | |||
| Qwen: Qwen3.5 397B A17Bqwen/qwen3.5-397b-a17b | 1178.0 | 262.144K | $0.55 | $3.5 | — | |||
| Anthropic: Claude Opus 4anthropic/claude-opus-4 | 1168.0 | 200K | $15 | $75 | — | |||
| Solar Pro 4upstage/solar-pro4 | 1165.0 | 524.288K | $0.3 | $1.2 | 2026-08-06 | |||
| Z.ai: GLM 4.6z-ai/glm-4.6 | 1162.0 | 198K | $0.43 | $1.75 | — | |||
| NVIDIA: Nemotron 3 Ultra (batch)nvidia/nemotron-3-ultra-550b-a55b:batch | 1161.0 | 512.288K | $0.6 | $3.6 | — | |||
| Z.ai: GLM 4.5z-ai/glm-4.5 | 1159.0 | 131.072K | $0.6 | $2.2 | — | |||
| Qwen: Qwen3 Coder 480B A35B (free)qwen/qwen3-coder:free | 1156.0 | 262K | Free | Free | — | |||
| Gemini 2.5 Progoogle/gemini-2.5-pro | 1154.0 | 1.04858M | $1.25 | $10 | 2025-06-17 | |||
| Google: Gemini 2.5 Pro (batch)google/gemini-2.5-pro:batch | 1154.0 | 1.04858M | $0.625 | $5 | — | |||
| GPT-5.3 Codexopenai/gpt-5.3-codex | 1152.0 | 400K | $1.75 | $14 | 2026-02-05 | |||
| NVIDIA: Nemotron 3 Ultra (free)nvidia/nemotron-3-ultra-550b-a55b:free | 1151.0 | 1M | Free | Free | — | |||
| MiniMax: MiniMax M2minimax/minimax-m2 | 1151.0 | 204.8K | $0.255 | $1.02 | — | |||
| Anthropic: Claude Sonnet 4anthropic/claude-sonnet-4 | 1143.0 | 200K | $3 | $15 | — | |||
| Z.ai: GLM 4.5 Airz-ai/glm-4.5-air | 1138.0 | 131.072K | $0.13 | $0.85 | — | |||
| Qwen: Qwen3 Coder 480B A35Bqwen/qwen3-coder | 1127.0 | 262.144K | $0.3 | $1 | — |