GPT-5.5 is OpenAI’s frontier model designed for complex professional workloads, building on GPT-5.4 with stronger reasoning, higher reliability, and improved token efficiency on hard tasks. It features a 1M+ token...
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
Claude Opus 4.8 is Anthropic's most capable generally available model in the Opus family. It supports text, image, and file inputs with text output, with reasoning support and a 1M-token...
Claude Opus 4.8 is Anthropic's most capable generally available model in the Opus family. It supports text, image, and file inputs with text output, with reasoning support and a 1M-token...
Reasoning-first Gemini preview for agentic coding and complex problem solving
Gemini 3.1 Pro Preview is Google’s frontier reasoning model, delivering enhanced software engineering performance, improved agentic reliability, and more efficient token usage across complex workflows. Building on the multimodal foundation...
Earlier Kimi frontier model for long-context agents, coding, and multimodal work
Claude Opus 4.5 is Anthropic’s frontier reasoning model optimized for complex software engineering, agentic workflows, and long-horizon computer use. It offers strong multimodal capabilities, competitive performance across real-world coding and...
Claude Opus 4.5 is Anthropic’s frontier reasoning model optimized for complex software engineering, agentic workflows, and long-horizon computer use. It offers strong multimodal capabilities, competitive performance across real-world coding and...
MiniMax-M2.7 is a next-generation large language model designed for autonomous, real-world productivity and continuous improvement. Built to actively participate in its own evolution, M2.7 integrates advanced agentic capabilities through multi-agent...
Qwen 3.6 Plus builds on a hybrid architecture that combines efficient linear attention with sparse mixture-of-experts routing, enabling strong scalability and high-performance inference. Compared to the 3.5 series, it delivers...
DeepSeek V4 Flash 0731 is a sparse mixture-of-experts model from DeepSeek, with 13B active parameters out of 284B total. This re-post-trained revision is suited for coding, reasoning, and agent workflows....
Official DeepSeek V4 Flash release with enhanced agentic capabilities and integrated DSpark speculative decoding
Open MoE flagship with million-token context for coding and long agent runs
GLM-5V-Turbo is Z.ai’s first native multimodal agent foundation model, built for vision-based coding and agent-driven tasks. It natively handles image, video, and text inputs, excels at long-horizon planning, complex coding,...
Grok 4.20 is a reasoning model from SpaceXAI with industry-leading speed and agentic tool calling capabilities. It combines the lowest hallucination rate on the market with strict prompt adherance, delivering...
GLM-4.7 is Z.ai’s latest flagship model, featuring upgrades in two key areas: enhanced programming capabilities and more stable multi-step reasoning/execution. It demonstrates significant improvements in executing complex agent tasks while...
Nex-N2-Pro is an agentic mixture-of-experts model from Nex AGI, with 17B active parameters out of 397B total. Built on the Qwen3.5 architecture, it accepts text and image input and produces...
MiniMax-M2.5 is a SOTA large language model designed for real-world productivity. Trained in a diverse range of complex real-world digital working environments, M2.5 builds upon the coding expertise of M2.1...
Agent-ready GPT for coding and computer-use workflows at a lower cost
GPT-5.4 is OpenAI’s latest frontier model, unifying the Codex and GPT lines into a single system. It features a 1M+ token context window (922K input, 128K output) with support for...
Hy3 is a 295B-parameter Mixture-of-Experts model from Tencent (21B active, 192 experts with top-8 routing) built for reasoning, agentic workflows, and real-world production use. It supports a configurable reasoning effort:...
Multimodal MoE reasoning model (975B total, 41B active) for text, image, and audio
Inkling is an open-weight multimodal mixture-of-experts model from Thinking Machines Lab, with 41B active parameters out of 975B total. It is designed for general-purpose reasoning, coding, agentic and tool-use systems,...
Inkling is an open-weight multimodal mixture-of-experts model from Thinking Machines Lab, with 41B active parameters out of 975B total. It is designed for general-purpose reasoning, coding, agentic and tool-use systems,...
Fast DeepSeek V4 lane for economical reasoning, coding, and long-context work
MiniMax-M2.1 is a lightweight, state-of-the-art large language model optimized for coding, agentic workflows, and modern application development. With only 10 billion activated parameters, it delivers a major jump in real-world...
Gemini 3 Flash Preview is a high speed, high value thinking model designed for agentic workflows, multi turn chat, and coding assistance. It delivers near Pro level reasoning and tool...
New Gemini flash lane bringing frontier-style multimodal reasoning to cheaper runs
GPT-5.2 is the latest frontier-grade model in the GPT-5 series, offering stronger agentic and long context perfomance compared to GPT-5.1. It uses adaptive reasoning to allocate computation dynamically, responding quickly...
Reliable GPT generation for broad coding, writing, and tool-assisted product work
Grok 4.3 is a reasoning model from SpaceXAI. It accepts text and image inputs with text output, and is suited for agentic workflows, instruction-following tasks, and applications requiring high factual...
Newer StepFun flash model for faster agents, coding, and multimodal prompts
Grok 4.3 is a reasoning model from SpaceXAI. It accepts text and image inputs with text output, and is suited for agentic workflows, instruction-following tasks, and applications requiring high factual...
The Qwen3.5 series 397B-A17B native vision-language model is built on a hybrid architecture that integrates a linear attention mechanism with a sparse mixture-of-experts model, achieving higher inference efficiency. It delivers...
Claude Sonnet 4.5 is Anthropic’s most advanced Sonnet model to date, optimized for real-world agents and coding workflows. It delivers state-of-the-art performance on coding benchmarks such as SWE-bench Verified, with...
Claude Sonnet 4.5 is Anthropic’s most advanced Sonnet model to date, optimized for real-world agents and coding workflows. It delivers state-of-the-art performance on coding benchmarks such as SWE-bench Verified, with...
The Qwen3.5 native vision-language series Plus models are built on a hybrid architecture that integrates linear attention mechanisms with sparse mixture-of-experts models, achieving higher inference efficiency. In a variety of...
GPT-5.1 is the latest frontier-grade model in the GPT-5 series, offering stronger general-purpose reasoning, improved instruction adherence, and a more natural conversational style compared to GPT-5. It uses adaptive reasoning...
Sharper GPT-5 generation for coding, product work, and tool-assisted tasks
Original GPT-5 workhorse for reasoning, coding, writing, and tool workflows
GPT-5 is OpenAI’s most advanced model, offering major improvements in reasoning, code quality, and user experience. It is optimized for complex tasks that require step-by-step reasoning, instruction following, and accuracy...
Tencent Hy reasoning model for coding, instruction following, and agent tasks
Qwen3-Coder-480B-A35B-Instruct is a Mixture-of-Experts (MoE) code generation model developed by the Qwen team. It is optimized for agentic coding tasks such as function calling, tool use, and long-context reasoning over...
Claude Opus 4.1 is an updated version of Anthropic’s flagship model, offering improved performance in coding, reasoning, and agentic tasks. It achieves 74.5% on SWE-bench Verified and shows notable gains...
Claude Opus 4.1 is an updated version of Anthropic’s flagship model, offering improved performance in coding, reasoning, and agentic tasks. It achieves 74.5% on SWE-bench Verified and shows notable gains...
Upstage's flagship model, specialized for agentic use
Google's proven reasoning model for coding, math, and multimodal analysis
Gemini 2.5 Pro is Google’s state-of-the-art AI model designed for advanced reasoning, coding, mathematics, and scientific tasks. It employs “thinking” capabilities, enabling it to reason through responses with enhanced accuracy...
Claude Opus 4 is benchmarked as the world’s best coding model, at time of release, bringing sustained performance on complex, long-running tasks and agent workflows. It sets new benchmarks in...
Mistral Large 3 2512 is Mistral’s most capable model to date, featuring a sparse mixture-of-experts architecture with 41B active parameters (675B total), and released under the Apache 2.0 license.
| Model | Creator | Score | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|---|
| OpenAI: GPT-5.5 (batch)openai/gpt-5.5:batch | 1269.0 | 1.05M | $2.5 | $15 | — | |||
| Anthropic: Claude Opus 4.8anthropic/claude-opus-4.8 | 1267.0 | 1M | $5 | $25 | — | |||
| Anthropic: Claude Opus 4.8 (batch)anthropic/claude-opus-4.8:batch | 1267.0 | 1M | $2.5 | $12.5 | — | |||
| Gemini 3.1 Pro Previewgoogle/gemini-3.1-pro-preview | 1265.0 | 1.04858M | $2 | $12 | 2026-02-19 | |||
| Google: Gemini 3.1 Pro Preview (batch)google/gemini-3.1-pro-preview:batch | 1265.0 | 1.04858M | $1 | $6 | — | |||
| Kimi K2.5moonshotai/kimi-k2.5 | 1261.0 | 262.144K | $0.3 | $1.9 | 2026-01 | |||
| Anthropic: Claude Opus 4.5 (batch)anthropic/claude-opus-4.5:batch | 1259.0 | 200K | $2.5 | $12.5 | — | |||
| Anthropic: Claude Opus 4.5anthropic/claude-opus-4.5 | 1259.0 | 200K | $5 | $25 | — | |||
| MiniMax: MiniMax M2.7minimax/minimax-m2.7 | 1258.0 | 204.8K | $0.3 | $1.2 | — | |||
| Qwen: Qwen3.6 Plusqwen/qwen3.6-plus | 1253.0 | 1M | $0.325 | $1.95 | — | |||
| DeepSeek: DeepSeek V4 Flash 0731 (batch)deepseek/deepseek-v4-flash-0731:batch | 1251.0 | 1.04858M | $0.11 | $0.33 | — | |||
| DeepSeek V4 Flash 0731deepseek/deepseek-v4-flash-0731 | 1251.0 | 1M | $0.05 | $0.16 | 2026-07-31 | |||
| DeepSeek V4 Prodeepseek/deepseek-v4-pro | 1247.0 | 1M | $0.435 | $0.87 | 2026-04-24 | |||
| Z.ai: GLM 5V Turboz-ai/glm-5v-turbo | 1243.0 | 202.752K | $1.2 | $4 | — | |||
| SpaceXAI: Grok 4.20x-ai/grok-4.20 | 1242.0 | 2M | $1.25 | $2.5 | — | |||
| Z.ai: GLM 4.7z-ai/glm-4.7 | 1238.0 | 202.752K | $0.4 | $1.75 | — | |||
| Nex AGI: Nex-N2-Pronex-agi/nex-n2-pro | 1236.0 | 262.144K | $0.25 | $1 | — | |||
| MiniMax: MiniMax M2.5minimax/minimax-m2.5 | 1235.0 | 200K | $0.27 | $1.08 | — | |||
| GPT-5.4openai/gpt-5.4 | 1232.0 | 1.05M | $2.5 | $15 | 2026-03-05 | |||
| OpenAI: GPT-5.4 (batch)openai/gpt-5.4:batch | 1232.0 | 1.05M | $1.25 | $7.5 | — | |||
| Tencent: Hy3 (free)tencent/hy3:free | 1230.0 | 262.144K | Free | Free | — | |||
| Inklingthinkingmachines/inkling | 1226.0 | 1.04858M | $1.87 | $4.68 | 2026-07-15 | |||
| Thinking Machines: Inkling (batch)thinkingmachines/inkling:batch | 1226.0 | 524.288K | $1 | $4.05 | — | |||
| Thinking Machines: Inkling (free)thinkingmachines/inkling:free | 1226.0 | 1.04858M | Free | Free | — | |||
| DeepSeek V4 Flashdeepseek/deepseek-v4-flash | 1220.0 | 1M | $0.15 | $0.6 | 2026-04-24 | |||
| MiniMax: MiniMax M2.1minimax/minimax-m2.1 | 1214.0 | 204.8K | $0.3 | $1.2 | — | |||
| Google: Gemini 3 Flash Preview (batch)google/gemini-3-flash-preview:batch | 1208.0 | 1.04858M | $0.25 | $1.5 | — | |||
| Gemini 3 Flash Previewgoogle/gemini-3-flash-preview | 1208.0 | 1.04858M | $0.5 | $3 | 2025-12-17 | |||
| OpenAI: GPT-5.2 (batch)openai/gpt-5.2:batch | 1207.0 | 400K | $0.875 | $7 | — | |||
| GPT-5.2openai/gpt-5.2 | 1207.0 | 400K | $1.75 | $14 | 2025-12-11 | |||
| SpaceXAI: Grok 4.3x-ai/grok-4.3 | 1206.0 | 1M | $1.25 | $2.5 | — | |||
| Step 3.7 Flashstepfun/step-3.7-flash | 1206.0 | 256K | $0.185 | $1.11 | 2026-05-29 | |||
| SpaceXAI: Grok 4.3 (batch)x-ai/grok-4.3:batch | 1206.0 | 1M | $1 | $2 | — | |||
| Qwen: Qwen3.5 397B A17Bqwen/qwen3.5-397b-a17b | 1203.0 | 262.144K | $0.55 | $3.5 | — | |||
| Anthropic: Claude Sonnet 4.5anthropic/claude-sonnet-4.5 | 1202.0 | 1M | $3 | $15 | — | |||
| Anthropic: Claude Sonnet 4.5 (batch)anthropic/claude-sonnet-4.5:batch | 1202.0 | 1M | $1.5 | $7.5 | — | |||
| Qwen: Qwen3.5 Plus 2026-02-15qwen/qwen3.5-plus-02-15 | 1200.0 | 1M | $0.26 | $1.56 | — | |||
| OpenAI: GPT-5.1 (batch)openai/gpt-5.1:batch | 1199.0 | 400K | $0.625 | $5 | — | |||
| GPT-5.1openai/gpt-5.1 | 1199.0 | 400K | $1.25 | $10 | 2025-11-13 | |||
| GPT-5openai/gpt-5 | 1197.0 | 400K | $1.25 | $10 | 2025-08-07 | |||
| OpenAI: GPT-5 (batch)openai/gpt-5:batch | 1197.0 | 400K | $0.625 | $5 | — | |||
| Hy3tencent/hy3 | 1194.0 | 256K | $0.066 | $0.26 | 2026-07-06 | |||
| Qwen: Qwen3 Coder 480B A35B (free)qwen/qwen3-coder:free | 1189.0 | 262K | Free | Free | — | |||
| Anthropic: Claude Opus 4.1 (batch)anthropic/claude-opus-4.1:batch | 1189.0 | 200K | $7.5 | $37.5 | — | |||
| Anthropic: Claude Opus 4.1anthropic/claude-opus-4.1 | 1189.0 | 200K | $15 | $75 | — | |||
| Solar Pro 4upstage/solar-pro4 | 1188.0 | 524.288K | $0.3 | $1.2 | 2026-08-06 | |||
| Gemini 2.5 Progoogle/gemini-2.5-pro | 1179.0 | 1.04858M | $1.25 | $10 | 2025-06-17 | |||
| Google: Gemini 2.5 Pro (batch)google/gemini-2.5-pro:batch | 1179.0 | 1.04858M | $0.625 | $5 | — | |||
| Anthropic: Claude Opus 4anthropic/claude-opus-4 | 1177.0 | 200K | $15 | $75 | — | |||
| Mistral: Mistral Large 3 2512mistralai/mistral-large-2512 | 1175.0 | 262.144K | $0.5 | $1.5 | — |