Codex GPT for repository edits, code review, and practical software agents
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
Earlier Kimi frontier model for long-context agents, coding, and multimodal work
Qwen 3.6 Plus builds on a hybrid architecture that combines efficient linear attention with sparse mixture-of-experts routing, enabling strong scalability and high-performance inference. Compared to the 3.5 series, it delivers...
Nex-N2-Pro is an agentic mixture-of-experts model from Nex AGI, with 17B active parameters out of 397B total. Built on the Qwen3.5 architecture, it accepts text and image input and produces...
GLM-5V-Turbo is Z.ai’s first native multimodal agent foundation model, built for vision-based coding and agent-driven tasks. It natively handles image, video, and text inputs, excels at long-horizon planning, complex coding,...
Grok 4.20 is a reasoning model from SpaceXAI with industry-leading speed and agentic tool calling capabilities. It combines the lowest hallucination rate on the market with strict prompt adherance, delivering...
Inkling is an open-weight multimodal mixture-of-experts model from Thinking Machines Lab, with 41B active parameters out of 975B total. It is designed for general-purpose reasoning, coding, agentic and tool-use systems,...
Multimodal MoE reasoning model (975B total, 41B active) for text, image, and audio
Inkling is an open-weight multimodal mixture-of-experts model from Thinking Machines Lab, with 41B active parameters out of 975B total. It is designed for general-purpose reasoning, coding, agentic and tool-use systems,...
Grok 4.3 is a reasoning model from SpaceXAI. It accepts text and image inputs with text output, and is suited for agentic workflows, instruction-following tasks, and applications requiring high factual...
Grok 4.3 is a reasoning model from SpaceXAI. It accepts text and image inputs with text output, and is suited for agentic workflows, instruction-following tasks, and applications requiring high factual...
Reliable GPT generation for broad coding, writing, and tool-assisted product work
GPT-5.2 is the latest frontier-grade model in the GPT-5 series, offering stronger agentic and long context perfomance compared to GPT-5.1. It uses adaptive reasoning to allocate computation dynamically, responding quickly...
Newer StepFun flash model for faster agents, coding, and multimodal prompts
GPT-5 is OpenAI’s most advanced model, offering major improvements in reasoning, code quality, and user experience. It is optimized for complex tasks that require step-by-step reasoning, instruction following, and accuracy...
Original GPT-5 workhorse for reasoning, coding, writing, and tool workflows
The Qwen3.5 native vision-language series Plus models are built on a hybrid architecture that integrates linear attention mechanisms with sparse mixture-of-experts models, achieving higher inference efficiency. In a variety of...
Claude Sonnet 4.5 is Anthropic’s most advanced Sonnet model to date, optimized for real-world agents and coding workflows. It delivers state-of-the-art performance on coding benchmarks such as SWE-bench Verified, with...
Claude Sonnet 4.5 is Anthropic’s most advanced Sonnet model to date, optimized for real-world agents and coding workflows. It delivers state-of-the-art performance on coding benchmarks such as SWE-bench Verified, with...
Sharper GPT-5 generation for coding, product work, and tool-assisted tasks
GPT-5.1 is the latest frontier-grade model in the GPT-5 series, offering stronger general-purpose reasoning, improved instruction adherence, and a more natural conversational style compared to GPT-5. It uses adaptive reasoning...
Claude Opus 4.1 is an updated version of Anthropic’s flagship model, offering improved performance in coding, reasoning, and agentic tasks. It achieves 74.5% on SWE-bench Verified and shows notable gains...
Claude Opus 4.1 is an updated version of Anthropic’s flagship model, offering improved performance in coding, reasoning, and agentic tasks. It achieves 74.5% on SWE-bench Verified and shows notable gains...
The Qwen3.5 series 397B-A17B native vision-language model is built on a hybrid architecture that integrates a linear attention mechanism with a sparse mixture-of-experts model, achieving higher inference efficiency. It delivers...
Claude Opus 4 is benchmarked as the world’s best coding model, at time of release, bringing sustained performance on complex, long-running tasks and agent workflows. It sets new benchmarks in...
Gemini 2.5 Pro is Google’s state-of-the-art AI model designed for advanced reasoning, coding, mathematics, and scientific tasks. It employs “thinking” capabilities, enabling it to reason through responses with enhanced accuracy...
Google's proven reasoning model for coding, math, and multimodal analysis
Coding-optimized GPT model for repository edits, reviews, and agentic software work
Claude Sonnet 4 significantly enhances the capabilities of its predecessor, Sonnet 3.7, excelling in both coding and reasoning tasks with improved precision and controllability. Achieving state-of-the-art performance on SWE-bench (72.7%),...
GPT-5 Mini is a compact version of GPT-5, designed to handle lighter-weight reasoning tasks. It provides the same instruction-following and safety-tuning benefits as GPT-5, but with reduced latency and cost....
Small GPT-5 for responsive agents, coding help, and everyday automation
Mistral Large 3 2512 is Mistral’s most capable model to date, featuring a sparse mixture-of-experts architecture with 41B active parameters (675B total), and released under the Apache 2.0 license.
Mistral Large 3 2512 is Mistral’s most capable model to date, featuring a sparse mixture-of-experts architecture with 41B active parameters (675B total), and released under the Apache 2.0 license.
Claude Haiku 4.5 is Anthropic’s fastest and most efficient model, delivering near-frontier intelligence at a fraction of the cost and latency of larger Claude models. Matching Claude Sonnet 4’s performance...
Claude Haiku 4.5 is Anthropic’s fastest and most efficient model, delivering near-frontier intelligence at a fraction of the cost and latency of larger Claude models. Matching Claude Sonnet 4’s performance...
Mistral Medium 3.1 is an updated version of Mistral Medium 3, which is a high-performance enterprise-grade language model designed to deliver frontier-level capabilities at significantly reduced operational cost. It balances...
Mistral Medium 3.1 is an updated version of Mistral Medium 3, which is a high-performance enterprise-grade language model designed to deliver frontier-level capabilities at significantly reduced operational cost. It balances...
Fast Gemini workhorse for multimodal apps where latency and price matter
Gemini 2.5 Flash is Google's state-of-the-art workhorse model, specifically designed for advanced reasoning, coding, mathematics, and scientific tasks. It includes built-in "thinking" capabilities, enabling it to provide responses with greater...
Coding-optimized GPT model for repository edits, reviews, and agentic software work
Low-latency Gemini model for high-volume multimodal and agent workloads
GPT-5-Nano is the smallest and fastest variant in the GPT-5 system, optimized for developer tools, rapid interactions, and ultra-low latency environments. While limited in reasoning depth compared to its larger...
Tiny GPT-5 lane for routing, extraction, classification, and bulk jobs
Mistral Medium 3 is a high-performance enterprise-grade language model designed to deliver frontier-level capabilities at significantly reduced operational cost. It balances state-of-the-art reasoning and multimodal performance with 8× lower cost...
Deliberate o-series reasoner for hard math, coding, and multi-step analysis
o3 is a well-rounded and powerful model across domains. It sets a new standard for math, science, coding, and visual reasoning tasks. It also excels at technical writing and instruction-following....
Long-lived GPT workhorse for coding, instruction following, and production apps
GPT-4.1 is a flagship large language model optimized for advanced instruction following, real-world software engineering, and long-context reasoning. It supports a 1 million token context window and outperforms GPT-4o and...
OpenAI o4-mini is a compact reasoning model in the o-series, optimized for fast, cost-efficient performance while retaining strong multimodal and agentic capabilities. It supports tool use and demonstrates competitive reasoning...
Fast o-series model for compact reasoning, coding, and tool use
| Model | Creator | Score | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|---|
| GPT-5.1 Codexopenai/gpt-5.1-codex | 1252.0 | 400K | $1.07 | $8.5 | 2025-11-13 | |||
| Kimi K2.5moonshotai/kimi-k2.5 | 1251.0 | 262.144K | $0.3 | $1.9 | 2026-01 | |||
| Qwen: Qwen3.6 Plusqwen/qwen3.6-plus | 1249.0 | 1M | $0.325 | $1.95 | — | |||
| Nex AGI: Nex-N2-Pronex-agi/nex-n2-pro | 1237.0 | 262.144K | $0.25 | $1 | — | |||
| Z.ai: GLM 5V Turboz-ai/glm-5v-turbo | 1227.0 | 202.752K | $1.2 | $4 | — | |||
| SpaceXAI: Grok 4.20x-ai/grok-4.20 | 1213.0 | 2M | $1.25 | $2.5 | — | |||
| Thinking Machines: Inkling (batch)thinkingmachines/inkling:batch | 1209.0 | 524.288K | $1 | $4.05 | — | |||
| Inklingthinkingmachines/inkling | 1209.0 | 1.04858M | $1.87 | $4.68 | 2026-07-15 | |||
| Thinking Machines: Inkling (free)thinkingmachines/inkling:free | 1209.0 | 1.04858M | Free | Free | — | |||
| SpaceXAI: Grok 4.3x-ai/grok-4.3 | 1206.0 | 1M | $1.25 | $2.5 | — | |||
| SpaceXAI: Grok 4.3 (batch)x-ai/grok-4.3:batch | 1206.0 | 1M | $1 | $2 | — | |||
| GPT-5.2openai/gpt-5.2 | 1205.0 | 400K | $1.75 | $14 | 2025-12-11 | |||
| OpenAI: GPT-5.2 (batch)openai/gpt-5.2:batch | 1205.0 | 400K | $0.875 | $7 | — | |||
| Step 3.7 Flashstepfun/step-3.7-flash | 1196.0 | 256K | $0.185 | $1.11 | 2026-05-29 | |||
| OpenAI: GPT-5 (batch)openai/gpt-5:batch | 1194.0 | 400K | $0.625 | $5 | — | |||
| GPT-5openai/gpt-5 | 1194.0 | 400K | $1.25 | $10 | 2025-08-07 | |||
| Qwen: Qwen3.5 Plus 2026-02-15qwen/qwen3.5-plus-02-15 | 1192.0 | 1M | $0.26 | $1.56 | — | |||
| Anthropic: Claude Sonnet 4.5 (batch)anthropic/claude-sonnet-4.5:batch | 1187.0 | 1M | $1.5 | $7.5 | — | |||
| Anthropic: Claude Sonnet 4.5anthropic/claude-sonnet-4.5 | 1187.0 | 1M | $3 | $15 | — | |||
| GPT-5.1openai/gpt-5.1 | 1181.0 | 400K | $1.25 | $10 | 2025-11-13 | |||
| OpenAI: GPT-5.1 (batch)openai/gpt-5.1:batch | 1181.0 | 400K | $0.625 | $5 | — | |||
| Anthropic: Claude Opus 4.1 (batch)anthropic/claude-opus-4.1:batch | 1179.0 | 200K | $7.5 | $37.5 | — | |||
| Anthropic: Claude Opus 4.1anthropic/claude-opus-4.1 | 1179.0 | 200K | $15 | $75 | — | |||
| Qwen: Qwen3.5 397B A17Bqwen/qwen3.5-397b-a17b | 1178.0 | 262.144K | $0.55 | $3.5 | — | |||
| Anthropic: Claude Opus 4anthropic/claude-opus-4 | 1168.0 | 200K | $15 | $75 | — | |||
| Google: Gemini 2.5 Pro (batch)google/gemini-2.5-pro:batch | 1154.0 | 1.04858M | $0.625 | $5 | — | |||
| Gemini 2.5 Progoogle/gemini-2.5-pro | 1154.0 | 1.04858M | $1.25 | $10 | 2025-06-17 | |||
| GPT-5.3 Codexopenai/gpt-5.3-codex | 1152.0 | 400K | $1.75 | $14 | 2026-02-05 | |||
| Anthropic: Claude Sonnet 4anthropic/claude-sonnet-4 | 1143.0 | 200K | $3 | $15 | — | |||
| OpenAI: GPT-5 Mini (batch)openai/gpt-5-mini:batch | 1121.0 | 400K | $0.125 | $1 | — | |||
| GPT-5 Miniopenai/gpt-5-mini | 1121.0 | 400K | $0.25 | $2 | 2025-08-07 | |||
| Mistral: Mistral Large 3 2512 (batch)mistralai/mistral-large-2512:batch | 1117.0 | 262.144K | $0.25 | $0.75 | — | |||
| Mistral: Mistral Large 3 2512mistralai/mistral-large-2512 | 1117.0 | 262.144K | $0.5 | $1.5 | — | |||
| Anthropic: Claude Haiku 4.5anthropic/claude-haiku-4.5 | 1115.0 | 200K | $1 | $5 | — | |||
| Anthropic: Claude Haiku 4.5 (batch)anthropic/claude-haiku-4.5:batch | 1115.0 | 200K | $0.5 | $2.5 | — | |||
| Mistral: Mistral Medium 3.1mistralai/mistral-medium-3.1 | 1114.0 | 131.072K | $0.4 | $2 | — | |||
| Mistral: Mistral Medium 3.1 (batch)mistralai/mistral-medium-3.1:batch | 1114.0 | 131.072K | $0.2 | $1 | — | |||
| Gemini 2.5 Flashgoogle/gemini-2.5-flash | 1107.0 | 1.04858M | $0.3 | $2.5 | 2025-06-17 | |||
| Google: Gemini 2.5 Flash (batch)google/gemini-2.5-flash:batch | 1107.0 | 1.04858M | $0.15 | $1.25 | — | |||
| GPT-5.1 Codex miniopenai/gpt-5.1-codex-mini | 1096.0 | 400K | $0.22 | $1.8 | 2025-11-13 | |||
| Gemini 3.1 Flash Lite Previewgoogle/gemini-3.1-flash-lite-preview | 1086.0 | 1.04858M | $0.25 | $1.5 | 2026-03-03 | |||
| OpenAI: GPT-5 Nano (batch)openai/gpt-5-nano:batch | 1081.0 | 400K | $0.025 | $0.2 | — | |||
| GPT-5 Nanoopenai/gpt-5-nano | 1081.0 | 400K | $0.05 | $0.4 | 2025-08-07 | |||
| Mistral: Mistral Medium 3mistralai/mistral-medium-3 | 1039.0 | 131.072K | $0.4 | $2 | — | |||
| o3openai/o3 | 1031.0 | 200K | $2 | $8 | 2025-04-16 | |||
| OpenAI: o3 (batch)openai/o3:batch | 1031.0 | 200K | $1 | $4 | — | |||
| GPT-4.1openai/gpt-4.1 | 1015.0 | 1.04758M | $2 | $8 | 2025-04-14 | |||
| OpenAI: GPT-4.1 (batch)openai/gpt-4.1:batch | 1015.0 | 1.04758M | $1 | $4 | — | |||
| OpenAI: o4 Mini (batch)openai/o4-mini:batch | 998.0 | 200K | $0.55 | $2.2 | — | |||
| o4-miniopenai/o4-mini | 998.0 | 200K | $1.1 | $4.4 | 2025-04-16 |