Default frontier GPT for coding, computer use, research, and knowledge work
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
GPT-5.5 is OpenAI’s frontier model designed for complex professional workloads, building on GPT-5.4 with stronger reasoning, higher reliability, and improved token efficiency on hard tasks. It features a 1M+ token...
Grok 4.20 is a reasoning model from SpaceXAI with industry-leading speed and agentic tool calling capabilities. It combines the lowest hallucination rate on the market with strict prompt adherance, delivering...
Gemini 3 Flash Preview is a high speed, high value thinking model designed for agentic workflows, multi turn chat, and coding assistance. It delivers near Pro level reasoning and tool...
New Gemini flash lane bringing frontier-style multimodal reasoning to cheaper runs
The Qwen3.5 series 397B-A17B native vision-language model is built on a hybrid architecture that integrates a linear attention mechanism with a sparse mixture-of-experts model, achieving higher inference efficiency. It delivers...
Inkling is an open-weight multimodal mixture-of-experts model from Thinking Machines Lab, with 41B active parameters out of 975B total. It is designed for general-purpose reasoning, coding, agentic and tool-use systems,...
Claude Sonnet 4.5 is Anthropic’s most advanced Sonnet model to date, optimized for real-world agents and coding workflows. It delivers state-of-the-art performance on coding benchmarks such as SWE-bench Verified, with...
Inkling is an open-weight multimodal mixture-of-experts model from Thinking Machines Lab, with 41B active parameters out of 975B total. It is designed for general-purpose reasoning, coding, agentic and tool-use systems,...
Multimodal MoE reasoning model (975B total, 41B active) for text, image, and audio
Claude Sonnet 4.5 is Anthropic’s most advanced Sonnet model to date, optimized for real-world agents and coding workflows. It delivers state-of-the-art performance on coding benchmarks such as SWE-bench Verified, with...
Claude Opus 4.1 is an updated version of Anthropic’s flagship model, offering improved performance in coding, reasoning, and agentic tasks. It achieves 74.5% on SWE-bench Verified and shows notable gains...
Claude Opus 4.1 is an updated version of Anthropic’s flagship model, offering improved performance in coding, reasoning, and agentic tasks. It achieves 74.5% on SWE-bench Verified and shows notable gains...
Claude Sonnet 4 significantly enhances the capabilities of its predecessor, Sonnet 3.7, excelling in both coding and reasoning tasks with improved precision and controllability. Achieving state-of-the-art performance on SWE-bench (72.7%),...
Claude Opus 4 is benchmarked as the world’s best coding model, at time of release, bringing sustained performance on complex, long-running tasks and agent workflows. It sets new benchmarks in...
Newer StepFun flash model for faster agents, coding, and multimodal prompts
Grok 4.3 is a reasoning model from SpaceXAI. It accepts text and image inputs with text output, and is suited for agentic workflows, instruction-following tasks, and applications requiring high factual...
Grok 4.3 is a reasoning model from SpaceXAI. It accepts text and image inputs with text output, and is suited for agentic workflows, instruction-following tasks, and applications requiring high factual...
The Qwen3.5 native vision-language series Plus models are built on a hybrid architecture that integrates linear attention mechanisms with sparse mixture-of-experts models, achieving higher inference efficiency. In a variety of...
GPT-5.4 is OpenAI’s latest frontier model, unifying the Codex and GPT lines into a single system. It features a 1M+ token context window (922K input, 128K output) with support for...
Agent-ready GPT for coding and computer-use workflows at a lower cost
Mistral Large 3 2512 is Mistral’s most capable model to date, featuring a sparse mixture-of-experts architecture with 41B active parameters (675B total), and released under the Apache 2.0 license.
Mistral Large 3 2512 is Mistral’s most capable model to date, featuring a sparse mixture-of-experts architecture with 41B active parameters (675B total), and released under the Apache 2.0 license.
Mistral Medium 3 is a high-performance enterprise-grade language model designed to deliver frontier-level capabilities at significantly reduced operational cost. It balances state-of-the-art reasoning and multimodal performance with 8× lower cost...
Mistral Medium 3.1 is an updated version of Mistral Medium 3, which is a high-performance enterprise-grade language model designed to deliver frontier-level capabilities at significantly reduced operational cost. It balances...
Mistral Medium 3.1 is an updated version of Mistral Medium 3, which is a high-performance enterprise-grade language model designed to deliver frontier-level capabilities at significantly reduced operational cost. It balances...
Google's proven reasoning model for coding, math, and multimodal analysis
Gemini 2.5 Pro is Google’s state-of-the-art AI model designed for advanced reasoning, coding, mathematics, and scientific tasks. It employs “thinking” capabilities, enabling it to reason through responses with enhanced accuracy...
GPT-5.2 is the latest frontier-grade model in the GPT-5 series, offering stronger agentic and long context perfomance compared to GPT-5.1. It uses adaptive reasoning to allocate computation dynamically, responding quickly...
Reliable GPT generation for broad coding, writing, and tool-assisted product work
Claude Haiku 4.5 is Anthropic’s fastest and most efficient model, delivering near-frontier intelligence at a fraction of the cost and latency of larger Claude models. Matching Claude Sonnet 4’s performance...
Claude Haiku 4.5 is Anthropic’s fastest and most efficient model, delivering near-frontier intelligence at a fraction of the cost and latency of larger Claude models. Matching Claude Sonnet 4’s performance...
Gemini 2.5 Flash is Google's state-of-the-art workhorse model, specifically designed for advanced reasoning, coding, mathematics, and scientific tasks. It includes built-in "thinking" capabilities, enabling it to provide responses with greater...
Fast Gemini workhorse for multimodal apps where latency and price matter
Sharper GPT-5 generation for coding, product work, and tool-assisted tasks
GPT-5.1 is the latest frontier-grade model in the GPT-5 series, offering stronger general-purpose reasoning, improved instruction adherence, and a more natural conversational style compared to GPT-5. It uses adaptive reasoning...
Original GPT-5 workhorse for reasoning, coding, writing, and tool workflows
GPT-5 is OpenAI’s most advanced model, offering major improvements in reasoning, code quality, and user experience. It is optimized for complex tasks that require step-by-step reasoning, instruction following, and accuracy...
Low-latency Gemini model for high-volume multimodal and agent workloads
Small GPT-5 for responsive agents, coding help, and everyday automation
GPT-5 Mini is a compact version of GPT-5, designed to handle lighter-weight reasoning tasks. It provides the same instruction-following and safety-tuning benefits as GPT-5, but with reduced latency and cost....
A balanced model in the Ministral 3 family, Ministral 3 8B is a powerful, efficient tiny language model with vision capabilities.
A balanced model in the Ministral 3 family, Ministral 3 8B is a powerful, efficient tiny language model with vision capabilities.
Coding-optimized GPT model for repository edits, reviews, and agentic software work
The largest model in the Ministral 3 family, Ministral 3 14B offers frontier capabilities and performance comparable to its larger Mistral Small 3.2 24B counterpart. A powerful and efficient language...
Coding-optimized GPT model for repository edits, reviews, and agentic software work
GPT-5-Nano is the smallest and fastest variant in the GPT-5 system, optimized for developer tools, rapid interactions, and ultra-low latency environments. While limited in reasoning depth compared to its larger...
Tiny GPT-5 lane for routing, extraction, classification, and bulk jobs
The smallest model in the Ministral 3 family, Ministral 3 3B is a powerful, efficient tiny language model with vision capabilities.
For tasks that demand low latency, GPT‑4.1 nano is the fastest and cheapest model in the GPT-4.1 series. It delivers exceptional performance at a small size with its 1 million...
| Model | Creator | Score | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|---|
| GPT-5.5openai/gpt-5.5 | 1227.0 | 1.05M | $5 | $30 | 2026-04-23 | |||
| OpenAI: GPT-5.5 (batch)openai/gpt-5.5:batch | 1227.0 | 1.05M | $2.5 | $15 | — | |||
| SpaceXAI: Grok 4.20x-ai/grok-4.20 | 1222.0 | 2M | $1.25 | $2.5 | — | |||
| Google: Gemini 3 Flash Preview (batch)google/gemini-3-flash-preview:batch | 1212.0 | 1.04858M | $0.25 | $1.5 | — | |||
| Gemini 3 Flash Previewgoogle/gemini-3-flash-preview | 1212.0 | 1.04858M | $0.5 | $3 | 2025-12-17 | |||
| Qwen: Qwen3.5 397B A17Bqwen/qwen3.5-397b-a17b | 1190.0 | 262.144K | $0.55 | $3.5 | — | |||
| Thinking Machines: Inkling (batch)thinkingmachines/inkling:batch | 1187.0 | 524.288K | $1 | $4.05 | — | |||
| Anthropic: Claude Sonnet 4.5anthropic/claude-sonnet-4.5 | 1187.0 | 1M | $3 | $15 | — | |||
| Thinking Machines: Inkling (free)thinkingmachines/inkling:free | 1187.0 | 1.04858M | Free | Free | — | |||
| Inklingthinkingmachines/inkling | 1187.0 | 1.04858M | $1.87 | $4.68 | 2026-07-15 | |||
| Anthropic: Claude Sonnet 4.5 (batch)anthropic/claude-sonnet-4.5:batch | 1187.0 | 1M | $1.5 | $7.5 | — | |||
| Anthropic: Claude Opus 4.1 (batch)anthropic/claude-opus-4.1:batch | 1182.0 | 200K | $7.5 | $37.5 | — | |||
| Anthropic: Claude Opus 4.1anthropic/claude-opus-4.1 | 1182.0 | 200K | $15 | $75 | — | |||
| Anthropic: Claude Sonnet 4anthropic/claude-sonnet-4 | 1168.0 | 200K | $3 | $15 | — | |||
| Anthropic: Claude Opus 4anthropic/claude-opus-4 | 1168.0 | 200K | $15 | $75 | — | |||
| Step 3.7 Flashstepfun/step-3.7-flash | 1156.0 | 256K | $0.185 | $1.11 | 2026-05-29 | |||
| SpaceXAI: Grok 4.3x-ai/grok-4.3 | 1154.0 | 1M | $1.25 | $2.5 | — | |||
| SpaceXAI: Grok 4.3 (batch)x-ai/grok-4.3:batch | 1154.0 | 1M | $1 | $2 | — | |||
| Qwen: Qwen3.5 Plus 2026-02-15qwen/qwen3.5-plus-02-15 | 1148.0 | 1M | $0.26 | $1.56 | — | |||
| OpenAI: GPT-5.4 (batch)openai/gpt-5.4:batch | 1130.0 | 1.05M | $1.25 | $7.5 | — | |||
| GPT-5.4openai/gpt-5.4 | 1130.0 | 1.05M | $2.5 | $15 | 2026-03-05 | |||
| Mistral: Mistral Large 3 2512 (batch)mistralai/mistral-large-2512:batch | 1128.0 | 262.144K | $0.25 | $0.75 | — | |||
| Mistral: Mistral Large 3 2512mistralai/mistral-large-2512 | 1128.0 | 262.144K | $0.5 | $1.5 | — | |||
| Mistral: Mistral Medium 3mistralai/mistral-medium-3 | 1112.0 | 131.072K | $0.4 | $2 | — | |||
| Mistral: Mistral Medium 3.1 (batch)mistralai/mistral-medium-3.1:batch | 1110.0 | 131.072K | $0.2 | $1 | — | |||
| Mistral: Mistral Medium 3.1mistralai/mistral-medium-3.1 | 1110.0 | 131.072K | $0.4 | $2 | — | |||
| Gemini 2.5 Progoogle/gemini-2.5-pro | 1109.0 | 1.04858M | $1.25 | $10 | 2025-06-17 | |||
| Google: Gemini 2.5 Pro (batch)google/gemini-2.5-pro:batch | 1109.0 | 1.04858M | $0.625 | $5 | — | |||
| OpenAI: GPT-5.2 (batch)openai/gpt-5.2:batch | 1108.0 | 400K | $0.875 | $7 | — | |||
| GPT-5.2openai/gpt-5.2 | 1108.0 | 400K | $1.75 | $14 | 2025-12-11 | |||
| Anthropic: Claude Haiku 4.5 (batch)anthropic/claude-haiku-4.5:batch | 1101.0 | 200K | $0.5 | $2.5 | — | |||
| Anthropic: Claude Haiku 4.5anthropic/claude-haiku-4.5 | 1101.0 | 200K | $1 | $5 | — | |||
| Google: Gemini 2.5 Flash (batch)google/gemini-2.5-flash:batch | 1100.0 | 1.04858M | $0.15 | $1.25 | — | |||
| Gemini 2.5 Flashgoogle/gemini-2.5-flash | 1100.0 | 1.04858M | $0.3 | $2.5 | 2025-06-17 | |||
| GPT-5.1openai/gpt-5.1 | 1090.0 | 400K | $1.25 | $10 | 2025-11-13 | |||
| OpenAI: GPT-5.1 (batch)openai/gpt-5.1:batch | 1090.0 | 400K | $0.625 | $5 | — | |||
| GPT-5openai/gpt-5 | 1084.0 | 400K | $1.25 | $10 | 2025-08-07 | |||
| OpenAI: GPT-5 (batch)openai/gpt-5:batch | 1084.0 | 400K | $0.625 | $5 | — | |||
| Gemini 3.1 Flash Lite Previewgoogle/gemini-3.1-flash-lite-preview | 1077.0 | 1.04858M | $0.25 | $1.5 | 2026-03-03 | |||
| GPT-5 Miniopenai/gpt-5-mini | 1066.0 | 400K | $0.25 | $2 | 2025-08-07 | |||
| OpenAI: GPT-5 Mini (batch)openai/gpt-5-mini:batch | 1066.0 | 400K | $0.125 | $1 | — | |||
| Mistral: Ministral 3 8B 2512mistralai/ministral-8b-2512 | 1061.0 | 262.144K | $0.15 | $0.15 | — | |||
| Mistral: Ministral 3 8B 2512 (batch)mistralai/ministral-8b-2512:batch | 1061.0 | 262.144K | $0.075 | $0.075 | — | |||
| GPT-5.3 Codexopenai/gpt-5.3-codex | 1037.0 | 400K | $1.75 | $14 | 2026-02-05 | |||
| Mistral: Ministral 3 14B 2512mistralai/ministral-14b-2512 | 1019.0 | 262.144K | $0.2 | $0.2 | — | |||
| GPT-5.1 Codex miniopenai/gpt-5.1-codex-mini | 1017.0 | 400K | $0.22 | $1.8 | 2025-11-13 | |||
| OpenAI: GPT-5 Nano (batch)openai/gpt-5-nano:batch | 994.0 | 400K | $0.025 | $0.2 | — | |||
| GPT-5 Nanoopenai/gpt-5-nano | 994.0 | 400K | $0.05 | $0.4 | 2025-08-07 | |||
| Mistral: Ministral 3 3B 2512mistralai/ministral-3b-2512 | 991.0 | 131.072K | $0.1 | $0.1 | — | |||
| OpenAI: GPT-4.1 Nano (batch)openai/gpt-4.1-nano:batch | 953.0 | 1.04758M | $0.05 | $0.2 | — |