Coding-optimized GPT model for repository edits, reviews, and agentic software work
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
Claude Sonnet 4 significantly enhances the capabilities of its predecessor, Sonnet 3.7, excelling in both coding and reasoning tasks with improved precision and controllability. Achieving state-of-the-art performance on SWE-bench (72.7%),...
GPT-5 Mini is a compact version of GPT-5, designed to handle lighter-weight reasoning tasks. It provides the same instruction-following and safety-tuning benefits as GPT-5, but with reduced latency and cost....
Mistral Large 3 2512 is Mistral’s most capable model to date, featuring a sparse mixture-of-experts architecture with 41B active parameters (675B total), and released under the Apache 2.0 license.
Mistral Large 3 2512 is Mistral’s most capable model to date, featuring a sparse mixture-of-experts architecture with 41B active parameters (675B total), and released under the Apache 2.0 license.
Claude Haiku 4.5 is Anthropic’s fastest and most efficient model, delivering near-frontier intelligence at a fraction of the cost and latency of larger Claude models. Matching Claude Sonnet 4’s performance...
Claude Haiku 4.5 is Anthropic’s fastest and most efficient model, delivering near-frontier intelligence at a fraction of the cost and latency of larger Claude models. Matching Claude Sonnet 4’s performance...
Mistral Medium 3.1 is an updated version of Mistral Medium 3, which is a high-performance enterprise-grade language model designed to deliver frontier-level capabilities at significantly reduced operational cost. It balances...
Mistral Medium 3.1 is an updated version of Mistral Medium 3, which is a high-performance enterprise-grade language model designed to deliver frontier-level capabilities at significantly reduced operational cost. It balances...
Gemini 2.5 Flash is Google's state-of-the-art workhorse model, specifically designed for advanced reasoning, coding, mathematics, and scientific tasks. It includes built-in "thinking" capabilities, enabling it to provide responses with greater...
Fast Gemini workhorse for multimodal apps where latency and price matter
Low-latency Gemini model for high-volume multimodal and agent workloads
GPT-5-Nano is the smallest and fastest variant in the GPT-5 system, optimized for developer tools, rapid interactions, and ultra-low latency environments. While limited in reasoning depth compared to its larger...
Mistral Medium 3 is a high-performance enterprise-grade language model designed to deliver frontier-level capabilities at significantly reduced operational cost. It balances state-of-the-art reasoning and multimodal performance with 8× lower cost...
Mistral's cutting-edge language model for coding released end of July 2025. Codestral specializes in low-latency, high-frequency tasks such as fill-in-the-middle (FIM), code correction and test generation. [Blog Post](https://mistral.ai/news/codestral-25-08)
Mistral's cutting-edge language model for coding released end of July 2025. Codestral specializes in low-latency, high-frequency tasks such as fill-in-the-middle (FIM), code correction and test generation. [Blog Post](https://mistral.ai/news/codestral-25-08)
Deliberate o-series reasoner for hard math, coding, and multi-step analysis
o3 is a well-rounded and powerful model across domains. It sets a new standard for math, science, coding, and visual reasoning tasks. It also excels at technical writing and instruction-following....
Long-lived GPT workhorse for coding, instruction following, and production apps
GPT-4.1 is a flagship large language model optimized for advanced instruction following, real-world software engineering, and long-context reasoning. It supports a 1 million token context window and outperforms GPT-4o and...
OpenAI o4-mini is a compact reasoning model in the o-series, optimized for fast, cost-efficient performance while retaining strong multimodal and agentic capabilities. It supports tool use and demonstrates competitive reasoning...
GPT-4.1 Mini is a mid-sized model delivering performance competitive with GPT-4o at substantially lower latency and cost. It retains a 1 million token context window and scores 45.1% on hard...
Affordable GPT-4.1 lane for fast coding help and structured extraction
For tasks that demand low latency, GPT‑4.1 nano is the fastest and cheapest model in the GPT-4.1 series. It delivers exceptional performance at a small size with its 1 million...
Omni-era GPT for multimodal chat, practical coding, and general assistants
GPT-4o ("o" for "omni") is OpenAI's latest AI model, supporting both text and image inputs with text outputs. It maintains the intelligence level of [GPT-4 Turbo](/models/openai/gpt-4-turbo) while being twice as...
| Model | Creator | Score | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|---|
| GPT-5.3 Codexopenai/gpt-5.3-codex | 1152.0 | 400K | $1.75 | $14 | 2026-02-05 | |||
| Anthropic: Claude Sonnet 4anthropic/claude-sonnet-4 | 1143.0 | 200K | $3 | $15 | — | |||
| OpenAI: GPT-5 Mini (batch)openai/gpt-5-mini:batch | 1121.0 | 400K | $0.125 | $1 | — | |||
| Mistral: Mistral Large 3 2512mistralai/mistral-large-2512 | 1117.0 | 262.144K | $0.5 | $1.5 | — | |||
| Mistral: Mistral Large 3 2512 (batch)mistralai/mistral-large-2512:batch | 1117.0 | 262.144K | $0.25 | $0.75 | — | |||
| Anthropic: Claude Haiku 4.5anthropic/claude-haiku-4.5 | 1115.0 | 200K | $1 | $5 | — | |||
| Anthropic: Claude Haiku 4.5 (batch)anthropic/claude-haiku-4.5:batch | 1115.0 | 200K | $0.5 | $2.5 | — | |||
| Mistral: Mistral Medium 3.1mistralai/mistral-medium-3.1 | 1114.0 | 131.072K | $0.4 | $2 | — | |||
| Mistral: Mistral Medium 3.1 (batch)mistralai/mistral-medium-3.1:batch | 1114.0 | 131.072K | $0.2 | $1 | — | |||
| Google: Gemini 2.5 Flash (batch)google/gemini-2.5-flash:batch | 1107.0 | 1.04858M | $0.15 | $1.25 | — | |||
| Gemini 2.5 Flashgoogle/gemini-2.5-flash | 1107.0 | 1.04858M | $0.3 | $2.5 | 2025-06-17 | |||
| Gemini 3.1 Flash Lite Previewgoogle/gemini-3.1-flash-lite-preview | 1086.0 | 1.04858M | $0.25 | $1.5 | 2026-03-03 | |||
| OpenAI: GPT-5 Nano (batch)openai/gpt-5-nano:batch | 1081.0 | 400K | $0.025 | $0.2 | — | |||
| Mistral: Mistral Medium 3mistralai/mistral-medium-3 | 1039.0 | 131.072K | $0.4 | $2 | — | |||
| Mistral: Codestral 2508mistralai/codestral-2508 | 1032.0 | 256K | $0.3 | $0.9 | — | |||
| Mistral: Codestral 2508 (batch)mistralai/codestral-2508:batch | 1032.0 | 256K | $0.15 | $0.45 | — | |||
| o3openai/o3 | 1031.0 | 200K | $2 | $8 | 2025-04-16 | |||
| OpenAI: o3 (batch)openai/o3:batch | 1031.0 | 200K | $1 | $4 | — | |||
| GPT-4.1openai/gpt-4.1 | 1015.0 | 1.04758M | $2 | $8 | 2025-04-14 | |||
| OpenAI: GPT-4.1 (batch)openai/gpt-4.1:batch | 1015.0 | 1.04758M | $1 | $4 | — | |||
| OpenAI: o4 Mini (batch)openai/o4-mini:batch | 998.0 | 200K | $0.55 | $2.2 | — | |||
| OpenAI: GPT-4.1 Mini (batch)openai/gpt-4.1-mini:batch | 976.0 | 1.04758M | $0.2 | $0.8 | — | |||
| GPT-4.1 miniopenai/gpt-4.1-mini | 976.0 | 1.04758M | $0.4 | $1.6 | 2025-04-14 | |||
| OpenAI: GPT-4.1 Nano (batch)openai/gpt-4.1-nano:batch | 931.0 | 1.04758M | $0.05 | $0.2 | — | |||
| GPT-4oopenai/gpt-4o | 900.0 | 128K | $2.5 | $10 | 2024-05-13 | |||
| OpenAI: GPT-4o (batch)openai/gpt-4o:batch | 900.0 | 128K | $1.25 | $5 | — |