GPT-5 is OpenAI’s most advanced model, offering major improvements in reasoning, code quality, and user experience. It is optimized for complex tasks that require step-by-step reasoning, instruction following, and accuracy...
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
The o-series of models are trained with reinforcement learning to think before they answer and perform complex reasoning. The o3-pro model uses more compute to think harder and provide consistently...
Gemini 2.5 Pro is Google’s state-of-the-art AI model designed for advanced reasoning, coding, mathematics, and scientific tasks. It employs “thinking” capabilities, enabling it to reason through responses with enhanced accuracy...
o3 is a well-rounded and powerful model across domains. It sets a new standard for math, science, coding, and visual reasoning tasks. It also excels at technical writing and instruction-following....
DeepSeek reasoning model for multi-step analysis, math, coding, and tools
Flagship Claude model for deep reasoning, coding, and long-horizon agents
Flagship Claude model for deep reasoning, coding, and long-horizon agents
OpenAI o4-mini is a compact reasoning model in the o-series, optimized for fast, cost-efficient performance while retaining strong multimodal and agentic capabilities. It supports tool use and demonstrates competitive reasoning...
Balanced Claude model for coding, analysis, agent workflows, and cost control
The latest and strongest model family from OpenAI, o1 is designed to spend more time thinking before responding. The o1 model series is trained with large-scale reinforcement learning to reason...
Balanced Claude model for coding, analysis, agent workflows, and cost control
Balanced Claude model for coding, analysis, agent workflows, and cost control
OpenAI o3-mini is a cost-efficient language model optimized for STEM reasoning tasks, particularly excelling in science, mathematics, and coding. This model supports the `reasoning_effort` parameter, which can be set to...
Gemini 2.5 Flash is Google's state-of-the-art workhorse model, specifically designed for advanced reasoning, coding, mathematics, and scientific tasks. It includes built-in "thinking" capabilities, enabling it to provide responses with greater...
GPT-4.1 is a flagship large language model optimized for advanced instruction following, real-world software engineering, and long-context reasoning. It supports a 1 million token context window and outperforms GPT-4o and...
Balanced Claude model for coding, analysis, agent workflows, and cost control
GPT-4.1 Mini is a mid-sized model delivering performance competitive with GPT-4o at substantially lower latency and cost. It retains a 1 million token context window and scores 45.1% on hard...
Fast Claude model for responsive assistance, classification, and lightweight agents
Open multimodal Llama for strong reasoning with efficient everyday serving
Cohere command model for multilingual enterprise agents, tools, and chat
Mistral code model for completions, refactors, and developer IDE workflows
For tasks that demand low latency, GPT‑4.1 nano is the fastest and cheapest model in the GPT-4.1 series. It delivers exceptional performance at a small size with its 1 million...
| Model | Creator | Score | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|---|
| OpenAI: GPT-5openai/gpt-5 | 88.0 | 400K | $1.25 | $10 | 2025-08-07 | |||
| OpenAI: o3 Proopenai/o3-pro | 84.9 | 200K | $20 | $80 | 2025-06-10 | |||
| Google: Gemini 2.5 Progoogle/gemini-2.5-pro | 83.1 | 1.04858M | $1.25 | $10 | 2025-06-17 | |||
| OpenAI: o3openai/o3 | 81.3 | 200K | $2 | $8 | 2025-04-16 | |||
| DeepSeek Reasonerdeepseek/deepseek-reasoner | 74.2 | 1M | $0.147 | $0.295 | 2025-12-01 | |||
| Claude Opus 4anthropic/claude-opus-4-20250514 | 72.0 | 200K | $15 | $75 | 2025-05-22 | |||
| Claude Opus 4 (latest)anthropic/claude-opus-4-0 | 72.0 | 200K | — | — | 2025-05-22 | |||
| OpenAI: o4 Miniopenai/o4-mini | 72.0 | 200K | $1.1 | $4.4 | 2025-04-16 | |||
| Claude Sonnet 3.7anthropic/claude-3-7-sonnet-20250219 | 64.9 | 200K | $3 | $15 | 2025-02-19 | |||
| OpenAI: o1openai/o1 | 61.7 | 200K | $15 | $60 | 2024-12-05 | |||
| Claude Sonnet 4 (latest)anthropic/claude-sonnet-4-0 | 61.3 | 200K | $2.898 | $14.493 | 2025-05-22 | |||
| Claude Sonnet 4anthropic/claude-sonnet-4-20250514 | 61.3 | 200K | $3 | $15 | 2025-05-22 | |||
| OpenAI: o3 Miniopenai/o3-mini | 60.4 | 200K | $1.1 | $4.4 | 2024-12-20 | |||
| Google: Gemini 2.5 Flashgoogle/gemini-2.5-flash | 55.1 | 1.04858M | $0.3 | $2.5 | 2025-06-17 | |||
| OpenAI: GPT-4.1openai/gpt-4.1 | 52.4 | 1.04758M | $2 | $8 | 2025-04-14 | |||
| Claude Sonnet 3.5 v2anthropic/claude-3-5-sonnet-20241022 | 51.6 | 200K | — | — | 2024-10-22 | |||
| OpenAI: GPT-4.1 Miniopenai/gpt-4.1-mini | 32.4 | 1.04758M | $0.4 | $1.6 | 2025-04-14 | |||
| Claude Haiku 3.5anthropic/claude-3-5-haiku-20241022 | 28.0 | 200K | $0.8 | $4 | 2024-10-22 | |||
| Llama 4 Maverick 17B Instructmeta/llama-4-maverick-17b-instruct | 15.6 | 1M | $0.14 | $0.59 | 2025-04-05 | |||
| Command Acohere/command-a-03-2025 | 12.0 | 256K | $2.5 | $10 | 2025-03-13 | |||
| Codestral (latest)mistral/codestral-latest | 11.1 | 256K | $0.3 | $0.9 | 2024-05-29 | |||
| OpenAI: GPT-4.1 Nanoopenai/gpt-4.1-nano | 8.9 | 1.04758M | $0.1 | $0.4 | 2025-04-14 |