GPT-5 is OpenAI’s most advanced model, offering major improvements in reasoning, code quality, and user experience. It is optimized for complex tasks that require step-by-step reasoning, instruction following, and accuracy...
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
The o-series of models are trained with reinforcement learning to think before they answer and perform complex reasoning. The o3-pro model uses more compute to think harder and provide consistently...
Gemini 2.5 Pro is Google’s state-of-the-art AI model designed for advanced reasoning, coding, mathematics, and scientific tasks. It employs “thinking” capabilities, enabling it to reason through responses with enhanced accuracy...
o3 is a well-rounded and powerful model across domains. It sets a new standard for math, science, coding, and visual reasoning tasks. It also excels at technical writing and instruction-following....
Flagship Claude model for deep reasoning, coding, and long-horizon agents
OpenAI o4-mini is a compact reasoning model in the o-series, optimized for fast, cost-efficient performance while retaining strong multimodal and agentic capabilities. It supports tool use and demonstrates competitive reasoning...
Balanced Claude model for coding, analysis, agent workflows, and cost control
Balanced Claude model for coding, analysis, agent workflows, and cost control
OpenAI o3-mini is a cost-efficient language model optimized for STEM reasoning tasks, particularly excelling in science, mathematics, and coding. This model supports the `reasoning_effort` parameter, which can be set to...
Gemini 2.5 Flash is Google's state-of-the-art workhorse model, specifically designed for advanced reasoning, coding, mathematics, and scientific tasks. It includes built-in "thinking" capabilities, enabling it to provide responses with greater...
GPT-4.1 is a flagship large language model optimized for advanced instruction following, real-world software engineering, and long-context reasoning. It supports a 1 million token context window and outperforms GPT-4o and...
GPT-4.1 Mini is a mid-sized model delivering performance competitive with GPT-4o at substantially lower latency and cost. It retains a 1 million token context window and scores 45.1% on hard...
Open multimodal Llama for strong reasoning with efficient everyday serving
For tasks that demand low latency, GPT‑4.1 nano is the fastest and cheapest model in the GPT-4.1 series. It delivers exceptional performance at a small size with its 1 million...
| Model | Creator | Score | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|---|
| OpenAI: GPT-5openai/gpt-5 | 88.0 | 400K | $1.25 | $10 | 2025-08-07 | |||
| OpenAI: o3 Proopenai/o3-pro | 84.9 | 200K | $20 | $80 | 2025-06-10 | |||
| Google: Gemini 2.5 Progoogle/gemini-2.5-pro | 83.1 | 1.04858M | $1.25 | $10 | 2025-06-17 | |||
| OpenAI: o3openai/o3 | 81.3 | 200K | $2 | $8 | 2025-04-16 | |||
| Claude Opus 4anthropic/claude-opus-4-20250514 | 72.0 | 200K | $15 | $75 | 2025-05-22 | |||
| OpenAI: o4 Miniopenai/o4-mini | 72.0 | 200K | $1.1 | $4.4 | 2025-04-16 | |||
| Claude Sonnet 3.7anthropic/claude-3-7-sonnet-20250219 | 64.9 | 200K | $3 | $15 | 2025-02-19 | |||
| Claude Sonnet 4anthropic/claude-sonnet-4-20250514 | 61.3 | 200K | $3 | $15 | 2025-05-22 | |||
| OpenAI: o3 Miniopenai/o3-mini | 60.4 | 200K | $1.1 | $4.4 | 2024-12-20 | |||
| Google: Gemini 2.5 Flashgoogle/gemini-2.5-flash | 55.1 | 1.04858M | $0.3 | $2.5 | 2025-06-17 | |||
| OpenAI: GPT-4.1openai/gpt-4.1 | 52.4 | 1.04758M | $2 | $8 | 2025-04-14 | |||
| OpenAI: GPT-4.1 Miniopenai/gpt-4.1-mini | 32.4 | 1.04758M | $0.4 | $1.6 | 2025-04-14 | |||
| Llama 4 Maverick 17B Instructmeta/llama-4-maverick-17b-instruct | 15.6 | 1M | $0.14 | $0.59 | 2025-04-05 | |||
| OpenAI: GPT-4.1 Nanoopenai/gpt-4.1-nano | 8.9 | 1.04758M | $0.1 | $0.4 | 2025-04-14 |