GPT-5 is OpenAI’s most advanced model, offering major improvements in reasoning, code quality, and user experience. It is optimized for complex tasks that require step-by-step reasoning, instruction following, and accuracy...
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
The o-series of models are trained with reinforcement learning to think before they answer and perform complex reasoning. The o3-pro model uses more compute to think harder and provide consistently...
Gemini 2.5 Pro is Google’s state-of-the-art AI model designed for advanced reasoning, coding, mathematics, and scientific tasks. It employs “thinking” capabilities, enabling it to reason through responses with enhanced accuracy...
o3 is a well-rounded and powerful model across domains. It sets a new standard for math, science, coding, and visual reasoning tasks. It also excels at technical writing and instruction-following....
OpenAI o4-mini is a compact reasoning model in the o-series, optimized for fast, cost-efficient performance while retaining strong multimodal and agentic capabilities. It supports tool use and demonstrates competitive reasoning...
Flagship Claude model for deep reasoning, coding, and long-horizon agents
Balanced Claude model for coding, analysis, agent workflows, and cost control
The latest and strongest model family from OpenAI, o1 is designed to spend more time thinking before responding. The o1 model series is trained with large-scale reinforcement learning to reason...
Balanced Claude model for coding, analysis, agent workflows, and cost control
OpenAI o3-mini is a cost-efficient language model optimized for STEM reasoning tasks, particularly excelling in science, mathematics, and coding. This model supports the `reasoning_effort` parameter, which can be set to...
Gemini 2.5 Flash is Google's state-of-the-art workhorse model, specifically designed for advanced reasoning, coding, mathematics, and scientific tasks. It includes built-in "thinking" capabilities, enabling it to provide responses with greater...
GPT-4.1 is a flagship large language model optimized for advanced instruction following, real-world software engineering, and long-context reasoning. It supports a 1 million token context window and outperforms GPT-4o and...
GPT-4.1 Mini is a mid-sized model delivering performance competitive with GPT-4o at substantially lower latency and cost. It retains a 1 million token context window and scores 45.1% on hard...
GPT-4o ("o" for "omni") is OpenAI's latest AI model, supporting both text and image inputs with text outputs. It maintains the intelligence level of [GPT-4 Turbo](/models/openai/gpt-4-turbo) while being twice as...
The 2024-08-06 version of GPT-4o offers improved performance in structured outputs, with the ability to supply a JSON schema in the respone_format. Read more [here](https://openai.com/index/introducing-structured-outputs-in-the-api/). GPT-4o ("o" for "omni") is...
Flagship Qwen model for complex reasoning, coding, and agentic workflows
The 2024-11-20 version of GPT-4o offers a leveled-up creative writing ability with more natural, engaging, and tailored writing to improve relevance & readability. It’s also better at working with uploaded...
For tasks that demand low latency, GPT‑4.1 nano is the fastest and cheapest model in the GPT-4.1 series. It delivers exceptional performance at a small size with its 1 million...
GPT-4o mini is OpenAI's newest model after [GPT-4 Omni](/models/openai/gpt-4o), supporting both text and image inputs with text outputs. As their most advanced small model, it is many multiples more affordable...
| Model | Creator | Score | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|---|
| OpenAI: GPT-5openai/gpt-5 | 88.0 | 400K | $1.25 | $10 | 2025-08-07 | |||
| OpenAI: o3 Proopenai/o3-pro | 84.9 | 200K | $20 | $80 | 2025-06-10 | |||
| Google: Gemini 2.5 Progoogle/gemini-2.5-pro | 83.1 | 1.04858M | $1.25 | $10 | 2025-06-17 | |||
| OpenAI: o3openai/o3 | 81.3 | 200K | $2 | $8 | 2025-04-16 | |||
| OpenAI: o4 Miniopenai/o4-mini | 72.0 | 200K | $1.1 | $4.4 | 2025-04-16 | |||
| Claude Opus 4anthropic/claude-opus-4-20250514 | 72.0 | 200K | $15 | $75 | 2025-05-22 | |||
| Claude Sonnet 3.7anthropic/claude-3-7-sonnet-20250219 | 64.9 | 200K | $3 | $15 | 2025-02-19 | |||
| OpenAI: o1openai/o1 | 61.7 | 200K | $15 | $60 | 2024-12-05 | |||
| Claude Sonnet 4 (latest)anthropic/claude-sonnet-4-0 | 61.3 | 200K | $2.898 | $14.493 | 2025-05-22 | |||
| OpenAI: o3 Miniopenai/o3-mini | 60.4 | 200K | $1.1 | $4.4 | 2024-12-20 | |||
| Google: Gemini 2.5 Flashgoogle/gemini-2.5-flash | 55.1 | 1.04858M | $0.3 | $2.5 | 2025-06-17 | |||
| OpenAI: GPT-4.1openai/gpt-4.1 | 52.4 | 1.04758M | $2 | $8 | 2025-04-14 | |||
| OpenAI: GPT-4.1 Miniopenai/gpt-4.1-mini | 32.4 | 1.04758M | $0.4 | $1.6 | 2025-04-14 | |||
| OpenAI: GPT-4oopenai/gpt-4o | 23.1 | 128K | $2.5 | $10 | 2024-05-13 | |||
| OpenAI: GPT-4o (2024-08-06)openai/gpt-4o-2024-08-06 | 23.1 | 128K | $2.5 | $10 | 2024-08-06 | |||
| Qwen Maxalibaba/qwen-max | 21.8 | 32.768K | $1.6 | $6.4 | 2024-04-03 | |||
| OpenAI: GPT-4o (2024-11-20)openai/gpt-4o-2024-11-20 | 18.2 | 128K | $2.5 | $10 | 2024-11-20 | |||
| OpenAI: GPT-4.1 Nanoopenai/gpt-4.1-nano | 8.9 | 1.04758M | $0.1 | $0.4 | 2025-04-14 | |||
| OpenAI: GPT-4o-miniopenai/gpt-4o-mini | 3.6 | 128K | $0.15 | $0.6 | 2024-07-18 |