GPT-5 is OpenAI’s most advanced model, offering major improvements in reasoning, code quality, and user experience. It is optimized for complex tasks that require step-by-step reasoning, instruction following, and accuracy...
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
The o-series of models are trained with reinforcement learning to think before they answer and perform complex reasoning. The o3-pro model uses more compute to think harder and provide consistently...
Gemini 2.5 Pro is Google’s state-of-the-art AI model designed for advanced reasoning, coding, mathematics, and scientific tasks. It employs “thinking” capabilities, enabling it to reason through responses with enhanced accuracy...
o3 is a well-rounded and powerful model across domains. It sets a new standard for math, science, coding, and visual reasoning tasks. It also excels at technical writing and instruction-following....
DeepSeek reasoning model for multi-step analysis, math, coding, and tools
OpenAI o4-mini is a compact reasoning model in the o-series, optimized for fast, cost-efficient performance while retaining strong multimodal and agentic capabilities. It supports tool use and demonstrates competitive reasoning...
Flagship Claude model for deep reasoning, coding, and long-horizon agents
Flagship Claude model for deep reasoning, coding, and long-horizon agents
Balanced Claude model for coding, analysis, agent workflows, and cost control
The latest and strongest model family from OpenAI, o1 is designed to spend more time thinking before responding. The o1 model series is trained with large-scale reinforcement learning to reason...
Balanced Claude model for coding, analysis, agent workflows, and cost control
Balanced Claude model for coding, analysis, agent workflows, and cost control
OpenAI o3-mini is a cost-efficient language model optimized for STEM reasoning tasks, particularly excelling in science, mathematics, and coding. This model supports the `reasoning_effort` parameter, which can be set to...
Large open Qwen MoE for multilingual reasoning, coding, and tool use
Gemini 2.5 Flash is Google's state-of-the-art workhorse model, specifically designed for advanced reasoning, coding, mathematics, and scientific tasks. It includes built-in "thinking" capabilities, enabling it to provide responses with greater...
Dense open Qwen model for self-hosted chat, reasoning, and coding
Cohere command model for multilingual enterprise agents, tools, and chat
| Model | Creator | Score | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|---|
| OpenAI: GPT-5openai/gpt-5 | 88.0 | 400K | $1.25 | $10 | 2025-08-07 | |||
| OpenAI: o3 Proopenai/o3-pro | 84.9 | 200K | $20 | $80 | 2025-06-10 | |||
| Google: Gemini 2.5 Progoogle/gemini-2.5-pro | 83.1 | 1.04858M | $1.25 | $10 | 2025-06-17 | |||
| OpenAI: o3openai/o3 | 81.3 | 200K | $2 | $8 | 2025-04-16 | |||
| DeepSeek Reasonerdeepseek/deepseek-reasoner | 74.2 | 1M | $0.147 | $0.295 | 2025-12-01 | |||
| OpenAI: o4 Miniopenai/o4-mini | 72.0 | 200K | $1.1 | $4.4 | 2025-04-16 | |||
| Claude Opus 4 (latest)anthropic/claude-opus-4-0 | 72.0 | 200K | — | — | 2025-05-22 | |||
| Claude Opus 4anthropic/claude-opus-4-20250514 | 72.0 | 200K | $15 | $75 | 2025-05-22 | |||
| Claude Sonnet 3.7anthropic/claude-3-7-sonnet-20250219 | 64.9 | 200K | $3 | $15 | 2025-02-19 | |||
| OpenAI: o1openai/o1 | 61.7 | 200K | $15 | $60 | 2024-12-05 | |||
| Claude Sonnet 4anthropic/claude-sonnet-4-20250514 | 61.3 | 200K | $3 | $15 | 2025-05-22 | |||
| Claude Sonnet 4 (latest)anthropic/claude-sonnet-4-0 | 61.3 | 200K | $2.898 | $14.493 | 2025-05-22 | |||
| OpenAI: o3 Miniopenai/o3-mini | 60.4 | 200K | $1.1 | $4.4 | 2024-12-20 | |||
| Qwen3 235B-A22Balibaba/qwen3-235b-a22b | 59.6 | 131.072K | $0.7 | $2.8 | 2025-04 | |||
| Google: Gemini 2.5 Flashgoogle/gemini-2.5-flash | 55.1 | 1.04858M | $0.3 | $2.5 | 2025-06-17 | |||
| Qwen3 32Balibaba/qwen3-32b | 40.0 | 131.072K | $0.7 | $2.8 | 2025-04 | |||
| Command Acohere/command-a-03-2025 | 12.0 | 256K | $2.5 | $10 | 2025-03-13 |