19 models
Ranked by Aider Polyglot
Tools 88.0

GPT-5 is OpenAI’s most advanced model, offering major improvements in reasoning, code quality, and user experience. It is optimized for complex tasks that require step-by-step reasoning, instruction following, and accuracy...

openai/gpt-5 2025-08-07 400K context $1.25/M input $10/M output
30 providers
Reasoning Tools 84.9

The o-series of models are trained with reinforcement learning to think before they answer and perform complex reasoning. The o3-pro model uses more compute to think harder and provide consistently...

openai/o3-pro 2025-06-10 200K context $20/M input $80/M output
8 providers
Reasoning 81.3

o3 is a well-rounded and powerful model across domains. It sets a new standard for math, science, coding, and visual reasoning tasks. It also excels at technical writing and instruction-following....

openai/o3 2025-04-16 200K context $2/M input $8/M output
20 providers
Open weights 74.2

DeepSeek reasoning model for multi-step analysis, math, coding, and tools

deepseek/deepseek-reasoner 2025-12-01 1M context $0.147/M input $0.295/M output
6 providers
Reasoning Tools JSON 72.0

OpenAI o4-mini is a compact reasoning model in the o-series, optimized for fast, cost-efficient performance while retaining strong multimodal and agentic capabilities. It supports tool use and demonstrates competitive reasoning...

openai/o4-mini 2025-04-16 200K context $1.1/M input $4.4/M output
20 providers
Tools Open weights 70.2

DeepSeek-V3 is the latest model from the DeepSeek team, building upon the instruction following and coding abilities of the previous versions. Pre-trained on nearly 15 trillion tokens, the reported evaluations...

deepseek/deepseek-chat 2025-12-01 128K context $0.257/M input $1.029/M output
8 providers
Reasoning 61.7

The latest and strongest model family from OpenAI, o1 is designed to spend more time thinking before responding. The o1 model series is trained with large-scale reinforcement learning to reason...

openai/o1 2024-12-05 200K context $15/M input $60/M output
16 providers
Reasoning Tools JSON 60.4

OpenAI o3-mini is a cost-efficient language model optimized for STEM reasoning tasks, particularly excelling in science, mathematics, and coding. This model supports the `reasoning_effort` parameter, which can be set to...

openai/o3-mini 2024-12-20 200K context $1.1/M input $4.4/M output
19 providers
Reasoning Tools Open weights 56.9

DeepSeek R1 is here: Performance on par with [OpenAI o1](/openai/o1), but open-sourced and with fully open reasoning tokens. It's 671B parameters in size, with 37B active in an inference pass....

deepseek/deepseek-r1 2025-01-20 64K context $0.7/M input $2.5/M output
14 providers
Tools JSON 52.4

GPT-4.1 is a flagship large language model optimized for advanced instruction following, real-world software engineering, and long-context reasoning. It supports a 1 million token context window and outperforms GPT-4o and...

openai/gpt-4.1 2025-04-14 1.04758M context $2/M input $8/M output
28 providers
32.4

GPT-4.1 Mini is a mid-sized model delivering performance competitive with GPT-4o at substantially lower latency and cost. It retains a 1 million token context window and scores 45.1% on hard...

openai/gpt-4.1-mini 2025-04-14 1.04758M context $0.4/M input $1.6/M output
25 providers
Tools 23.1

The 2024-08-06 version of GPT-4o offers improved performance in structured outputs, with the ability to supply a JSON schema in the respone_format. Read more [here](https://openai.com/index/introducing-structured-outputs-in-the-api/). GPT-4o ("o" for "omni") is...

openai/gpt-4o-2024-08-06 2024-08-06 128K context $2.5/M input $10/M output
7 providers
Tools 23.1

GPT-4o ("o" for "omni") is OpenAI's latest AI model, supporting both text and image inputs with text outputs. It maintains the intelligence level of [GPT-4 Turbo](/models/openai/gpt-4-turbo) while being twice as...

openai/gpt-4o 2024-05-13 128K context $2.5/M input $10/M output
23 providers
Tools 21.8

Flagship Qwen model for complex reasoning, coding, and agentic workflows

alibaba/qwen-max 2024-04-03 32.768K context $1.6/M input $6.4/M output
6 providers
Tools 18.2

The 2024-11-20 version of GPT-4o offers a leveled-up creative writing ability with more natural, engaging, and tailored writing to improve relevance & readability. It’s also better at working with uploaded...

openai/gpt-4o-2024-11-20 2024-11-20 128K context $2.5/M input $10/M output
8 providers
Reasoning Tools Open weights 12.0

Cohere command model for multilingual enterprise agents, tools, and chat

cohere/command-a-03-2025 2025-03-13 256K context $2.5/M input $10/M output
6 providers
Tools Open weights 11.1

Mistral code model for completions, refactors, and developer IDE workflows

mistral/codestral-latest 2024-05-29 256K context $0.3/M input $0.9/M output
5 providers
8.9

For tasks that demand low latency, GPT‑4.1 nano is the fastest and cheapest model in the GPT-4.1 series. It delivers exceptional performance at a small size with its 1 million...

openai/gpt-4.1-nano 2025-04-14 1.04758M context $0.1/M input $0.4/M output
20 providers
3.6

GPT-4o mini is OpenAI's newest model after [GPT-4 Omni](/models/openai/gpt-4o), supporting both text and image inputs with text outputs. As their most advanced small model, it is many multiples more affordable...

openai/gpt-4o-mini 2024-07-18 128K context $0.15/M input $0.6/M output
23 providers