22 models
Ranked by Artificial Analysis Coding Index
Reasoning Tools JSON 38.9

Coding-optimized GPT model for repository edits, reviews, and agentic software work

openai/gpt-5-codex 2025-09-15 400K context $1.1/M input $9/M output
13 providers
Reasoning Tools JSON Open weights 37.1

Step 3.7 Flash is StepFun's latest high-efficiency multimodal Mixture-of-Experts model. It pairs a 196B-parameter language backbone with a vision encoder for native image and video understanding, activating roughly 11B parameters...

stepfun/step-3.7-flash 2026-05-29 256K context $0.2/M input $1.15/M output
17 providers
Reasoning Tools JSON 32.0

Gemini 2.5 Pro is Google’s state-of-the-art AI model designed for advanced reasoning, coding, mathematics, and scientific tasks. It employs “thinking” capabilities, enabling it to reason through responses with enhanced accuracy...

google/gemini-2.5-pro 2025-06-17 1.04858M context $1.25/M input $10/M output
29 providers
Open weights 31.6

Step 3.5 Flash is StepFun's most capable open-source foundation model. Built on a sparse Mixture of Experts (MoE) architecture, it selectively activates only 11B of its 196B parameters per token....

stepfun/step-3.5-flash 2026-01-29 262.144K context $0.1/M input $0.3/M output
14 providers
Reasoning Tools Open weights 29.5

Late GLM-4 workhorse for coding agents, reasoning, and structured tasks

zhipuai/glm-4.6 2025-09-30 204.8K context $0.6/M input $2.2/M output
18 providers
26.4

Flagship Qwen3 model for coding agents, complex reasoning, and tool use

alibaba/qwen3-max 2025-09-23 262.144K context $1.2/M input $6/M output
20 providers
Reasoning Tools Open weights 26.3

Hybrid-reasoning GLM release that made the 4.5 line broadly useful

zhipuai/glm-4.5 2025-07-28 131.072K context $0.6/M input $2.2/M output
14 providers
Reasoning Tools Open weights 24.3

Fast Mistral production model for chat, extraction, and cost-sensitive agents

mistral/mistral-small-2603 2026-03-16 256K context $0.15/M input $0.6/M output
12 providers
Tools JSON 24.2

GPT-4o ("o" for "omni") is OpenAI's latest AI model, supporting both text and image inputs with text outputs. It maintains the intelligence level of [GPT-4 Turbo](/models/openai/gpt-4-turbo) while being twice as...

openai/gpt-4o-2024-05-13 2024-05-13 128K context $5/M input $15/M output
5 providers
Reasoning Tools Open weights 23.8

Lighter GLM-4.5 variant for fast coding assistance and cheaper agents

zhipuai/glm-4.5-air 2025-07-28 131.072K context $0.2/M input $1.1/M output
11 providers
Tools Open weights 23.7

Mistral's coding-agent model for repository work, terminal tasks, and software fixes

mistral/devstral-2512 2025-12-09 262.144K context $0.4/M input $2/M output
13 providers
Open weights 22.7

Mistral's largest general model for enterprise agents, coding, and multilingual reasoning

mistral/mistral-large-2512 2024-11-01 262.144K context $0.5/M input $1.5/M output
12 providers
Reasoning Tools JSON 22.2

Gemini 2.5 Flash is Google's state-of-the-art workhorse model, specifically designed for advanced reasoning, coding, mathematics, and scientific tasks. It includes built-in "thinking" capabilities, enabling it to provide responses with greater...

google/gemini-2.5-flash 2025-06-17 1.04858M context $0.3/M input $2.5/M output
30 providers
21.5

The latest GPT-4 Turbo model with vision capabilities. Vision requests can now use JSON mode and function calling. Training data: up to December 2023.

openai/gpt-4-turbo 2023-11-06 128K context $10/M input $30/M output
15 providers
Tools Open weights 19.4

Smaller Qwen coder for efficient local agents and repo-level fixes

alibaba/qwen3-coder-30b-a3b-instruct 2025-04 262.144K context $0.45/M input $2.25/M output
13 providers
Tools 16.7

The 2024-11-20 version of GPT-4o offers a leveled-up creative writing ability with more natural, engaging, and tailored writing to improve relevance & readability. It’s also better at working with uploaded...

openai/gpt-4o-2024-11-20 2024-11-20 128K context $2.5/M input $10/M output
8 providers
Tools 16.6

The 2024-08-06 version of GPT-4o offers improved performance in structured outputs, with the ability to supply a JSON schema in the respone_format. Read more [here](https://openai.com/index/introducing-structured-outputs-in-the-api/). GPT-4o ("o" for "omni") is...

openai/gpt-4o-2024-08-06 2024-08-06 128K context $2.5/M input $10/M output
7 providers
Reasoning Tools Open weights 15.9

DeepSeek R1 is here: Performance on par with [OpenAI o1](/openai/o1), but open-sourced and with fully open reasoning tokens. It's 671B parameters in size, with 37B active in an inference pass....

deepseek/deepseek-r1 2025-01-20 64K context $0.7/M input $2.5/M output
14 providers
Reasoning Tools Open weights 10.9

GLM vision model for visual reasoning, documents, and multimodal agents

zhipuai/glm-4.5v 2025-08-11 64K context $0.6/M input $1.8/M output
7 providers
10.7

GPT-3.5 Turbo is OpenAI's fastest model. It can understand and generate natural language or code, and is optimized for chat and traditional completion tasks. Training data up to Sep 2021.

openai/gpt-3.5-turbo 2023-03-01 16.385K context $0.5/M input $1.5/M output
13 providers
Tools Open weights 10.7

Popular open Llama workhorse for multilingual chat, coding, and self-hosting

meta/llama-3.3-70b-instruct 2024-12-06 128K context $0.1/M input $0.32/M output
25 providers
Reasoning Tools JSON 9.5

Gemini 2.5 Flash-Lite is a lightweight reasoning model in the Gemini 2.5 family, optimized for ultra-low latency and cost efficiency. It offers improved throughput, faster token generation, and better performance...

google/gemini-2.5-flash-lite 2025-06-17 1.04858M context $0.1/M input $0.4/M output
18 providers