13 models
Ranked by Artificial Analysis Coding Index
Reasoning Tools JSON 38.9

Coding-optimized GPT model for repository edits, reviews, and agentic software work

openai/gpt-5-codex 2025-09-15 400K context $1.1/M input $9/M output
13 providers
Reasoning Tools JSON Open weights 37.1

Newer StepFun flash model for faster agents, coding, and multimodal prompts

stepfun/step-3.7-flash 2026-05-29 256K context $0.185/M input $1.11/M output
17 providers
Reasoning Tools JSON 32.0

Google's proven reasoning model for coding, math, and multimodal analysis

google/gemini-2.5-pro 2025-06-17 1.04858M context $1.25/M input $10/M output
29 providers
Open weights 31.6

StepFun flash lane for quick multimodal reasoning and coding assistance

stepfun/step-3.5-flash 2026-01-29 256K context $0.1/M input $0.3/M output
14 providers
26.4

Flagship Qwen3 model for coding agents, complex reasoning, and tool use

alibaba/qwen3-max 2025-09-23 262.144K context $1.2/M input $6/M output
20 providers
Tools Open weights 23.7

Mistral's coding-agent model for repository work, terminal tasks, and software fixes

mistral/devstral-2512 2025-12-09 262.144K context $0.4/M input $2/M output
13 providers
Open weights 22.7

Mistral's largest general model for enterprise agents, coding, and multilingual reasoning

mistral/mistral-large-2512 2024-11-01 262.144K context $0.5/M input $1.5/M output
12 providers
Reasoning Tools JSON 22.2

Fast Gemini workhorse for multimodal apps where latency and price matter

google/gemini-2.5-flash 2025-06-17 1.04858M context $0.3/M input $2.5/M output
30 providers
21.5

Compact GPT model for low-latency assistance and high-volume workloads

openai/gpt-4-turbo 2023-11-06 128K context $10/M input $30/M output
15 providers
Reasoning Tools Open weights 15.9

Classic open reasoning model for transparent math, coding, and deliberate problem solving

deepseek/deepseek-r1 2025-01-20 128K context $0.7/M input $2.5/M output
14 providers
Tools Open weights 10.7

Popular open Llama workhorse for multilingual chat, coding, and self-hosting

meta/llama-3.3-70b-instruct 2024-12-06 128K context $0.1/M input $0.32/M output
25 providers
10.7

Compact GPT model for low-latency assistance and high-volume workloads

openai/gpt-3.5-turbo 2023-03-01 16.385K context $0.5/M input $1.5/M output
13 providers
Reasoning Tools JSON 9.5

Lean Gemini 2.5 lane for cheap multimodal traffic and quick agents

google/gemini-2.5-flash-lite 2025-06-17 1.04858M context $0.1/M input $0.4/M output
18 providers