Coding-optimized GPT model for repository edits, reviews, and agentic software work
Model catalog
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
13 providers
Newer StepFun flash model for faster agents, coding, and multimodal prompts
17 providers
StepFun flash lane for quick multimodal reasoning and coding assistance
14 providers
Google's proven reasoning model for coding, math, and multimodal analysis
29 providers
Flagship Qwen3 model for coding agents, complex reasoning, and tool use
20 providers
Fast Gemini workhorse for multimodal apps where latency and price matter
30 providers
Classic open reasoning model for transparent math, coding, and deliberate problem solving
14 providers
Lean Gemini 2.5 lane for cheap multimodal traffic and quick agents
18 providers
Popular open Llama workhorse for multilingual chat, coding, and self-hosting
25 providers
| Model | Creator | Score | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|---|
| GPT-5-Codexopenai/gpt-5-codex | 37.9 | 400K | $1.1 | $9 | 2025-09-15 | |||
| Step 3.7 Flashstepfun/step-3.7-flash | 35.6 | 256K | $0.185 | $1.11 | 2026-05-29 | |||
| Step 3.5 Flashstepfun/step-3.5-flash | 27.3 | 256K | $0.1 | $0.3 | 2026-01-29 | |||
| Gemini 2.5 Progoogle/gemini-2.5-pro | 26.5 | 1.04858M | $1.25 | $10 | 2025-06-17 | |||
| Qwen3 Maxalibaba/qwen3-max | 20.5 | 262.144K | $1.2 | $6 | 2025-09-23 | |||
| Gemini 2.5 Flashgoogle/gemini-2.5-flash | 13.6 | 1.04858M | $0.3 | $2.5 | 2025-06-17 | |||
| DeepSeek-R1deepseek/deepseek-r1 | 6.1 | 128K | $0.7 | $2.5 | 2025-01-20 | |||
| Gemini 2.5 Flash-Litegoogle/gemini-2.5-flash-lite | 4.5 | 1.04858M | $0.1 | $0.4 | 2025-06-17 | |||
| Llama-3.3-70B-Instructmeta/llama-3.3-70b-instruct | 3.0 | 128K | $0.1 | $0.32 | 2024-12-06 |