Coding-optimized GPT model for repository edits, reviews, and agentic software work
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
Step 3.7 Flash is StepFun's latest high-efficiency multimodal Mixture-of-Experts model. It pairs a 196B-parameter language backbone with a vision encoder for native image and video understanding, activating roughly 11B parameters...
Step 3.5 Flash is StepFun's most capable open-source foundation model. Built on a sparse Mixture of Experts (MoE) architecture, it selectively activates only 11B of its 196B parameters per token....
Gemini 2.5 Pro is Google’s state-of-the-art AI model designed for advanced reasoning, coding, mathematics, and scientific tasks. It employs “thinking” capabilities, enabling it to reason through responses with enhanced accuracy...
Late GLM-4 workhorse for coding agents, reasoning, and structured tasks
Flagship Qwen3 model for coding agents, complex reasoning, and tool use
Mistral's coding-agent model for repository work, terminal tasks, and software fixes
Fast Mistral production model for chat, extraction, and cost-sensitive agents
Mistral's largest general model for enterprise agents, coding, and multilingual reasoning
Smaller Qwen coder for efficient local agents and repo-level fixes
Gemini 2.5 Flash is Google's state-of-the-art workhorse model, specifically designed for advanced reasoning, coding, mathematics, and scientific tasks. It includes built-in "thinking" capabilities, enabling it to provide responses with greater...
Gemini 2.5 Flash-Lite is a lightweight reasoning model in the Gemini 2.5 family, optimized for ultra-low latency and cost efficiency. It offers improved throughput, faster token generation, and better performance...
| Model | Creator | Score | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|---|
| GPT-5-Codexopenai/gpt-5-codex | 37.9 | 400K | $1.1 | $9 | 2025-09-15 | |||
| StepFun: Step 3.7 Flashstepfun/step-3.7-flash | 35.6 | 256K | $0.2 | $1.15 | 2026-05-29 | |||
| StepFun: Step 3.5 Flashstepfun/step-3.5-flash | 27.3 | 262.144K | $0.1 | $0.3 | 2026-01-29 | |||
| Google: Gemini 2.5 Progoogle/gemini-2.5-pro | 26.5 | 1.04858M | $1.25 | $10 | 2025-06-17 | |||
| GLM-4.6zhipuai/glm-4.6 | 25.0 | 204.8K | $0.6 | $2.2 | 2025-09-30 | |||
| Qwen3 Maxalibaba/qwen3-max | 20.5 | 262.144K | $1.2 | $6 | 2025-09-23 | |||
| Devstral 2mistral/devstral-2512 | 18.9 | 262.144K | $0.4 | $2 | 2025-12-09 | |||
| Mistral Small 4mistral/mistral-small-2603 | 17.4 | 256K | $0.15 | $0.6 | 2026-03-16 | |||
| Mistral Large 3mistral/mistral-large-2512 | 15.9 | 262.144K | $0.5 | $1.5 | 2024-11-01 | |||
| Qwen3-Coder 30B-A3B Instructalibaba/qwen3-coder-30b-a3b-instruct | 15.2 | 262.144K | $0.45 | $2.25 | 2025-04 | |||
| Google: Gemini 2.5 Flashgoogle/gemini-2.5-flash | 13.6 | 1.04858M | $0.3 | $2.5 | 2025-06-17 | |||
| Google: Gemini 2.5 Flash Litegoogle/gemini-2.5-flash-lite | 4.5 | 1.04858M | $0.1 | $0.4 | 2025-06-17 |