GPT-5-Codex is a specialized version of GPT-5 optimized for software engineering and coding workflows. It is designed for both interactive development sessions and long, independent execution of complex engineering tasks....
Every result stays attached to its source, version, metric and harness. Scores from different versions are never merged, and the profile below plots each benchmark against its own population rather than on a shared scale.
Each strip shows every published score for that benchmark, with this model marked. The lighter shape behind the ticks is the density of results, and the dashed line is the median.
Cost against capability
Price and performance
Listed input price plotted against Design Arena: webapps, the benchmark with the widest published coverage that this model appears in.
The stepped line is the efficient frontier: at each price, the best score available for that money or less. A model sitting on it is not being beaten by anything cheaper. Price is log-scaled because listed rates span four orders of magnitude. Only models with both a listed price and a score on this benchmark can appear.
Catalog activity
Change log
Field-level changes detected between successful source imports.
Answered directly from the stored record — nothing here is generated beyond the catalog's own fields.
What is OpenAI: GPT-5 Codex (batch)?
GPT-5-Codex is a specialized version of GPT-5 optimized for software engineering and coding workflows. It is designed for both interactive development sessions and long, independent execution of complex engineering tasks. It is published by OpenAI and catalogued here from OpenRouter.
What is the context length of OpenAI: GPT-5 Codex (batch)?
OpenAI: GPT-5 Codex (batch) accepts up to 400K tokens of context and returns up to 128K output tokens.
Does OpenAI: GPT-5 Codex (batch) support tool calling and structured output?