Coding-optimized GPT model for repository edits, reviews, and agentic software work
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
Step 3.7 Flash is StepFun's latest high-efficiency multimodal Mixture-of-Experts model. It pairs a 196B-parameter language backbone with a vision encoder for native image and video understanding, activating roughly 11B parameters...
Gemini 2.5 Pro is Google’s state-of-the-art AI model designed for advanced reasoning, coding, mathematics, and scientific tasks. It employs “thinking” capabilities, enabling it to reason through responses with enhanced accuracy...
Fast Mistral production model for chat, extraction, and cost-sensitive agents
GPT-4o ("o" for "omni") is OpenAI's latest AI model, supporting both text and image inputs with text outputs. It maintains the intelligence level of [GPT-4 Turbo](/models/openai/gpt-4-turbo) while being twice as...
Mistral's largest general model for enterprise agents, coding, and multilingual reasoning
Gemini 2.5 Flash is Google's state-of-the-art workhorse model, specifically designed for advanced reasoning, coding, mathematics, and scientific tasks. It includes built-in "thinking" capabilities, enabling it to provide responses with greater...
The latest GPT-4 Turbo model with vision capabilities. Vision requests can now use JSON mode and function calling. Training data: up to December 2023.
The 2024-11-20 version of GPT-4o offers a leveled-up creative writing ability with more natural, engaging, and tailored writing to improve relevance & readability. It’s also better at working with uploaded...
The 2024-08-06 version of GPT-4o offers improved performance in structured outputs, with the ability to supply a JSON schema in the respone_format. Read more [here](https://openai.com/index/introducing-structured-outputs-in-the-api/). GPT-4o ("o" for "omni") is...
Mistral model for multilingual chat, reasoning, and tool-assisted workflows
GLM vision model for visual reasoning, documents, and multimodal agents
Gemini 2.5 Flash-Lite is a lightweight reasoning model in the Gemini 2.5 family, optimized for ultra-low latency and cost efficiency. It offers improved throughput, faster token generation, and better performance...
| Model | Creator | Score | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|---|
| GPT-5-Codexopenai/gpt-5-codex | 38.9 | 400K | $1.1 | $9 | 2025-09-15 | |||
| StepFun: Step 3.7 Flashstepfun/step-3.7-flash | 37.1 | 256K | $0.2 | $1.15 | 2026-05-29 | |||
| Google: Gemini 2.5 Progoogle/gemini-2.5-pro | 32.0 | 1.04858M | $1.25 | $10 | 2025-06-17 | |||
| Mistral Small 4mistral/mistral-small-2603 | 24.3 | 256K | $0.15 | $0.6 | 2026-03-16 | |||
| OpenAI: GPT-4o (2024-05-13)openai/gpt-4o-2024-05-13 | 24.2 | 128K | $5 | $15 | 2024-05-13 | |||
| Mistral Large 3mistral/mistral-large-2512 | 22.7 | 262.144K | $0.5 | $1.5 | 2024-11-01 | |||
| Google: Gemini 2.5 Flashgoogle/gemini-2.5-flash | 22.2 | 1.04858M | $0.3 | $2.5 | 2025-06-17 | |||
| OpenAI: GPT-4 Turboopenai/gpt-4-turbo | 21.5 | 128K | $10 | $30 | 2023-11-06 | |||
| OpenAI: GPT-4o (2024-11-20)openai/gpt-4o-2024-11-20 | 16.7 | 128K | $2.5 | $10 | 2024-11-20 | |||
| OpenAI: GPT-4o (2024-08-06)openai/gpt-4o-2024-08-06 | 16.6 | 128K | $2.5 | $10 | 2024-08-06 | |||
| Mistral Medium 3mistral/mistral-medium-2505 | 13.6 | 131.072K | $0.4 | $2 | 2025-05-07 | |||
| GLM-4.5Vzhipuai/glm-4.5v | 10.9 | 64K | $0.6 | $1.8 | 2025-08-11 | |||
| Google: Gemini 2.5 Flash Litegoogle/gemini-2.5-flash-lite | 9.5 | 1.04858M | $0.1 | $0.4 | 2025-06-17 |