Tiny GPT-5 lane for routing, extraction, classification, and bulk jobs
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
Original GPT-5 workhorse for reasoning, coding, writing, and tool workflows
Efficient Qwen model for fast chat, extraction, and high-volume workloads
Efficient Mistral model for fast chat, extraction, and production assistants
Deliberate o-series reasoner for hard math, coding, and multi-step analysis
Fast o-series model for compact reasoning, coding, and tool use
Smaller o-series reasoner for economical coding, math, and planning tasks
O-series reasoning model for hard analysis, math, coding, and planning
Qwen instruction model for multilingual chat, reasoning, and tool use
Web-grounded Sonar for multi-step research questions that need cited reasoning
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| GPT-5 Nanoopenai/gpt-5-nano | 400K | $0.05 | $0.4 | 2025-08-07 | |||
| GPT-5openai/gpt-5 | 400K | $1.25 | $10 | 2025-08-07 | |||
| Qwen Flashalibaba/qwen-flash | 1M | $0.05 | $0.4 | 2025-07-28 | |||
| Mistral Small 3.2mistral/mistral-small-2506 | 128K | $0.1 | $0.3 | 2025-06-20 | |||
| o3openai/o3 | 200K | $2 | $8 | 2025-04-16 | |||
| o4-miniopenai/o4-mini | 200K | $1.1 | $4.4 | 2025-04-16 | |||
| o3-miniopenai/o3-mini | 200K | $1.1 | $4.4 | 2024-12-20 | |||
| o1openai/o1 | 200K | $15 | $60 | 2024-12-05 | |||
| Qwen Plusalibaba/qwen-plus | 1M | $0.4 | $1.2 | 2024-01-25 | |||
| Sonar Reasoning Properplexity/sonar-reasoning-pro | 128K | $2 | $8 | 2024-01-01 |