Compact GPT model for low-latency assistance and high-volume workloads
15 providers
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
Compact GPT model for low-latency assistance and high-volume workloads
Small omni GPT for cheap multimodal assistance and production-scale traffic
| Model | Creator | Score | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|---|
| GPT-4 Turboopenai/gpt-4-turbo | 31.9 | 128K | $10 | $30 | 2023-11-06 | |||
| GPT-4o miniopenai/gpt-4o-mini | 22.9 | 128K | $0.15 | $0.6 | 2024-07-18 |