Small Nemotron 3 MoE for efficient coding, math, and long-context agents
11 providers
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
Small Nemotron 3 MoE for efficient coding, math, and long-context agents
Small omni GPT for cheap multimodal assistance and production-scale traffic
Tiny GPT-4.1 option for classification, routing, and very high-volume tasks
Compact GPT model for low-latency assistance and high-volume workloads
| Model | Creator | Score | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|---|
| Nemotron 3 Nano 30B A3Bnvidia/nemotron-3-nano-30b-a3b | 14.4 | 262.144K | $0.05 | $0.2 | 2025-12-15 | |||
| GPT-4o miniopenai/gpt-4o-mini | 11.4 | 128K | $0.15 | $0.6 | 2024-07-18 | |||
| GPT-4.1 nanoopenai/gpt-4.1-nano | 11.1 | 1.04758M | $0.1 | $0.4 | 2025-04-14 | |||
| GPT-3.5-turboopenai/gpt-3.5-turbo | 10.7 | 16.385K | $0.5 | $1.5 | 2023-03-01 |