Google's proven reasoning model for coding, math, and multimodal analysis
29 providers
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
Google's proven reasoning model for coding, math, and multimodal analysis
Fast Gemini workhorse for multimodal apps where latency and price matter
Long-lived GPT workhorse for coding, instruction following, and production apps
Affordable GPT-4.1 lane for fast coding help and structured extraction
Open multimodal Llama for strong reasoning with efficient everyday serving
Tiny GPT-4.1 option for classification, routing, and very high-volume tasks
| Model | Creator | Score | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|---|
| Gemini 2.5 Progoogle/gemini-2.5-pro | 83.1 | 1.04858M | $1.25 | $10 | 2025-06-17 | |||
| Gemini 2.5 Flashgoogle/gemini-2.5-flash | 55.1 | 1.04858M | $0.3 | $2.5 | 2025-06-17 | |||
| GPT-4.1openai/gpt-4.1 | 52.4 | 1.04758M | $2 | $8 | 2025-04-14 | |||
| GPT-4.1 miniopenai/gpt-4.1-mini | 32.4 | 1.04758M | $0.4 | $1.6 | 2025-04-14 | |||
| Llama 4 Maverick 17B Instructmeta/llama-4-maverick-17b-instruct | 15.6 | 1M | $0.14 | $0.59 | 2025-04-05 | |||
| GPT-4.1 nanoopenai/gpt-4.1-nano | 8.9 | 1.04758M | $0.1 | $0.4 | 2025-04-14 |