Gemini 2.5 Pro is Google’s state-of-the-art AI model designed for advanced reasoning, coding, mathematics, and scientific tasks. It employs “thinking” capabilities, enabling it to reason through responses with enhanced accuracy...
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
Google's proven reasoning model for coding, math, and multimodal analysis
Qwen3.5-9B is a multimodal foundation model from the Qwen3.5 family, designed to deliver strong reasoning, coding, and visual understanding in an efficient 9B-parameter architecture. It uses a unified vision-language design...
Qwen3.5-9B is a multimodal foundation model from the Qwen3.5 family, designed to deliver strong reasoning, coding, and visual understanding in an efficient 9B-parameter architecture. It uses a unified vision-language design...
Mistral Small 4 is the next major release in the Mistral Small family, unifying the capabilities of several flagship Mistral models into a single system. It combines strong reasoning from...
Mistral Small 4 is the next major release in the Mistral Small family, unifying the capabilities of several flagship Mistral models into a single system. It combines strong reasoning from...
GPT model for general reasoning, writing, coding, and tool-assisted tasks
Nova 2 Lite is a fast, cost-effective reasoning model for everyday workloads that can process text, images, and videos to generate text. Nova 2 Lite demonstrates standout capabilities in processing...
Compact GPT model for low-latency assistance and high-volume workloads
The latest GPT-4 Turbo model with vision capabilities. Vision requests can now use JSON mode and function calling. Training data: up to December 2023.
Mistral Medium 3.1 is an updated version of Mistral Medium 3, which is a high-performance enterprise-grade language model designed to deliver frontier-level capabilities at significantly reduced operational cost. It balances...
Mistral Medium 3.1 is an updated version of Mistral Medium 3, which is a high-performance enterprise-grade language model designed to deliver frontier-level capabilities at significantly reduced operational cost. It balances...
GPT-4.1 Mini is a mid-sized model delivering performance competitive with GPT-4o at substantially lower latency and cost. It retains a 1 million token context window and scores 45.1% on hard...
Affordable GPT-4.1 lane for fast coding help and structured extraction
Mistral Large 3 2512 is Mistral’s most capable model to date, featuring a sparse mixture-of-experts architecture with 41B active parameters (675B total), and released under the Apache 2.0 license.
Mistral Large 3 2512 is Mistral’s most capable model to date, featuring a sparse mixture-of-experts architecture with 41B active parameters (675B total), and released under the Apache 2.0 license.
Llama 4 Maverick 17B Instruct (128E) is a high-capacity multimodal language model from Meta, built on a mixture-of-experts (MoE) architecture with 128 experts and 17 billion active parameters per forward...
Small GPT-5 for responsive agents, coding help, and everyday automation
GPT-5 Mini is a compact version of GPT-5, designed to handle lighter-weight reasoning tasks. It provides the same instruction-following and safety-tuning benefits as GPT-5, but with reduced latency and cost....
The largest model in the Ministral 3 family, Ministral 3 14B offers frontier capabilities and performance comparable to its larger Mistral Small 3.2 24B counterpart. A powerful and efficient language...
NVIDIA Nemotron™ 3 Nano Omni is a 30B-A3B open multimodal model designed to function as a perception and context sub-agent in enterprise agent systems. It accepts text, image, video, and...
Small omni GPT for cheap multimodal assistance and production-scale traffic
GPT-4o mini is OpenAI's newest model after [GPT-4 Omni](/models/openai/gpt-4o), supporting both text and image inputs with text outputs. As their most advanced small model, it is many multiples more affordable...
For tasks that demand low latency, GPT‑4.1 nano is the fastest and cheapest model in the GPT-4.1 series. It delivers exceptional performance at a small size with its 1 million...
Tiny GPT-4.1 option for classification, routing, and very high-volume tasks
Largest open Gemma 3 instruction model for multilingual text generation and visual understanding
A balanced model in the Ministral 3 family, Ministral 3 8B is a powerful, efficient tiny language model with vision capabilities.
A balanced model in the Ministral 3 family, Ministral 3 8B is a powerful, efficient tiny language model with vision capabilities.
Llama 4 Scout 17B Instruct (16E) is a mixture-of-experts (MoE) language model developed by Meta, activating 17 billion parameters out of a total of 109B. It supports native multimodal input...
Open multimodal Gemma instruction model for multilingual text generation and image understanding
The smallest model in the Ministral 3 family, Ministral 3 3B is a powerful, efficient tiny language model with vision capabilities.
Open multimodal Gemma instruction model for efficient text generation and image understanding
| Model | Creator | Score | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|---|
| Google: Gemini 2.5 Pro (batch)google/gemini-2.5-pro:batch | 33.3 | 1.04858M | $0.625 | $5 | — | |||
| Gemini 2.5 Progoogle/gemini-2.5-pro | 33.3 | 1.04858M | $1.25 | $10 | 2025-06-17 | |||
| Qwen: Qwen3.5-9B (batch)qwen/qwen3.5-9b:batch | 28.7 | 262.144K | $0.17 | $0.25 | — | |||
| Qwen: Qwen3.5-9Bqwen/qwen3.5-9b | 28.7 | 262.144K | $0.1 | $0.15 | — | |||
| Mistral: Mistral Small 4 (batch)mistralai/mistral-small-2603:batch | 26.6 | 262.144K | $0.075 | $0.3 | — | |||
| Mistral: Mistral Small 4mistralai/mistral-small-2603 | 26.6 | 262.144K | $0.15 | $0.6 | — | |||
| GPT-4o (2024-05-13)openai/gpt-4o-2024-05-13 | 24.2 | 128K | $5 | $15 | 2024-05-13 | |||
| Amazon: Nova 2 Liteamazon/nova-2-lite-v1 | 23.0 | 1M | $0.3 | $2.5 | — | |||
| GPT-4 Turboopenai/gpt-4-turbo | 21.5 | 128K | $10 | $30 | 2023-11-06 | |||
| OpenAI: GPT-4 Turbo (batch)openai/gpt-4-turbo:batch | 21.5 | 128K | $5 | $15 | — | |||
| Mistral: Mistral Medium 3.1mistralai/mistral-medium-3.1 | 20.5 | 131.072K | $0.4 | $2 | — | |||
| Mistral: Mistral Medium 3.1 (batch)mistralai/mistral-medium-3.1:batch | 20.5 | 131.072K | $0.2 | $1 | — | |||
| OpenAI: GPT-4.1 Mini (batch)openai/gpt-4.1-mini:batch | 20.2 | 1.04758M | $0.2 | $0.8 | — | |||
| GPT-4.1 miniopenai/gpt-4.1-mini | 20.2 | 1.04758M | $0.4 | $1.6 | 2025-04-14 | |||
| Mistral: Mistral Large 3 2512mistralai/mistral-large-2512 | 20.1 | 262.144K | $0.5 | $1.5 | — | |||
| Mistral: Mistral Large 3 2512 (batch)mistralai/mistral-large-2512:batch | 20.1 | 262.144K | $0.25 | $0.75 | — | |||
| Meta: Llama 4 Maverickmeta-llama/llama-4-maverick | 16.3 | 128K | $0.2 | $0.696 | — | |||
| GPT-5 Miniopenai/gpt-5-mini | 15.6 | 400K | $0.25 | $2 | 2025-08-07 | |||
| OpenAI: GPT-5 Mini (batch)openai/gpt-5-mini:batch | 15.6 | 400K | $0.125 | $1 | — | |||
| Mistral: Ministral 3 14B 2512mistralai/ministral-14b-2512 | 14.4 | 262.144K | $0.2 | $0.2 | — | |||
| NVIDIA: Nemotron 3 Nano Omni (free)nvidia/nemotron-3-nano-omni-30b-a3b-reasoning:free | 13.8 | 256K | Free | Free | — | |||
| GPT-4o miniopenai/gpt-4o-mini | 11.4 | 128K | $0.15 | $0.6 | 2024-07-18 | |||
| OpenAI: GPT-4o-mini (batch)openai/gpt-4o-mini:batch | 11.4 | 128K | $0.075 | $0.3 | — | |||
| OpenAI: GPT-4.1 Nano (batch)openai/gpt-4.1-nano:batch | 11.1 | 1.04758M | $0.05 | $0.2 | — | |||
| GPT-4.1 nanoopenai/gpt-4.1-nano | 11.1 | 1.04758M | $0.1 | $0.4 | 2025-04-14 | |||
| Gemma 3 27B ITgoogle/gemma-3-27b-it | 10.1 | 131.072K | $0.08 | $0.16 | 2025-03-12 | |||
| Mistral: Ministral 3 8B 2512mistralai/ministral-8b-2512 | 9.7 | 262.144K | $0.15 | $0.15 | — | |||
| Mistral: Ministral 3 8B 2512 (batch)mistralai/ministral-8b-2512:batch | 9.7 | 262.144K | $0.075 | $0.075 | — | |||
| Meta: Llama 4 Scoutmeta-llama/llama-4-scout | 8.2 | 327.68K | $0.1 | $0.3 | — | |||
| Gemma 3 12B ITgoogle/gemma-3-12b-it | 5.8 | 131.072K | $0.05 | $0.1 | 2025-03-12 | |||
| Mistral: Ministral 3 3B 2512mistralai/ministral-3b-2512 | 4.8 | 131.072K | $0.1 | $0.1 | — | |||
| Gemma 3 4B ITgoogle/gemma-3-4b-it | 2.7 | 131.072K | $0.04 | $0.08 | 2025-03-12 |