Gemini 3.1 Flash Lite Preview is Google's high-efficiency model optimized for high-volume use cases. It outperforms Gemini 2.5 Flash Lite on overall quality and approaches Gemini 2.5 Flash performance across...
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
Gemma 4 31B Instruct is Google DeepMind's 30.7B dense multimodal model supporting text and image input with text output. Features a 256K token context window, configurable thinking/reasoning mode, native function...
Gemma 4 31B Instruct is Google DeepMind's 30.7B dense multimodal model supporting text and image input with text output. Features a 256K token context window, configurable thinking/reasoning mode, native function...
Gemma 4 31B Instruct is Google DeepMind's 30.7B dense multimodal model supporting text and image input with text output. Features a 256K token context window, configurable thinking/reasoning mode, native function...
Mistral Medium 3.5 is a dense 128B instruction-following model from Mistral AI. It supports text and image inputs with text output, and is designed for agentic workflows, coding, and complex...
Mistral Medium 3.5 is a dense 128B instruction-following model from Mistral AI. It supports text and image inputs with text output, and is designed for agentic workflows, coding, and complex...
GPT-4.1 Mini is a mid-sized model delivering performance competitive with GPT-4o at substantially lower latency and cost. It retains a 1 million token context window and scores 45.1% on hard...
GPT-4.1 Mini is a mid-sized model delivering performance competitive with GPT-4o at substantially lower latency and cost. It retains a 1 million token context window and scores 45.1% on hard...
Mistral Medium 3.1 is an updated version of Mistral Medium 3, which is a high-performance enterprise-grade language model designed to deliver frontier-level capabilities at significantly reduced operational cost. It balances...
Mistral Medium 3.1 is an updated version of Mistral Medium 3, which is a high-performance enterprise-grade language model designed to deliver frontier-level capabilities at significantly reduced operational cost. It balances...
Mistral Small 4 is the next major release in the Mistral Small family, unifying the capabilities of several flagship Mistral models into a single system. It combines strong reasoning from...
Mistral Small 4 is the next major release in the Mistral Small family, unifying the capabilities of several flagship Mistral models into a single system. It combines strong reasoning from...
Mistral Large 3 2512 is Mistral’s most capable model to date, featuring a sparse mixture-of-experts architecture with 41B active parameters (675B total), and released under the Apache 2.0 license.
Mistral Large 3 2512 is Mistral’s most capable model to date, featuring a sparse mixture-of-experts architecture with 41B active parameters (675B total), and released under the Apache 2.0 license.
For tasks that demand low latency, GPT‑4.1 nano is the fastest and cheapest model in the GPT-4.1 series. It delivers exceptional performance at a small size with its 1 million...
For tasks that demand low latency, GPT‑4.1 nano is the fastest and cheapest model in the GPT-4.1 series. It delivers exceptional performance at a small size with its 1 million...
Llama 4 Maverick 17B Instruct (128E) is a high-capacity multimodal language model from Meta, built on a mixture-of-experts (MoE) architecture with 128 experts and 17 billion active parameters per forward...
Llama 4 Scout 17B Instruct (16E) is a mixture-of-experts (MoE) language model developed by Meta, activating 17 billion parameters out of a total of 109B. It supports native multimodal input...
The largest model in the Ministral 3 family, Ministral 3 14B offers frontier capabilities and performance comparable to its larger Mistral Small 3.2 24B counterpart. A powerful and efficient language...
A balanced model in the Ministral 3 family, Ministral 3 8B is a powerful, efficient tiny language model with vision capabilities.
A balanced model in the Ministral 3 family, Ministral 3 8B is a powerful, efficient tiny language model with vision capabilities.
Gemma 3 introduces multimodality, supporting vision-language input and text outputs. It handles context windows up to 128k tokens, understands over 140 languages, and offers improved math, reasoning, and chat capabilities,...
The smallest model in the Ministral 3 family, Ministral 3 3B is a powerful, efficient tiny language model with vision capabilities.
Gemma 3 introduces multimodality, supporting vision-language input and text outputs. It handles context windows up to 128k tokens, understands over 140 languages, and offers improved math, reasoning, and chat capabilities,...
| Model | Creator | Score | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|---|
| Google: Gemini 3.1 Flash Lite Previewgoogle/gemini-3.1-flash-lite-preview | 16.0 | 1.04858M | $0.25 | $1.5 | 2026-03-03 | |||
| Google: Gemma 4 31B (batch)google/gemma-4-31b-it:batch | 15.4 | 262.144K | $0.39 | $0.97 | — | |||
| Google: Gemma 4 31B (free)google/gemma-4-31b-it:free | 15.4 | 262.144K | Free | Free | — | |||
| Google: Gemma 4 31Bgoogle/gemma-4-31b-it | 15.4 | 262.144K | $0.09 | $0.34 | 2026-04-02 | |||
| Mistral: Mistral Medium 3.5mistralai/mistral-medium-3-5 | 14.9 | 262.144K | $1.5 | $7.5 | — | |||
| Mistral: Mistral Medium 3.5 (batch)mistralai/mistral-medium-3-5:batch | 14.9 | 262.144K | $0.75 | $3.75 | — | |||
| OpenAI: GPT-4.1 Mini (batch)openai/gpt-4.1-mini:batch | 14.8 | 1.04758M | $0.2 | $0.8 | — | |||
| OpenAI: GPT-4.1 Miniopenai/gpt-4.1-mini | 14.8 | 1.04758M | $0.4 | $1.6 | 2025-04-14 | |||
| Mistral: Mistral Medium 3.1mistralai/mistral-medium-3.1 | 14.7 | 131.072K | $0.4 | $2 | — | |||
| Mistral: Mistral Medium 3.1 (batch)mistralai/mistral-medium-3.1:batch | 14.7 | 131.072K | $0.2 | $1 | — | |||
| Mistral: Mistral Small 4 (batch)mistralai/mistral-small-2603:batch | 11.5 | 262.144K | $0.075 | $0.3 | — | |||
| Mistral: Mistral Small 4mistralai/mistral-small-2603 | 11.5 | 262.144K | $0.15 | $0.6 | — | |||
| Mistral: Mistral Large 3 2512 (batch)mistralai/mistral-large-2512:batch | 9.7 | 262.144K | $0.25 | $0.75 | — | |||
| Mistral: Mistral Large 3 2512mistralai/mistral-large-2512 | 9.7 | 262.144K | $0.5 | $1.5 | — | |||
| OpenAI: GPT-4.1 Nanoopenai/gpt-4.1-nano | 9.6 | 1.04758M | $0.1 | $0.4 | 2025-04-14 | |||
| OpenAI: GPT-4.1 Nano (batch)openai/gpt-4.1-nano:batch | 9.6 | 1.04758M | $0.05 | $0.2 | — | |||
| Meta: Llama 4 Maverickmeta-llama/llama-4-maverick | 9.3 | 128K | $0.2 | $0.696 | — | |||
| Meta: Llama 4 Scoutmeta-llama/llama-4-scout | 6.5 | 327.68K | $0.1 | $0.3 | — | |||
| Mistral: Ministral 3 14B 2512mistralai/ministral-14b-2512 | 6.0 | 262.144K | $0.2 | $0.2 | — | |||
| Mistral: Ministral 3 8B 2512mistralai/ministral-8b-2512 | 5.5 | 262.144K | $0.15 | $0.15 | — | |||
| Mistral: Ministral 3 8B 2512 (batch)mistralai/ministral-8b-2512:batch | 5.5 | 262.144K | $0.075 | $0.075 | — | |||
| Google: Gemma 3 27Bgoogle/gemma-3-27b-it | 4.9 | 131.072K | $0.08 | $0.45 | 2025-03-12 | |||
| Mistral: Ministral 3 3B 2512mistralai/ministral-3b-2512 | 4.8 | 131.072K | $0.1 | $0.1 | — | |||
| Google: Gemma 3 12Bgoogle/gemma-3-12b-it | 3.8 | 131.072K | $0.05 | $0.15 | 2025-03-12 |