Open multimodal Gemma instruction model for multilingual text generation and image understanding
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
Meta's latest class of model (Llama 3.1) launched with a variety of sizes & flavors. This 8B instruct-tuned version is fast and efficient. It has demonstrated strong performance compared to...
The smallest model in the Ministral 3 family, Ministral 3 3B is a powerful, efficient tiny language model with vision capabilities.
Gemma 3n E4B-it is optimized for efficient execution on mobile and low-resource devices, such as phones, laptops, and tablets. It supports multimodal inputs—including text, visual data, and audio—enabling diverse tasks...
Open multimodal Gemma instruction model for efficient text generation and image understanding
| Model | Creator | Score | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|---|
| Gemma 3 12B ITgoogle/gemma-3-12b-it | 5.8 | 131.072K | $0.05 | $0.1 | 2025-03-12 | |||
| Meta: Llama 3.1 8B Instructmeta-llama/llama-3.1-8b-instruct | 5.4 | 131.072K | $0.05 | $0.08 | — | |||
| Mistral: Ministral 3 3B 2512mistralai/ministral-3b-2512 | 4.8 | 131.072K | $0.1 | $0.1 | — | |||
| Google: Gemma 3n 4Bgoogle/gemma-3n-e4b-it | 3.2 | 32.768K | $0.06 | $0.12 | — | |||
| Gemma 3 4B ITgoogle/gemma-3-4b-it | 2.7 | 131.072K | $0.04 | $0.08 | 2025-03-12 |