Open multimodal Llama model for image understanding, captioning, and visual QA
3 providers
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
Open multimodal Llama model for image understanding, captioning, and visual QA
Compact open Llama model for lightweight chat, drafting, and self-hosting
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| Llama-3.2-11B-Vision-Instructmeta/llama-3.2-11b-vision-instruct | 128K | $0.055 | $0.055 | 2024-09-25 | |||
| Llama-3.1-8B-Instructmeta/llama-3.1-8b-instruct | 128K | $0.02 | $0.04 | 2024-07-23 |