Open multilingual vision model for OCR, visual reasoning, and image question answering
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
Compact open multilingual vision model for OCR and visual question answering
Open Command R model optimized for Arabic enterprise chat, RAG, and cultural knowledge
DeepSeek R1 is here: Performance on par with [OpenAI o1](/openai/o1), but open-sourced and with fully open reasoning tokens. It's 671B parameters in size, with 37B active in an inference pass....
R1 reasoning distilled into Qwen 2.5 32B for efficient open-weight step-by-step problem solving
Open DeepSeek MoE chat model for coding, math, and general reasoning
Compact Microsoft instruction model tuned for efficient coding assistance, reasoning, and low-latency agent tasks
Popular open Llama workhorse for multilingual chat, coding, and self-hosting
Command R7B (12-2024) is a small, fast update of the Command R+ model, delivered in December 2024. It excels at RAG, tool use, agents, and similar tasks requiring complex reasoning...
Flagship Mistral model for advanced reasoning, coding, and multilingual work
Open coding-focused Qwen model for code generation, repair, and repository reasoning
Tiny open Qwen code model for lightweight completion and on-device coding
Mistral's larger vision model for document-heavy image understanding and chat
Flagship Mistral model for advanced reasoning, coding, and multilingual work
Mistral's largest general model for enterprise agents, coding, and multilingual reasoning
Open multilingual model optimized for generation across 23 languages
Efficient open Mistral edge model for on-device chat and function calling
Compact Mistral model for edge, latency-sensitive, and cost-efficient workloads
Open multimodal Llama model for image understanding, captioning, and visual QA
Small open Llama base model for lightweight text generation and self-hosting
Compact open Llama base model for lightweight and on-device use
Mistral vision-language model for image understanding and multimodal chat
Qwen vision-language model for visual reasoning, documents, and agent tasks
command-r-plus-08-2024 is an update of the [Command R+](/models/cohere/command-r-plus) with roughly 50% higher throughput and 25% lower latencies as compared to the previous Command R+ version, while keeping the hardware footprint...
command-r-08-2024 is an update of the [Command R](/models/cohere/command-r) with improved performance for multilingual retrieval-augmented generation (RAG) and tool use. More broadly, it is better at math, code and reasoning and...
Compact Nemotron model for efficient reasoning and deployable AI agents
Llama 3.1-based safety classifier for moderating prompts and model responses
Open Llama instruction model for multilingual chat, reasoning, and coding
Compact open Llama model for lightweight chat, drafting, and self-hosting
Efficient Mistral-NVIDIA open model for multilingual chat and local deployment
Open Mistral code model for fill-in-the-middle and 80+ programming languages
Mistral code model for completions, refactors, and developer IDE workflows
[Microsoft Research](/microsoft) Phi-4 is designed to perform well in complex reasoning tasks and can operate efficiently in situations with limited memory or where quick responses are needed. At 14 billion...
Granite 4.2 8B is a dense reasoning model from IBM. It is suited for mathematics, code generation, multilingual dialogue, and agentic workflows that need multi-step reasoning. It supports full, low-effort,...
Granite 4.1 8B is a dense, decoder-only 8-billion-parameter language model from IBM, part of the Granite 4.1 family. It supports a 131K-token context window and is designed for enterprise tasks...
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| Aya Vision 32Bcohere/c4ai-aya-vision-32b | 16K | — | — | 2025-03-04 | |||
| Aya Vision 8Bcohere/c4ai-aya-vision-8b | 16K | — | — | 2025-03-04 | |||
| Command R7B Arabiccohere/command-r7b-arabic-02-2025 | 128K | $0.037 | $0.15 | 2025-02-27 | |||
| DeepSeek: R1deepseek/deepseek-r1 | 64K | $0.7 | $2.5 | 2025-01-20 | |||
| DeepSeek-R1-Distill-Qwen-32Bdeepseek/deepseek-r1-distill-qwen-32b | 131.072K | $0.3 | $0.3 | 2025-01-20 | |||
| DeepSeek-V3deepseek/deepseek-v3 | 131.072K | $0.27 | $1.12 | 2024-12-26 | |||
| Phi-4-minimicrosoft/phi-4-mini | 128K | $0.075 | $0.3 | 2024-12-11 | |||
| Llama-3.3-70B-Instructmeta/llama-3.3-70b-instruct | 128K | $0.1 | $0.32 | 2024-12-06 | |||
| Cohere: Command R7B (12-2024)cohere/command-r7b-12-2024 | 128K | $0.037 | $0.15 | 2024-12-02 | |||
| Mistral Large 2.1mistral/mistral-large-2411 | 131.072K | $2 | $6 | 2024-11-18 | |||
| Qwen2.5-Coder-32B-Instructalibaba/qwen2.5-coder-32b-instruct | 131.072K | $0.06 | $0.2 | 2024-11-12 | |||
| Qwen2.5-Coder-0.5Balibaba/qwen2.5-coder-0.5b | 32.768K | $0.1 | $0.1 | 2024-11-12 | |||
| Pixtral Large (latest)mistral/pixtral-large-latest | 128K | $2 | $6 | 2024-11-01 | |||
| Mistral Large (latest)mistral/mistral-large-latest | 262.144K | $0.5 | $1.5 | 2024-11-01 | |||
| Mistral Large 3mistral/mistral-large-2512 | 262.144K | $0.5 | $1.5 | 2024-11-01 | |||
| Aya Expanse 32Bcohere/c4ai-aya-expanse-32b | 128K | — | — | 2024-10-24 | |||
| Ministral 8B Instructmistral/ministral-8b-instruct-2410 | 131.072K | $0.15 | $0.15 | 2024-10-16 | |||
| Ministral 3Bmistral/ministral-3b | 128K | $0.04 | $0.04 | 2024-10-16 | |||
| Llama-3.2-11B-Vision-Instructmeta/llama-3.2-11b-vision-instruct | 128K | $0.055 | $0.055 | 2024-09-25 | |||
| Llama-3.2-3Bmeta/llama-3.2-3b | 131.072K | $0.1 | $0.1 | 2024-09-25 | |||
| Llama-3.2-1Bmeta/llama-3.2-1b | 131.072K | $0.1 | $0.1 | 2024-09-25 | |||
| Pixtral 12Bmistral/pixtral-12b | 128K | $0.15 | $0.15 | 2024-09-01 | |||
| Qwen2.5-VL 72B Instructalibaba/qwen2-5-vl-72b-instruct | 131.072K | $2.8 | $8.4 | 2024-09 | |||
| Cohere: Command R+ (08-2024)cohere/command-r-plus-08-2024 | 128K | $2.5 | $10 | 2024-08-30 | |||
| Cohere: Command R (08-2024)cohere/command-r-08-2024 | 128K | $0.15 | $0.6 | 2024-08-30 | |||
| Nemotron Mini 4B Instructnvidia/nemotron-mini-4b-instruct | 128K | — | — | 2024-08-21 | |||
| Llama-Guard-3-8Bmeta/llama-guard-3-8b | 128K | — | — | 2024-07-23 | |||
| Llama-3.1-70B-Instructmeta/llama-3.1-70b-instruct | 128K | $0.4 | $0.4 | 2024-07-23 | |||
| Llama-3.1-8B-Instructmeta/llama-3.1-8b-instruct | 128K | $0.02 | $0.04 | 2024-07-23 | |||
| Mistral Nemomistral/mistral-nemo | 128K | $0.15 | $0.15 | 2024-07-01 | |||
| Codestral-22B-v0.1mistral/codestral-22b-v0.1 | 32.768K | $0.3 | $0.9 | 2024-05-29 | |||
| Codestral (latest)mistral/codestral-latest | 256K | $0.3 | $0.9 | 2024-05-29 | |||
| Microsoft: Phi 4microsoft/phi-4 | 16.384K | $0.07 | $0.14 | — | |||
| IBM: Granite 4.2 8Bibm-granite/granite-4.2-8b | 131.072K | $0.06 | $0.25 | — | |||
| IBM: Granite 4.1 8Bibm-granite/granite-4.1-8b | 131.072K | $0.05 | $0.1 | — |