Open weights

Open multilingual vision model for OCR, visual reasoning, and image question answering

cohere/c4ai-aya-vision-32b 2025-03-04 16K context Input not listed Output not listed
1 provider
Open weights

Compact open multilingual vision model for OCR and visual question answering

cohere/c4ai-aya-vision-8b 2025-03-04 16K context Input not listed Output not listed
1 provider
Tools Open weights

Open Command R model optimized for Arabic enterprise chat, RAG, and cultural knowledge

cohere/command-r7b-arabic-02-2025 2025-02-27 128K context $0.037/M input $0.15/M output
1 provider
Reasoning Tools Open weights

DeepSeek R1 is here: Performance on par with [OpenAI o1](/openai/o1), but open-sourced and with fully open reasoning tokens. It's 671B parameters in size, with 37B active in an inference pass....

deepseek/deepseek-r1 2025-01-20 64K context $0.7/M input $2.5/M output
14 providers
Open weights

R1 reasoning distilled into Qwen 2.5 32B for efficient open-weight step-by-step problem solving

deepseek/deepseek-r1-distill-qwen-32b 2025-01-20 131.072K context $0.3/M input $0.3/M output
1 provider
Tools Open weights

Open DeepSeek MoE chat model for coding, math, and general reasoning

deepseek/deepseek-v3 2024-12-26 131.072K context $0.27/M input $1.12/M output
8 providers
Tools Open weights

Compact Microsoft instruction model tuned for efficient coding assistance, reasoning, and low-latency agent tasks

microsoft/phi-4-mini 2024-12-11 128K context $0.075/M input $0.3/M output
2 providers
Tools Open weights

Popular open Llama workhorse for multilingual chat, coding, and self-hosting

meta/llama-3.3-70b-instruct 2024-12-06 128K context $0.1/M input $0.32/M output
25 providers
JSON Open weights

Command R7B (12-2024) is a small, fast update of the Command R+ model, delivered in December 2024. It excels at RAG, tool use, agents, and similar tasks requiring complex reasoning...

cohere/command-r7b-12-2024 2024-12-02 128K context $0.037/M input $0.15/M output
5 providers
Tools Open weights

Flagship Mistral model for advanced reasoning, coding, and multilingual work

mistral/mistral-large-2411 2024-11-18 131.072K context $2/M input $6/M output
2 providers
Tools JSON Open weights

Open coding-focused Qwen model for code generation, repair, and repository reasoning

alibaba/qwen2.5-coder-32b-instruct 2024-11-12 131.072K context $0.06/M input $0.2/M output
2 providers
Open weights

Tiny open Qwen code model for lightweight completion and on-device coding

alibaba/qwen2.5-coder-0.5b 2024-11-12 32.768K context $0.1/M input $0.1/M output
1 provider
Tools Open weights

Mistral's larger vision model for document-heavy image understanding and chat

mistral/pixtral-large-latest 2024-11-01 128K context $2/M input $6/M output
5 providers
Open weights

Flagship Mistral model for advanced reasoning, coding, and multilingual work

mistral/mistral-large-latest 2024-11-01 262.144K context $0.5/M input $1.5/M output
7 providers
Open weights

Mistral's largest general model for enterprise agents, coding, and multilingual reasoning

mistral/mistral-large-2512 2024-11-01 262.144K context $0.5/M input $1.5/M output
12 providers
Open weights

Open multilingual model optimized for generation across 23 languages

cohere/c4ai-aya-expanse-32b 2024-10-24 128K context Input not listed Output not listed
1 provider
Tools Open weights

Efficient open Mistral edge model for on-device chat and function calling

mistral/ministral-8b-instruct-2410 2024-10-16 131.072K context $0.15/M input $0.15/M output
1 provider
Open weights

Compact Mistral model for edge, latency-sensitive, and cost-efficient workloads

mistral/ministral-3b 2024-10-16 128K context $0.04/M input $0.04/M output
5 providers
Tools JSON Open weights

Open multimodal Llama model for image understanding, captioning, and visual QA

meta/llama-3.2-11b-vision-instruct 2024-09-25 128K context $0.055/M input $0.055/M output
3 providers
Open weights

Small open Llama base model for lightweight text generation and self-hosting

meta/llama-3.2-3b 2024-09-25 131.072K context $0.1/M input $0.1/M output
1 provider
Open weights

Compact open Llama base model for lightweight and on-device use

meta/llama-3.2-1b 2024-09-25 131.072K context $0.1/M input $0.1/M output
1 provider
Tools Open weights

Mistral vision-language model for image understanding and multimodal chat

mistral/pixtral-12b 2024-09-01 128K context $0.15/M input $0.15/M output
4 providers
Tools Open weights

Qwen vision-language model for visual reasoning, documents, and agent tasks

alibaba/qwen2-5-vl-72b-instruct 2024-09 131.072K context $2.8/M input $8.4/M output
3 providers
Open weights

command-r-plus-08-2024 is an update of the [Command R+](/models/cohere/command-r-plus) with roughly 50% higher throughput and 25% lower latencies as compared to the previous Command R+ version, while keeping the hardware footprint...

cohere/command-r-plus-08-2024 2024-08-30 128K context $2.5/M input $10/M output
9 providers
Tools JSON Open weights

command-r-08-2024 is an update of the [Command R](/models/cohere/command-r) with improved performance for multilingual retrieval-augmented generation (RAG) and tool use. More broadly, it is better at math, code and reasoning and...

cohere/command-r-08-2024 2024-08-30 128K context $0.15/M input $0.6/M output
7 providers
Tools Open weights

Compact Nemotron model for efficient reasoning and deployable AI agents

nvidia/nemotron-mini-4b-instruct 2024-08-21 128K context Input not listed Output not listed
1 provider
Open weights

Llama 3.1-based safety classifier for moderating prompts and model responses

meta/llama-guard-3-8b 2024-07-23 128K context Input not listed Output not listed
1 provider
Tools JSON Open weights

Open Llama instruction model for multilingual chat, reasoning, and coding

meta/llama-3.1-70b-instruct 2024-07-23 128K context $0.4/M input $0.4/M output
4 providers
Tools JSON Open weights

Compact open Llama model for lightweight chat, drafting, and self-hosting

meta/llama-3.1-8b-instruct 2024-07-23 128K context $0.02/M input $0.04/M output
7 providers
Open weights

Efficient Mistral-NVIDIA open model for multilingual chat and local deployment

mistral/mistral-nemo 2024-07-01 128K context $0.15/M input $0.15/M output
11 providers
Open weights

Open Mistral code model for fill-in-the-middle and 80+ programming languages

mistral/codestral-22b-v0.1 2024-05-29 32.768K context $0.3/M input $0.9/M output
1 provider
Tools Open weights

Mistral code model for completions, refactors, and developer IDE workflows

mistral/codestral-latest 2024-05-29 256K context $0.3/M input $0.9/M output
5 providers
Open weights

[Microsoft Research](/microsoft) Phi-4 is designed to perform well in complex reasoning tasks and can operate efficiently in situations with limited memory or where quick responses are needed. At 14 billion...

microsoft/phi-4 16.384K context $0.07/M input $0.14/M output
Open weights

Granite 4.2 8B is a dense reasoning model from IBM. It is suited for mathematics, code generation, multilingual dialogue, and agentic workflows that need multi-step reasoning. It supports full, low-effort,...

ibm-granite/granite-4.2-8b 131.072K context $0.06/M input $0.25/M output
Open weights

Granite 4.1 8B is a dense, decoder-only 8-billion-parameter language model from IBM, part of the Granite 4.1 family. It supports a 131K-token context window and is designed for enterprise tasks...

ibm-granite/granite-4.1-8b 131.072K context $0.05/M input $0.1/M output