64 models
Reasoning Tools Open weights

Cohere command model for multilingual enterprise agents, tools, and chat

cohere/command-a-03-2025 2025-03-13 256K context $2.5/M input $10/M output
6 providers
Reasoning Tools Open weights

Classic open reasoning model for transparent math, coding, and deliberate problem solving

deepseek/deepseek-r1 2025-01-20 128K context $0.7/M input $2.5/M output
14 providers
Reasoning Tools JSON

Smaller o-series reasoner for economical coding, math, and planning tasks

openai/o3-mini 2024-12-20 200K context $1.1/M input $4.4/M output
19 providers
Tools Open weights

Compact Microsoft instruction model tuned for efficient coding assistance, reasoning, and low-latency agent tasks

microsoft/phi-4-mini 2024-12-11 128K context $0.075/M input $0.3/M output
2 providers
Tools Open weights

Popular open Llama workhorse for multilingual chat, coding, and self-hosting

meta/llama-3.3-70b-instruct 2024-12-06 128K context $0.1/M input $0.32/M output
25 providers
Reasoning

O-series reasoning model for hard analysis, math, coding, and planning

openai/o1 2024-12-05 200K context $15/M input $60/M output
16 providers
Open weights

Compact Mistral model for edge, latency-sensitive, and cost-efficient workloads

mistral/ministral-3b 2024-10-16 128K context $0.04/M input $0.04/M output
5 providers
Tools JSON Open weights

Cohere retrieval model for long-context chat and enterprise RAG workflows

cohere/command-r-08-2024 2024-08-30 128K context $0.15/M input $0.6/M output
7 providers
Open weights

Cohere's RAG workhorse for long-context enterprise search and tool use

cohere/command-r-plus-08-2024 2024-08-30 128K context $2.5/M input $10/M output
9 providers

Small omni GPT for cheap multimodal assistance and production-scale traffic

openai/gpt-4o-mini 2024-07-18 128K context $0.15/M input $0.6/M output
23 providers
Open weights

Efficient Mistral-NVIDIA open model for multilingual chat and local deployment

mistral/mistral-nemo 2024-07-01 128K context $0.15/M input $0.15/M output
11 providers
Tools

Omni-era GPT for multimodal chat, practical coding, and general assistants

openai/gpt-4o 2024-05-13 128K context $2.5/M input $10/M output
23 providers

Compact GPT model for low-latency assistance and high-volume workloads

openai/gpt-4-turbo 2023-11-06 128K context $10/M input $30/M output
15 providers
Tools

GPT model for general reasoning, writing, coding, and tool-assisted tasks

openai/gpt-4 2023-11-06 8.192K context $30/M input $60/M output
11 providers