Open weights

Flagship Mistral model for advanced reasoning, coding, and multilingual work

mistral/mistral-large-latest 2024-11-01 262.144K context $0.5/M input $1.5/M output
7 providers
Open weights

Mistral's largest general model for enterprise agents, coding, and multilingual reasoning

mistral/mistral-large-2512 2024-11-01 262.144K context $0.5/M input $1.5/M output
12 providers
Tools JSON Open weights

Open multimodal Llama model for image understanding, captioning, and visual QA

meta/llama-3.2-11b-vision-instruct 2024-09-25 128K context $0.055/M input $0.055/M output
3 providers
Open weights

Cohere's RAG workhorse for long-context enterprise search and tool use

cohere/command-r-plus-08-2024 2024-08-30 128K context $2.5/M input $10/M output
9 providers
Tools JSON Open weights

Cohere retrieval model for long-context chat and enterprise RAG workflows

cohere/command-r-08-2024 2024-08-30 128K context $0.15/M input $0.6/M output
7 providers
Tools

GPT model for general reasoning, writing, coding, and tool-assisted tasks

openai/gpt-4o-2024-08-06 2024-08-06 128K context $2.5/M input $10/M output
7 providers
Tools JSON Open weights

Open Llama instruction model for multilingual chat, reasoning, and coding

meta/llama-3.1-70b-instruct 2024-07-23 128K context $0.4/M input $0.4/M output
4 providers
Tools JSON Open weights

Compact open Llama model for lightweight chat, drafting, and self-hosting

meta/llama-3.1-8b-instruct 2024-07-23 128K context $0.02/M input $0.04/M output
7 providers

Small omni GPT for cheap multimodal assistance and production-scale traffic

openai/gpt-4o-mini 2024-07-18 128K context $0.15/M input $0.6/M output
23 providers
Open weights

Efficient Mistral-NVIDIA open model for multilingual chat and local deployment

mistral/mistral-nemo 2024-07-01 128K context $0.15/M input $0.15/M output
11 providers
Reasoning Tools

Research model for long-horizon investigation, synthesis, and analytical reports

openai/o4-mini-deep-research 2024-06-26 200K context $1.8/M input $7.2/M output
5 providers
Reasoning Tools

Research model for long-horizon investigation, synthesis, and analytical reports

openai/o3-deep-research 2024-06-26 200K context $9/M input $36/M output
6 providers
Tools Open weights

Mistral code model for completions, refactors, and developer IDE workflows

mistral/codestral-latest 2024-05-29 256K context $0.3/M input $0.9/M output
5 providers
Tools

Omni-era GPT for multimodal chat, practical coding, and general assistants

openai/gpt-4o 2024-05-13 128K context $2.5/M input $10/M output
23 providers
Tools JSON

GPT model for general reasoning, writing, coding, and tool-assisted tasks

openai/gpt-4o-2024-05-13 2024-05-13 128K context $5/M input $15/M output
5 providers
Tools

Qwen vision-language model for visual reasoning, documents, and agent tasks

alibaba/qwen-vl-max 2024-04-08 131.072K context $0.8/M input $3.2/M output
5 providers
Tools

Flagship Qwen model for complex reasoning, coding, and agentic workflows

alibaba/qwen-max 2024-04-03 32.768K context $1.6/M input $6.4/M output
6 providers
Reasoning

Qwen instruction model for multilingual chat, reasoning, and tool use

alibaba/qwen-plus 2024-01-25 1M context $0.4/M input $1.2/M output
10 providers
JSON

Deeper Sonar search model with broader retrieval and stronger synthesis

perplexity/sonar-pro 2024-01-01 200K context $3/M input $15/M output
9 providers
JSON

Fast web-grounded Sonar for current answers, citations, and lightweight retrieval

perplexity/sonar 2024-01-01 128K context $1/M input $1/M output
9 providers
JSON

Web-grounded Sonar for multi-step research questions that need cited reasoning

perplexity/sonar-reasoning-pro 2024-01-01 128K context $2/M input $8/M output
9 providers
Tools

GPT model for general reasoning, writing, coding, and tool-assisted tasks

openai/gpt-4 2023-11-06 8.192K context $30/M input $60/M output
11 providers

Compact GPT model for low-latency assistance and high-volume workloads

openai/gpt-4-turbo 2023-11-06 128K context $10/M input $30/M output
15 providers

Compact GPT model for low-latency assistance and high-volume workloads

openai/gpt-3.5-turbo 2023-03-01 16.385K context $0.5/M input $1.5/M output
13 providers