Popular open Llama workhorse for multilingual chat, coding, and self-hosting
Model catalog
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
25 providers
O-series reasoning model for hard analysis, math, coding, and planning
16 providers
Cohere retrieval model for long-context chat and enterprise RAG workflows
7 providers
Cohere's RAG workhorse for long-context enterprise search and tool use
9 providers
Small omni GPT for cheap multimodal assistance and production-scale traffic
23 providers
Efficient Mistral-NVIDIA open model for multilingual chat and local deployment
11 providers
Omni-era GPT for multimodal chat, practical coding, and general assistants
23 providers
Compact GPT model for low-latency assistance and high-volume workloads
15 providers
GPT model for general reasoning, writing, coding, and tool-assisted tasks
11 providers
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| Llama-3.3-70B-Instructmeta/llama-3.3-70b-instruct | 128K | $0.1 | $0.32 | 2024-12-06 | |||
| o1openai/o1 | 200K | $15 | $60 | 2024-12-05 | |||
| Command Rcohere/command-r-08-2024 | 128K | $0.15 | $0.6 | 2024-08-30 | |||
| Command R+cohere/command-r-plus-08-2024 | 128K | $2.5 | $10 | 2024-08-30 | |||
| GPT-4o miniopenai/gpt-4o-mini | 128K | $0.15 | $0.6 | 2024-07-18 | |||
| Mistral Nemomistral/mistral-nemo | 128K | $0.15 | $0.15 | 2024-07-01 | |||
| GPT-4oopenai/gpt-4o | 128K | $2.5 | $10 | 2024-05-13 | |||
| GPT-4 Turboopenai/gpt-4-turbo | 128K | $10 | $30 | 2023-11-06 | |||
| GPT-4openai/gpt-4 | 8.192K | $30 | $60 | 2023-11-06 |