Smaller o-series reasoner for economical coding, math, and planning tasks
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
Low-latency Gemini model for high-volume multimodal and agent workloads
Earlier Gemini Flash workhorse for responsive multimodal apps and tool use
O-series reasoning model for hard analysis, math, coding, and planning
Efficient model for low-latency assistance, extraction, and routine automation
Flagship model for demanding analysis, coding, and production agent workflows
Efficient model for low-latency assistance, extraction, and routine automation
GPT model for general reasoning, writing, coding, and tool-assisted tasks
Efficient Qwen model for fast chat, extraction, and high-volume workloads
Balanced Claude model for coding, analysis, agent workflows, and cost control
Fast Claude model for responsive assistance, classification, and lightweight agents
GPT model for general reasoning, writing, coding, and tool-assisted tasks
Small omni GPT for cheap multimodal assistance and production-scale traffic
Research model for long-horizon investigation, synthesis, and analytical reports
Research model for long-horizon investigation, synthesis, and analytical reports
GPT model for general reasoning, writing, coding, and tool-assisted tasks
Omni-era GPT for multimodal chat, practical coding, and general assistants
Qwen vision-language model for visual reasoning, documents, and agent tasks
Legacy model retained for compatibility with older integrations
Qwen instruction model for multilingual chat, reasoning, and tool use
Qwen vision-language model for visual reasoning, documents, and agent tasks
Fast web-grounded Sonar for current answers, citations, and lightweight retrieval
Web-grounded Sonar for multi-step research questions that need cited reasoning
Deeper Sonar search model with broader retrieval and stronger synthesis
Compact GPT model for low-latency assistance and high-volume workloads
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| o3-miniopenai/o3-mini | 200K | $1.1 | $4.4 | 2024-12-20 | |||
| Gemini 2.0 Flash-Litegoogle/gemini-2.0-flash-lite | 1.04858M | $0.052 | $0.21 | 2024-12-11 | |||
| Gemini 2.0 Flashgoogle/gemini-2.0-flash | 1.04858M | $0.1 | $0.42 | 2024-12-11 | |||
| o1openai/o1 | 200K | $15 | $60 | 2024-12-05 | |||
| Nova Microamazon/nova-micro | 128K | $0.035 | $0.14 | 2024-12-03 | |||
| Nova Proamazon/nova-pro | 300K | $0.8 | $3.2 | 2024-12-03 | |||
| Nova Liteamazon/nova-lite | 300K | $0.06 | $0.24 | 2024-12-03 | |||
| GPT-4o (2024-11-20)openai/gpt-4o-2024-11-20 | 128K | $2.5 | $10 | 2024-11-20 | |||
| Qwen Turboalibaba/qwen-turbo | 1M | $0.05 | $0.2 | 2024-11-01 | |||
| Claude Sonnet 3.5 v2anthropic/claude-3-5-sonnet-20241022 | 200K | — | — | 2024-10-22 | |||
| Claude Haiku 3.5anthropic/claude-3-5-haiku-20241022 | 200K | $0.8 | $4 | 2024-10-22 | |||
| GPT-4o (2024-08-06)openai/gpt-4o-2024-08-06 | 128K | $2.5 | $10 | 2024-08-06 | |||
| GPT-4o miniopenai/gpt-4o-mini | 128K | $0.15 | $0.6 | 2024-07-18 | |||
| o4-mini-deep-researchopenai/o4-mini-deep-research | 200K | $1.8 | $7.2 | 2024-06-26 | |||
| o3-deep-researchopenai/o3-deep-research | 200K | $9 | $36 | 2024-06-26 | |||
| GPT-4o (2024-05-13)openai/gpt-4o-2024-05-13 | 128K | $5 | $15 | 2024-05-13 | |||
| GPT-4oopenai/gpt-4o | 128K | $2.5 | $10 | 2024-05-13 | |||
| Qwen-VL Maxalibaba/qwen-vl-max | 131.072K | $0.8 | $3.2 | 2024-04-08 | |||
| Claude Haiku 3anthropic/claude-3-haiku-20240307 | 200K | $0.25 | $1.25 | 2024-03-13 | |||
| Qwen Plusalibaba/qwen-plus | 1M | $0.4 | $1.2 | 2024-01-25 | |||
| Qwen-VL Plusalibaba/qwen-vl-plus | 131.072K | $0.21 | $0.63 | 2024-01-25 | |||
| Sonarperplexity/sonar | 128K | $1 | $1 | 2024-01-01 | |||
| Sonar Reasoning Properplexity/sonar-reasoning-pro | 128K | $2 | $8 | 2024-01-01 | |||
| Sonar Properplexity/sonar-pro | 200K | $3 | $15 | 2024-01-01 | |||
| GPT-4 Turboopenai/gpt-4-turbo | 128K | $10 | $30 | 2023-11-06 | |||
| meta-llama/llama-4-scout-17b-16e-instructgroq/meta-llama/llama-4-scout-17b-16e-instruct | 131.072K | $0.11 | $0.34 | — | |||
| meta-llama/llama-4-maverick-17b-128e-instructgroq/meta-llama/llama-4-maverick-17b-128e-instruct | 131.072K | $0.2 | $0.6 | — | |||
| anthropic/claude-opus-4.5gmi/anthropic/claude-opus-4.5 | 409.6K | $5 | $25 | — | |||
| llama-3.3-70b-versatilegroq/llama-3.3-70b-versatile | 128K | $0.59 | $0.79 | — | |||
| llama-3.1-8b-instantgroq/llama-3.1-8b-instant | 128K | $0.05 | $0.08 | — | |||
| GigaChat-2-Progigachat/gigachat-2-pro | 128K | — | — | — | |||
| gemini-2.5-computer-use-preview-10-2025gemini/gemini-2.5-computer-use-preview-10-2025 | 128K | $1.25 | $10 | — | |||
| nova-pro-v1amazon_nova/nova-pro-v1 | 300K | $0.8 | $3.2 | — | |||
| nova-premier-v1amazon_nova/nova-premier-v1 | 1M | $2.5 | $12.5 | — | |||
| GigaChat-2-Maxgigachat/gigachat-2-max | 128K | — | — | — | |||
| nova-lite-v1amazon_nova/nova-lite-v1 | 300K | $0.06 | $0.24 | — | |||
| nova-micro-v1amazon_nova/nova-micro-v1 | 128K | $0.035 | $0.14 | — | |||
| GigaChat-2-Litegigachat/gigachat-2-lite | 128K | — | — | — | |||
| gemini-2.5-progemini/gemini-2.5-pro | 1.04858M | $1.25 | $10 | — | |||
| gemini-3.1-pro-preview-customtoolsvertex_ai/gemini-3.1-pro-preview-customtools | 1.04858M | $2 | $12 | — | |||
| Qwen3-4B-Instruct-2507-GGUFlemonade/qwen3-4b-instruct-2507-gguf | 262.144K | — | — | — | |||
| Gemma-3-4b-it-GGUFlemonade/gemma-3-4b-it-gguf | 128K | — | — | — | |||
| mai-code-1-flash-internalgithub_copilot/mai-code-1-flash-internal | 128K | $0.75 | $4.5 | — | |||
| gpt-oss-120b-mxfp-GGUFlemonade/gpt-oss-120b-mxfp-gguf | 131.072K | — | — | — | |||
| gpt-oss-20b-mxfp4-GGUFlemonade/gpt-oss-20b-mxfp4-gguf | 131.072K | — | — | — | |||
| gemini-3-pro-previewgithub_copilot/gemini-3-pro-preview | 128K | — | — | — | |||
| gemini-2.5-flash-lite-preview-06-17gemini/gemini-2.5-flash-lite-preview-06-17 | 1.04858M | $0.1 | $0.4 | — | |||
| Qwen3-Coder-30B-A3B-Instruct-GGUFlemonade/qwen3-coder-30b-a3b-instruct-gguf | 262.144K | — | — | — | |||
| openai-o3-minigradient_ai/openai-o3-mini | 200K | $1.1 | $4.4 | — | |||
| gemini-2.5-progithub_copilot/gemini-2.5-pro | 128K | — | — | — |