Tools

Flagship model for demanding analysis, coding, and production agent workflows

amazon/nova-pro 2024-12-03 300K context $0.8/M input $3.2/M output
3 providers
Tools

GPT model for general reasoning, writing, coding, and tool-assisted tasks

openai/gpt-4o-2024-11-20 2024-11-20 128K context $2.5/M input $10/M output
8 providers
Tools

Fast Claude model for responsive assistance, classification, and lightweight agents

anthropic/claude-3-5-haiku-20241022 2024-10-22 200K context $0.8/M input $4/M output
3 providers

Balanced Claude model for coding, analysis, agent workflows, and cost control

anthropic/claude-3-5-sonnet-20241022 2024-10-22 200K context Input not listed Output not listed
1 provider
Tools

GPT model for general reasoning, writing, coding, and tool-assisted tasks

openai/gpt-4o-2024-08-06 2024-08-06 128K context $2.5/M input $10/M output
7 providers

Small omni GPT for cheap multimodal assistance and production-scale traffic

openai/gpt-4o-mini 2024-07-18 128K context $0.15/M input $0.6/M output
23 providers
Reasoning Tools

Research model for long-horizon investigation, synthesis, and analytical reports

openai/o3-deep-research 2024-06-26 200K context $9/M input $36/M output
6 providers
Reasoning Tools

Research model for long-horizon investigation, synthesis, and analytical reports

openai/o4-mini-deep-research 2024-06-26 200K context $1.8/M input $7.2/M output
5 providers
Tools

Omni-era GPT for multimodal chat, practical coding, and general assistants

openai/gpt-4o 2024-05-13 128K context $2.5/M input $10/M output
23 providers
Tools JSON

GPT model for general reasoning, writing, coding, and tool-assisted tasks

openai/gpt-4o-2024-05-13 2024-05-13 128K context $5/M input $15/M output
5 providers
Tools

Qwen vision-language model for visual reasoning, documents, and agent tasks

alibaba/qwen-vl-max 2024-04-08 131.072K context $0.8/M input $3.2/M output
5 providers
Tools

Legacy model retained for compatibility with older integrations

anthropic/claude-3-haiku-20240307 2024-03-13 200K context $0.25/M input $1.25/M output
2 providers
Tools

Qwen vision-language model for visual reasoning, documents, and agent tasks

alibaba/qwen-vl-plus 2024-01-25 131.072K context $0.21/M input $0.63/M output
4 providers
JSON

Web-grounded Sonar for multi-step research questions that need cited reasoning

perplexity/sonar-reasoning-pro 2024-01-01 128K context $2/M input $8/M output
9 providers
JSON

Deeper Sonar search model with broader retrieval and stronger synthesis

perplexity/sonar-pro 2024-01-01 200K context $3/M input $15/M output
9 providers

Compact GPT model for low-latency assistance and high-volume workloads

openai/gpt-4-turbo 2023-11-06 128K context $10/M input $30/M output
15 providers

Nex-N2-Pro is an agentic mixture-of-experts model from Nex AGI, with 17B active parameters out of 397B total. Built on the Qwen3.5 architecture, it accepts text and image input and produces...

nex-agi/nex-n2-pro 262.144K context $0.25/M input $1/M output

No provider description is available for this model yet.

bedrock_converse/us.xai.grok-4.6 500K context $2.2/M input $6.6/M output

This model always redirects to the latest model in the OpenAI GPT family.

~openai/gpt-latest 1.05M context $2/M input $10/M output

Nova 2 Lite is a fast, cost-effective reasoning model for everyday workloads that can process text, images, and videos to generate text. Nova 2 Lite demonstrates standout capabilities in processing...

amazon/nova-2-lite-v1 1M context $0.3/M input $2.5/M output

No provider description is available for this model yet.

bedrock_mantle/xai.grok-4.6 500K context $2.2/M input $6.6/M output

No provider description is available for this model yet.

openai/chat-latest 400K context $5/M input $30/M output

This model always redirects to the latest model in the Gemini Flash family.

~google/gemini-flash-latest 1.04858M context $0.75/M input $3.75/M output

GPT-5.2 Chat (AKA Instant) is the fast, lightweight member of the 5.2 family, optimized for low-latency chat while retaining strong general intelligence. It uses adaptive reasoning to selectively “think” on...

openai/gpt-5.2-chat 128K context $1.75/M input $14/M output

No provider description is available for this model yet.

openai/daybreak-blue-latest 1.05M context $5/M input $30/M output

No provider description is available for this model yet.

openai/daybreak-red-latest 400K context $12.5/M input $75/M output

This model always redirects to the latest model in the Kimi family.

~moonshotai/kimi-latest 1.04858M context $2.125/M input $11.9/M output

No provider description is available for this model yet.

openai/gpt-5.6-cyber 400K context $12.5/M input $75/M output

The Qwen3.5 native vision-language series Plus models are built on a hybrid architecture that integrates linear attention mechanisms with sparse mixture-of-experts models, achieving higher inference efficiency. In a variety of...

qwen/qwen3.5-plus-02-15 1M context $0.26/M input $1.56/M output

No provider description is available for this model yet.

azure_ai/gpt-6-astra 922K context $10/M input $50/M output

Ox Alpha is a reasoning model designed for coding, sustained agentic work, and production workloads. It is suited for long-horizon software engineering, complex reasoning, and workflows that combine text with...

stealth/ox-alpha 1.04858M context Input not listed Output not listed

GPT-5.4 Pro is OpenAI's most advanced model, building on GPT-5.4's unified architecture with enhanced reasoning capabilities for complex, high-stakes tasks. It features a 1M+ token context window (922K input, 128K...

openai/gpt-5.4-pro:batch 1.05M context $15/M input $90/M output

This model always redirects to the latest model in the Gemini Pro family.

~google/gemini-pro-latest 1.04858M context $2/M input $12/M output

The Qwen3.5 122B-A10B native vision-language model is built on a hybrid architecture that integrates a linear attention mechanism with a sparse mixture-of-experts model, achieving higher inference efficiency. In terms of...

qwen/qwen3.5-122b-a10b 262.144K context $0.26/M input $2.08/M output

Nex-N2-Mini is an open-source agentic mixture-of-experts model from Nex AGI, the smaller sibling in the Nex-N2 series. It accepts text and image input and is built for coding, tool use,...

nex-agi/nex-n2-mini 262.144K context $0.025/M input $0.1/M output

Fast-mode variant of [Opus 5](/anthropic/claude-opus-5) - identical capabilities with higher output speed at 2x pricing relative to regular Opus 5. Learn more in Anthropic's docs: https://platform.claude.com/docs/en/build-with-claude/fast-mode Note: As of September...

anthropic/claude-opus-5-fast 1M context $10/M input $50/M output

GPT-5.4 mini brings the core capabilities of GPT-5.4 to a faster, more efficient model optimized for high-throughput workloads. It supports text and image inputs with strong performance across reasoning, coding,...

openai/gpt-5.4-mini:batch 400K context $0.375/M input $2.25/M output

Grok 4.20 Multi-Agent is a variant of SpaceXAI’s Grok 4.20 designed for collaborative, agent-based workflows. Multiple agents operate in parallel to conduct deep research, coordinate tool use, and synthesize information...

x-ai/grok-4.20-multi-agent 2M context $1.25/M input $2.5/M output

This model always redirects to the latest model in the GPT Mini family.

~openai/gpt-mini-latest 400K context $0.75/M input $4.5/M output

Opus 4.7 is the next generation of Anthropic's Opus family, built for long-running, asynchronous agents. Building on the coding and agentic strengths of Opus 4.6, it delivers stronger performance on...

anthropic/claude-opus-4.7 1M context $5/M input $25/M output

Muse Spark 1.2 contributor tier is a reasoning model from Meta designed for developers who want to start building at an even lower cost. It’s meaningfully cheaper than Muse Spark...

meta/muse-spark-1.2-contributor 1.04858M context $0.1/M input $0.2/M output

The Qwen3.5 27B native vision-language Dense model incorporates a linear attention mechanism, delivering fast response times while balancing inference speed and performance. Its overall capabilities are comparable to those of...

qwen/qwen3.5-27b 262.144K context $0.195/M input $1.56/M output