Compact Nemotron model for efficient reasoning and deployable AI agents
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
GPT model for general reasoning, writing, coding, and tool-assisted tasks
Compact open Llama model for lightweight chat, drafting, and self-hosting
Llama 3.1-based safety classifier for moderating prompts and model responses
Open Llama instruction model for multilingual chat, reasoning, and coding
Small omni GPT for cheap multimodal assistance and production-scale traffic
Efficient Mistral-NVIDIA open model for multilingual chat and local deployment
Research model for long-horizon investigation, synthesis, and analytical reports
Research model for long-horizon investigation, synthesis, and analytical reports
Open Mistral code model for fill-in-the-middle and 80+ programming languages
Mistral code model for completions, refactors, and developer IDE workflows
GPT model for general reasoning, writing, coding, and tool-assisted tasks
Omni-era GPT for multimodal chat, practical coding, and general assistants
Qwen vision-language model for visual reasoning, documents, and agent tasks
Flagship Qwen model for complex reasoning, coding, and agentic workflows
Legacy model retained for compatibility with older integrations
Qwen vision-language model for visual reasoning, documents, and agent tasks
Qwen instruction model for multilingual chat, reasoning, and tool use
Fast web-grounded Sonar for current answers, citations, and lightweight retrieval
Deeper Sonar search model with broader retrieval and stronger synthesis
Web-grounded Sonar for multi-step research questions that need cited reasoning
Compact GPT model for low-latency assistance and high-volume workloads
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
GPT-6 Astra Pro is the same underlying model as [GPT-6 Astra](https://openrouter.ai/openai/gpt-6-astra), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. Learn more in OpenAI's docs: https://developers.openai.com/api/docs/guides/reasoning#reasoning-mode
No provider description is available for this model yet.
Nex-N2.5 is an agentic model built to turn goals into working, verified outcomes. Its core strength is agentic coding within a visual feedback loop: it can explore codebases, implement multi-file...
Gemini 3.6 Flash is a high-efficiency model from Google for coding, agentic workflows, and web and app development. It is designed to produce polished outputs with fewer unnecessary edits and...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Gemini 3.5 Flash Lite is a high-efficiency model from Google with upgraded agentic capabilities. It is suited for subagents that execute focused tasks within complex, multi-agent workflows.
Claude Fable 5 is a Mythos-class model from Anthropic, built for autonomous knowledge work and coding. It supports text, image, and file inputs with text output, with reasoning support and...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| Nemotron Mini 4B Instructnvidia/nemotron-mini-4b-instruct | 128K | — | — | 2024-08-21 | |||
| GPT-4o (2024-08-06)openai/gpt-4o-2024-08-06 | 128K | $2.5 | $10 | 2024-08-06 | |||
| Llama-3.1-8B-Instructmeta/llama-3.1-8b-instruct | 128K | $0.02 | $0.04 | 2024-07-23 | |||
| Llama-Guard-3-8Bmeta/llama-guard-3-8b | 128K | — | — | 2024-07-23 | |||
| Llama-3.1-70B-Instructmeta/llama-3.1-70b-instruct | 128K | $0.4 | $0.4 | 2024-07-23 | |||
| GPT-4o miniopenai/gpt-4o-mini | 128K | $0.15 | $0.6 | 2024-07-18 | |||
| Mistral Nemomistral/mistral-nemo | 128K | $0.15 | $0.15 | 2024-07-01 | |||
| o4-mini-deep-researchopenai/o4-mini-deep-research | 200K | $1.8 | $7.2 | 2024-06-26 | |||
| o3-deep-researchopenai/o3-deep-research | 200K | $9 | $36 | 2024-06-26 | |||
| Codestral-22B-v0.1mistral/codestral-22b-v0.1 | 32.768K | $0.3 | $0.9 | 2024-05-29 | |||
| Codestral (latest)mistral/codestral-latest | 256K | $0.3 | $0.9 | 2024-05-29 | |||
| GPT-4o (2024-05-13)openai/gpt-4o-2024-05-13 | 128K | $5 | $15 | 2024-05-13 | |||
| GPT-4oopenai/gpt-4o | 128K | $2.5 | $10 | 2024-05-13 | |||
| Qwen-VL Maxalibaba/qwen-vl-max | 131.072K | $0.8 | $3.2 | 2024-04-08 | |||
| Qwen Maxalibaba/qwen-max | 32.768K | $1.6 | $6.4 | 2024-04-03 | |||
| Claude Haiku 3anthropic/claude-3-haiku-20240307 | 200K | $0.25 | $1.25 | 2024-03-13 | |||
| Qwen-VL Plusalibaba/qwen-vl-plus | 131.072K | $0.21 | $0.63 | 2024-01-25 | |||
| Qwen Plusalibaba/qwen-plus | 1M | $0.4 | $1.2 | 2024-01-25 | |||
| Sonarperplexity/sonar | 128K | $1 | $1 | 2024-01-01 | |||
| Sonar Properplexity/sonar-pro | 200K | $3 | $15 | 2024-01-01 | |||
| Sonar Reasoning Properplexity/sonar-reasoning-pro | 128K | $2 | $8 | 2024-01-01 | |||
| GPT-4 Turboopenai/gpt-4-turbo | 128K | $10 | $30 | 2023-11-06 | |||
| ap-northeast-1/anthropic.claude-v2:1bedrock/ap-northeast-1/anthropic.claude-v2:1 | 100K | $8 | $24 | — | |||
| ap-northeast-1/anthropic.claude-v1bedrock/ap-northeast-1/anthropic.claude-v1 | 100K | $8 | $24 | — | |||
| us/gpt-5.5-2026-04-23azure/us/gpt-5.5-2026-04-23 | 1.05M | $5.5 | $33 | — | |||
| ap-northeast-1/anthropic.claude-instant-v1bedrock/ap-northeast-1/anthropic.claude-instant-v1 | 100K | $2.23 | $7.55 | — | |||
| ap-northeast-1/6-month-commitment/anthropic.claude-v2:1bedrock/ap-northeast-1/6-month-commitment/anthropic.claude-v2:1 | 100K | — | — | — | |||
| OpenAI: GPT-6 Astra Pro (batch)openai/gpt-6-astra-pro:batch | 1.05M | $5 | $25 | — | |||
| ap-northeast-1/6-month-commitment/anthropic.claude-v1bedrock/ap-northeast-1/6-month-commitment/anthropic.claude-v1 | 100K | — | — | — | |||
| Nex AGI: Nex-N2.5-Mini (free)nex-agi/nex-n2.5-mini:free | 262.144K | Free | Free | — | |||
| Google: Gemini 3.6 Flash (batch)google/gemini-3.6-flash:batch | 1.04858M | $0.375 | $1.875 | — | |||
| ap-northeast-1/6-month-commitment/anthropic.claude-instant-v1bedrock/ap-northeast-1/6-month-commitment/anthropic.claude-instant-v1 | 100K | — | — | — | |||
| gpt-5-miniazure/gpt-5-mini | 272K | $0.25 | $2 | — | |||
| gpt-audio-2025-08-28azure/gpt-audio-2025-08-28 | 128K | $2.5 | $10 | — | |||
| ap-northeast-1/1-month-commitment/anthropic.claude-v2:1bedrock/ap-northeast-1/1-month-commitment/anthropic.claude-v2:1 | 100K | — | — | — | |||
| ap-northeast-1/1-month-commitment/anthropic.claude-v1bedrock/ap-northeast-1/1-month-commitment/anthropic.claude-v1 | 100K | — | — | — | |||
| llama-3.3-70bcerebras/llama-3.3-70b | 128K | $0.85 | $1.2 | — | |||
| ap-northeast-1/1-month-commitment/anthropic.claude-instant-v1bedrock/ap-northeast-1/1-month-commitment/anthropic.claude-instant-v1 | 100K | — | — | — | |||
| gpt-5.5-2026-04-23azure/gpt-5.5-2026-04-23 | 1.05M | $5 | $30 | — | |||
| Google: Gemini 3.5 Flash Lite (batch)google/gemini-3.5-flash-lite:batch | 1.04858M | $0.15 | $1.25 | — | |||
| Anthropic: Claude Fable 5 (batch)anthropic/claude-fable-5:batch | 1M | $5 | $25 | — | |||
| mistral-small-2503azure_ai/mistral-small-2503 | 128K | $0.1 | $0.3 | — | |||
| eu/gpt-5.5azure/eu/gpt-5.5 | 1.05M | $5.5 | $33 | — | |||
| mistral-smallazure_ai/mistral-small | 32K | $1 | $3 | — | |||
| us/gpt-5.5azure/us/gpt-5.5 | 1.05M | $5.5 | $33 | — | |||
| @cf/zai-org/glm-4.7-flashcloudflare/@cf/zai-org/glm-4.7-flash | 131.072K | $0.061 | $0.4 | — | |||
| Prism-ML/Ternary-Bonsai-27Btogether_ai/prism-ml/ternary-bonsai-27b | 262.144K | — | — | — | |||
| gpt-4.1azure/gpt-4.1 | 1.04758M | $2 | $8 | — | |||
| mistral-nemoazure_ai/mistral-nemo | 131.072K | $0.15 | $0.15 | — | |||
| mistral-medium-2505azure_ai/mistral-medium-2505 | 131.072K | $0.4 | $2 | — |