No provider description is available for this model yet.
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
Dots3-Note Preview is an open-weight mixture-of-experts model from Dots Studio, with 16B active parameters out of 280B total. It is the lightest model in the Dots 3 family and is...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Command A is an open-weights 111B parameter model with a 256k context window focused on delivering great performance across agentic, multilingual, and coding use cases. Compared to other leading proprietary...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Gemini 3.5 Flash Lite is a high-efficiency model from Google with upgraded agentic capabilities. It is suited for subagents that execute focused tasks within complex, multi-agent workflows.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
GPT-6 Astra Pro is the same underlying model as [GPT-6 Astra](https://openrouter.ai/openai/gpt-6-astra), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. Learn more in OpenAI's docs: https://developers.openai.com/api/docs/guides/reasoning#reasoning-mode
No provider description is available for this model yet.
No provider description is available for this model yet.
MiniMax-M2.7 is a next-generation large language model designed for autonomous, real-world productivity and continuous improvement. Built to actively participate in its own evolution, M2.7 integrates advanced agentic capabilities through multi-agent...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
DeepSeek V3, a 685B-parameter, mixture-of-experts model, is the latest iteration of the flagship chat model family from the DeepSeek team. It succeeds the [DeepSeek V3](/deepseek/deepseek-chat-v3) model and performs really well...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Gemma 3n E4B-it is optimized for efficient execution on mobile and low-resource devices, such as phones, laptops, and tablets. It supports multimodal inputs—including text, visual data, and audio—enabling diverse tasks...
No provider description is available for this model yet.
Jamba Large 1.7 is the latest model in the Jamba open family, offering improvements in grounding, instruction-following, and overall efficiency. Built on a hybrid SSM-Transformer architecture with a 256K context...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Qwen3.6 27B is a dense 27-billion-parameter language model from the Qwen Team at Alibaba, released in April 2026. It features hybrid multimodal capabilities — accepting text, image, and video inputs...
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| gpt-audio-mini-2025-10-06azure/gpt-audio-mini-2025-10-06 | 128K | $0.6 | $2.4 | — | |||
| Dots Studio: Dots3-Note Preview (free)dots-studio/dots-3-note-preview:free | 512K | Free | Free | — | |||
| Qwen/Qwen3.7-Plustogether_ai/qwen/qwen3.7-plus | 1M | $0.32 | $1.28 | — | |||
| gpt-4o-audio-preview-2024-12-17azure/gpt-4o-audio-preview-2024-12-17 | 128K | $2.5 | $10 | — | |||
| Qwen/Qwen3.8-2.4T-A95Btogether_ai/qwen/qwen3.8-2.4t-a95b | 1.01M | $2.5 | $6.25 | — | |||
| arize-ai/qwen-2-1.5b-instructtogether_ai/arize-ai/qwen-2-1.5b-instruct | 32.768K | $0.1 | $0.1 | — | |||
| deepseek-ai/DeepSeek-V4-Flash-0731together_ai/deepseek-ai/deepseek-v4-flash-0731 | 1.04858M | $0.14 | $0.28 | — | |||
| gpt-4o-miniazure/gpt-4o-mini | 128K | $0.165 | $0.66 | — | |||
| us-gov.nvidia.nemotron-nano-12b-v2bedrock_converse/us-gov.nvidia.nemotron-nano-12b-v2 | 128K | $0.24 | $0.72 | — | |||
| Cohere: Command Acohere/command-a | 256K | $2.5 | $10 | — | |||
| gpt-4o-mini-2024-07-18azure/gpt-4o-mini-2024-07-18 | 128K | $0.165 | $0.66 | — | |||
| gpt-4o-mini-audio-preview-2024-12-17azure/gpt-4o-mini-audio-preview-2024-12-17 | 128K | $2.5 | $10 | — | |||
| us-west-2/deepseek.v3.2bedrock/us-west-2/deepseek.v3.2 | 163.84K | $0.62 | $1.85 | — | |||
| gpt-5.1-2025-11-13azure/gpt-5.1-2025-11-13 | 272K | $1.25 | $10 | — | |||
| claude-opus-4-6azure_ai/claude-opus-4-6 | 1M | $5 | $25 | — | |||
| gpt-5azure/gpt-5 | 272K | $1.25 | $10 | — | |||
| deepseek-ai/DeepSeek-V4-Pro-0813together_ai/deepseek-ai/deepseek-v4-pro-0813 | 1.04858M | $1.32 | $3.96 | — | |||
| claude-opus-4-5azure_ai/claude-opus-4-5 | 200K | $5 | $25 | — | |||
| google/gemma-3n-E4B-ittogether_ai/google/gemma-3n-e4b-it | 32.768K | $0.06 | $0.12 | — | |||
| gpt-5-2025-08-07azure/gpt-5-2025-08-07 | 272K | $1.25 | $10 | — | |||
| gpt-5-chatazure/gpt-5-chat | 128K | $1.25 | $10 | — | |||
| gpt-5-chat-latestazure/gpt-5-chat-latest | 128K | $1.25 | $10 | — | |||
| google/gemma-4-31B-ittogether_ai/google/gemma-4-31b-it | 262.144K | $0.39 | $0.97 | — | |||
| meta-llama/Llama-Guard-4-12Btogether_ai/meta-llama/llama-guard-4-12b | 1.04858M | $0.2 | $0.2 | — | |||
| @cf/zai-org/glm-4.7-flashcloudflare/@cf/zai-org/glm-4.7-flash | 131.072K | $0.061 | $0.4 | — | |||
| Google: Gemini 3.5 Flash Lite (batch)google/gemini-3.5-flash-lite:batch | 1.04858M | $0.15 | $1.25 | — | |||
| us-gov.nvidia.nemotron-nano-3-30bbedrock_converse/us-gov.nvidia.nemotron-nano-3-30b | 262.144K | $0.072 | $0.288 | — | |||
| gpt-5-miniazure/gpt-5-mini | 272K | $0.25 | $2 | — | |||
| claude-haiku-4-5azure_ai/claude-haiku-4-5 | 200K | $1 | $5 | — | |||
| gpt-5-nanoazure/gpt-5-nano | 272K | $0.05 | $0.4 | — | |||
| OpenAI: GPT-6 Astra Proopenai/gpt-6-astra-pro | 1.05M | $10 | $50 | — | |||
| meta-models/Muse-Glimmer-30Btogether_ai/meta-models/muse-glimmer-30b | 131.072K | $0.35 | $1.5 | — | |||
| moonshotai/Kimi-K2.7-Codetogether_ai/moonshotai/kimi-k2.7-code | 262.144K | $0.95 | $4 | — | |||
| MiniMax: MiniMax M2.7minimax/minimax-m2.7 | 204.8K | $0.3 | $1.2 | — | |||
| gpt-5-nano-2025-08-07azure/gpt-5-nano-2025-08-07 | 272K | $0.05 | $0.4 | — | |||
| gpt-5.1azure/gpt-5.1 | 272K | $1.25 | $10 | — | |||
| moonshotai/Kimi-K3together_ai/moonshotai/kimi-k3 | 1.04858M | $3 | $15 | — | |||
| gpt-5.1-chatazure/gpt-5.1-chat | 128K | $1.25 | $10 | — | |||
| DeepSeek: DeepSeek V3 0324deepseek/deepseek-chat-v3-0324 | 163.84K | $0.25 | $1 | — | |||
| gpt-5.2azure/gpt-5.2 | 272K | $1.75 | $14 | — | |||
| gpt-5.2-2025-12-11azure/gpt-5.2-2025-12-11 | 272K | $1.75 | $14 | — | |||
| gpt-5.2-chatazure/gpt-5.2-chat | 128K | $1.75 | $14 | — | |||
| nvidia/nemotron-3-ultra-550b-a55btogether_ai/nvidia/nemotron-3-ultra-550b-a55b | 512.288K | $0.6 | $3.6 | — | |||
| Google: Gemma 3n 4Bgoogle/gemma-3n-e4b-it | 32.768K | $0.06 | $0.12 | — | |||
| us-west-2/moonshotai.kimi-k2-thinkingbedrock/us-west-2/moonshotai.kimi-k2-thinking | 262.144K | $0.6 | $2.5 | — | |||
| AI21: Jamba Large 1.7ai21/jamba-large-1.7 | 256K | $2 | $8 | — | |||
| gpt-5.2-chat-2025-12-11azure/gpt-5.2-chat-2025-12-11 | 128K | $1.75 | $14 | — | |||
| gpt-5.3-chatazure/gpt-5.3-chat | 128K | $1.75 | $14 | — | |||
| gpt-5.4azure/gpt-5.4 | 1.05M | $2.5 | $15 | — | |||
| Qwen: Qwen3.6 27Bqwen/qwen3.6-27b | 262.144K | $0.3 | $2 | — |