No provider description is available for this model yet.
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Ox Alpha is a reasoning model designed for coding, sustained agentic work, and production workloads. It is suited for long-horizon software engineering, complex reasoning, and workflows that combine text with...
The largest model in the Ministral 3 family, Ministral 3 14B offers frontier capabilities and performance comparable to its larger Mistral Small 3.2 24B counterpart. A powerful and efficient language...
GLM-4.5 is our latest flagship foundation model, purpose-built for agent-based applications. It leverages a Mixture-of-Experts (MoE) architecture and supports a context length of up to 128k tokens. GLM-4.5 delivers significantly...
Ling-2.6-1T is an instant (instruct) model from inclusionAI and the company’s trillion-parameter flagship, designed for real-world agents that require fast execution and high efficiency at scale. It uses a “fast...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
NVIDIA Nemotron 3 Nano 30B A3B is a small language MoE model with highest compute efficiency and accuracy for developers to build specialized agentic AI systems. The model is fully...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Opus 4.6 is Anthropic’s strongest model for coding and long-running professional tasks. It is built for agents that operate across entire workflows rather than single prompts, making it especially effective...
gpt-oss-20b is an open-weight 21B parameter model released by OpenAI under the Apache 2.0 license. It uses a Mixture-of-Experts (MoE) architecture with 3.6B active parameters per forward pass, optimized for...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
The simplest way to get free inference. openrouter/free is a router that selects free models at random from the models available on OpenRouter. The router smartly filters for models that...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Qwen3-Coder-Next is an open-weight causal language model optimized for coding agents and local development workflows. It uses a sparse MoE design with 80B total parameters and only 3B activated per...
DeepSeek-V3.1 is a large hybrid reasoning model (671B parameters, 37B active) that supports both thinking and non-thinking modes via prompt templates. It extends the DeepSeek-V3 base with a two-phase long-context...
Fast-mode variant of [Opus 4.7](/anthropic/claude-opus-4.7) - identical capabilities with higher output speed at premium 6x pricing. Learn more in Anthropic's docs: https://platform.claude.com/docs/en/build-with-claude/fast-mode
Llama 3.2 3B is a 3-billion-parameter multilingual large language model, optimized for advanced natural language processing tasks like dialogue generation, reasoning, and summarization. Designed with the latest transformer architecture, it...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
MiniMax-M2.5 is a SOTA large language model designed for real-world productivity. Trained in a diverse range of complex real-world digital working environments, M2.5 builds upon the coding expertise of M2.1...
No provider description is available for this model yet.
No provider description is available for this model yet.
Hy-MT2-7B is a 7B-parameter translation model from Tencent. It supports 33 language pairs and five Chinese dialect and minority-language pairs, with workflows for structured, delimiter-based, contextual, glossary-based, and style-guided translation.
No provider description is available for this model yet.
The Qwen3.5 series 397B-A17B native vision-language model is built on a hybrid architecture that integrates a linear attention mechanism with a sparse mixture-of-experts model, achieving higher inference efficiency. It delivers...
Qwen3-VL-30B-A3B-Thinking is a multimodal model that unifies strong text generation with visual understanding for images and videos. Its Thinking variant enhances reasoning in STEM, math, and complex tasks. It excels...
No provider description is available for this model yet.
Muse Spark 1.2 contributor tier is a reasoning model from Meta designed for developers who want to start building at an even lower cost. It’s meaningfully cheaper than Muse Spark...
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| anthropic.claude-opus-4-5-20251101-v1:0bedrock_converse/anthropic.claude-opus-4-5-20251101-v1:0 | 200K | $5 | $25 | — | |||
| LGAI-EXAONE/K-EXAONE-2.0-750B-A37Bfriendliai/lgai-exaone/k-exaone-2.0-750b-a37b | 262.144K | $0.6 | $2.4 | — | |||
| zai-glm-4.7cerebras/zai-glm-4.7 | 128K | $2.25 | $2.75 | — | |||
| anthropic.claude-opus-4-20250514-v1:0bedrock_converse/anthropic.claude-opus-4-20250514-v1:0 | 200K | $15 | $75 | — | |||
| Ox Alphastealth/ox-alpha | 1.04858M | — | — | — | |||
| Mistral: Ministral 3 14B 2512mistralai/ministral-14b-2512 | 262.144K | $0.2 | $0.2 | — | |||
| Z.ai: GLM 4.5z-ai/glm-4.5 | 131.072K | $0.6 | $2.2 | — | |||
| inclusionAI: Ling-2.6-1Tinclusionai/ling-2.6-1t | 262.144K | $0.075 | $0.625 | — | |||
| deepseek-ai/ESFT-token-summary-litedeepseek-ai/ESFT-token-summary-lite | Not documented | — | — | — | |||
| anthropic.claude-opus-4-1-20250805-v1:0bedrock_converse/anthropic.claude-opus-4-1-20250805-v1:0 | 200K | $15 | $75 | — | |||
| zai-org/GLM-5.2friendliai/zai-org/glm-5.2 | 1.04858M | $1.4 | $4.4 | — | |||
| meta-llama/CodeLlama-34b-hfmeta-llama/CodeLlama-34b-hf | Not documented | — | — | — | |||
| deepseek-ai/ESFT-token-code-litedeepseek-ai/ESFT-token-code-lite | Not documented | — | — | — | |||
| google/gemma-4-31B-itfriendliai/google/gemma-4-31b-it | 262.144K | $0.14 | $0.4 | — | |||
| NVIDIA: Nemotron 3 Nano 30B A3B (free)nvidia/nemotron-3-nano-30b-a3b:free | 256K | Free | Free | — | |||
| anthropic.claude-instant-v1bedrock/anthropic.claude-instant-v1 | 100K | $0.8 | $2.4 | — | |||
| deepseek-ai/ESFT-token-math-litedeepseek-ai/ESFT-token-math-lite | Not documented | — | — | — | |||
| anthropic.claude-3-sonnet-20240229-v1:0bedrock/anthropic.claude-3-sonnet-20240229-v1:0 | 200K | $3 | $15 | — | |||
| Anthropic: Claude Opus 4.6anthropic/claude-opus-4.6 | 1M | $5 | $25 | — | |||
| OpenAI: gpt-oss-20b (free)openai/gpt-oss-20b:free | 131.072K | Free | Free | — | |||
| meta-llama/CodeLlama-34b-Python-hfmeta-llama/CodeLlama-34b-Python-hf | Not documented | — | — | — | |||
| meta-llama/CodeLlama-34b-Instruct-hfmeta-llama/CodeLlama-34b-Instruct-hf | Not documented | — | — | — | |||
| deepseek-ai/ESFT-token-translation-litedeepseek-ai/ESFT-token-translation-lite | Not documented | — | — | — | |||
| anthropic.claude-3-opus-20240229-v1:0bedrock/anthropic.claude-3-opus-20240229-v1:0 | 200K | $15 | $75 | — | |||
| Free Models Routeropenrouter/free | 200K | Free | Free | — | |||
| anthropic.claude-3-haiku-20240307-v1:0bedrock/anthropic.claude-3-haiku-20240307-v1:0 | 200K | $0.25 | $1.25 | — | |||
| anthropic.claude-3-7-sonnet-20250219-v1:0bedrock_converse/anthropic.claude-3-7-sonnet-20250219-v1:0 | 200K | $3 | $15 | — | |||
| llama3ollama/llama3 | 8.192K | — | — | — | |||
| anthropic.claude-3-7-sonnet-20240620-v1:0bedrock/anthropic.claude-3-7-sonnet-20240620-v1:0 | 200K | $3.6 | $18 | — | |||
| anthropic.claude-3-5-sonnet-20241022-v2:0bedrock/anthropic.claude-3-5-sonnet-20241022-v2:0 | 1M | $3 | $15 | — | |||
| llama2:7bollama/llama2:7b | 4.096K | — | — | — | |||
| Qwen: Qwen3 Coder Nextqwen/qwen3-coder-next | 262.144K | $0.12 | $0.8 | — | |||
| DeepSeek: DeepSeek V3.1deepseek/deepseek-chat-v3.1 | 163.84K | $0.25 | $0.95 | — | |||
| Anthropic: Claude Opus 4.7 (Fast)anthropic/claude-opus-4.7-fast | 1M | $30 | $150 | — | |||
| Meta: Llama 3.2 3B Instructmeta-llama/llama-3.2-3b-instruct | 131.072K | $0.05 | $0.33 | — | |||
| meta-llama/CodeLlama-70b-hfmeta-llama/CodeLlama-70b-hf | Not documented | — | — | — | |||
| meta-llama/CodeLlama-70b-Python-hfmeta-llama/CodeLlama-70b-Python-hf | Not documented | — | — | — | |||
| sa-east-1/deepseek.v3.2bedrock/sa-east-1/deepseek.v3.2 | 163.84K | $0.74 | $2.22 | — | |||
| meta-llama/CodeLlama-70b-Instruct-hfmeta-llama/CodeLlama-70b-Instruct-hf | Not documented | — | — | — | |||
| meta-llama/Llama-2-7bmeta-llama/Llama-2-7b | Not documented | — | — | — | |||
| llama2:70bollama/llama2:70b | 4.096K | — | — | — | |||
| MiniMax: MiniMax M2.5minimax/minimax-m2.5 | 200K | $0.27 | $1.08 | — | |||
| meta-llama/Llama-2-13bmeta-llama/Llama-2-13b | Not documented | — | — | — | |||
| us-gov-east-1/amazon.titan-text-premier-v1:0bedrock/us-gov-east-1/amazon.titan-text-premier-v1:0 | 42K | $0.5 | $1.5 | — | |||
| Tencent: Hy-MT2-7Btencent/hy-mt2-7b | 8.192K | $0.074 | $0.295 | — | |||
| sa-east-1/meta.llama3-8b-instruct-v1:0bedrock/sa-east-1/meta.llama3-8b-instruct-v1:0 | 8.192K | $0.5 | $1.01 | — | |||
| Qwen: Qwen3.5 397B A17Bqwen/qwen3.5-397b-a17b | 262.144K | $0.55 | $3.5 | — | |||
| Qwen: Qwen3 VL 30B A3B Thinkingqwen/qwen3-vl-30b-a3b-thinking | 131.072K | $0.2 | $2.4 | — | |||
| us-gov-east-1/amazon.titan-text-lite-v1bedrock/us-gov-east-1/amazon.titan-text-lite-v1 | 42K | $0.3 | $0.4 | — | |||
| Meta: Muse Spark 1.2 Contributormeta/muse-spark-1.2-contributor | 1.04858M | $0.1 | $0.2 | — |