No provider description is available for this model yet.
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
Qwen3.8 Max 0902 is an updated snapshot of Qwen3.8 Max from Alibaba's Qwen team. It is a 2.4-trillion-parameter mixture-of-experts model that accepts text, image, and video input and returns text,...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
GPT-5.6 Terra Pro is the same underlying model as [GPT-5.6 Terra](https://openrouter.ai/openai/gpt-5.6-terra), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. Learn more in OpenAI's docs: https://developers.openai.com/api/docs/guides/reasoning#reasoning-mode
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Hy-MT2-1.8B is a compact 1.8B-parameter translation model from Tencent. It supports 33 language pairs and five Chinese dialect and minority-language pairs, with workflows for structured, delimiter-based, contextual, glossary-based, and style-guided...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Qwen Plus 0728, based on the Qwen3 foundation model, is a 1 million context hybrid reasoning model with a balanced performance, speed, and cost combination.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Qwen3-Next-80B-A3B-Instruct is an instruction-tuned chat model in the Qwen3-Next series optimized for fast, stable responses without “thinking” traces. It targets complex tasks across reasoning, code generation, knowledge QA, and multilingual...
Morph's high-accuracy apply model for complex code edits. ~4,500 tokens/sec with 98% accuracy for precise code transformations. The model requires the prompt to be in the following format: <instruction>{instruction}</instruction> <code>{initial_code}</code>...
No provider description is available for this model yet.
Coder‑Large is a 32 B‑parameter offspring of Qwen 2.5‑Instruct that has been further trained on permissively‑licensed GitHub, CodeSearchNet and synthetic bug‑fix corpora. It supports a 32k context window, enabling multi‑file...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Schematron V2 Small is a 3B-parameter HTML-to-JSON extraction model from Inference.net. It prioritizes extraction quality for complex schemas and long pages. Extraction instructions must be supplied through a JSON schema...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| @cf/nvidia/nemotron-3-120b-a12bcloudflare/@cf/nvidia/nemotron-3-120b-a12b | 256K | $0.5 | $1.5 | — | |||
| Qwen: Qwen3.8 Max (0902)qwen/qwen3.8-max-0902 | 1M | $2 | $6 | — | |||
| us.openai.gpt-5.6-terrabedrock_converse/us.openai.gpt-5.6-terra | 1M | $2.2 | $13.2 | — | |||
| deepseek-v4-flashdashscope/deepseek-v4-flash | 1M | $0.2 | $0.4 | — | |||
| deepseek-v4-flash-0731dashscope/deepseek-v4-flash-0731 | 1M | $0.2 | $0.4 | — | |||
| global.openai.gpt-5.6-terrabedrock_converse/global.openai.gpt-5.6-terra | 1M | $2 | $12 | — | |||
| google/gemma-4-31B-it-assistantgoogle/gemma-4-31B-it-assistant | Not documented | — | — | — | |||
| us-gov-west-1/anthropic.claude-3-7-sonnet-20250219-v1:0bedrock/us-gov-west-1/anthropic.claude-3-7-sonnet-20250219-v1:0 | 200K | $3.6 | $18 | — | |||
| deepseek-v4-prodashscope/deepseek-v4-pro | 1M | $2.4 | $4.8 | — | |||
| OpenAI: GPT-5.6 Terra Pro (batch)openai/gpt-5.6-terra-pro:batch | 1.05M | $1 | $6 | — | |||
| us-east-1/6-month-commitment/anthropic.claude-instant-v1bedrock/us-east-1/6-month-commitment/anthropic.claude-instant-v1 | 100K | — | — | — | |||
| us.openai.gpt-5.6-lunabedrock_converse/us.openai.gpt-5.6-luna | 1M | $0.22 | $1.32 | — | |||
| global.openai.gpt-5.6-lunabedrock_converse/global.openai.gpt-5.6-luna | 1M | $0.2 | $1.2 | — | |||
| global.anthropic.claude-fable-5bedrock_converse/global.anthropic.claude-fable-5 | 1M | $10 | $50 | — | |||
| @cf/aisingapore/gemma-sea-lion-v4-27b-itcloudflare/@cf/aisingapore/gemma-sea-lion-v4-27b-it | 128K | $0.351 | $0.555 | — | |||
| us.anthropic.claude-fable-5bedrock_converse/us.anthropic.claude-fable-5 | 1M | $11 | $55 | — | |||
| eu.anthropic.claude-fable-5bedrock_converse/eu.anthropic.claude-fable-5 | 1M | $11 | $55 | — | |||
| anthropic.claude-opus-5bedrock_converse/anthropic.claude-opus-5 | 1M | $5 | $25 | — | |||
| global.anthropic.claude-opus-5bedrock_converse/global.anthropic.claude-opus-5 | 1M | $5 | $25 | — | |||
| us.anthropic.claude-opus-5bedrock_converse/us.anthropic.claude-opus-5 | 1M | $5.5 | $27.5 | — | |||
| meta-llama/CodeLlama-7b-Python-hfmeta-llama/CodeLlama-7b-Python-hf | Not documented | — | — | — | |||
| glm-5.2dashscope/glm-5.2 | 1.04858M | $1.4 | $4.4 | — | |||
| kimi-k2.7-codedashscope/kimi-k2.7-code | 229.376K | $0.95 | $4 | — | |||
| Tencent: Hy-MT2-1.8Btencent/hy-mt2-1.8b | 8.192K | $0.044 | $0.177 | — | |||
| us-gov-west-1/anthropic.claude-3-5-sonnet-20240620-v1:0bedrock/us-gov-west-1/anthropic.claude-3-5-sonnet-20240620-v1:0 | 200K | $3.6 | $18 | — | |||
| qwen3.8-maxdashscope/qwen3.8-max | 991.808K | $2 | $6 | — | |||
| nvidia/NVIDIA-Nemotron-3.5-Lightningdeepinfra/nvidia/nvidia-nemotron-3.5-lightning | 262.144K | $0.05 | $0.2 | — | |||
| gemini-3.7-flashvertex_ai/gemini-3.7-flash | 1.04858M | $0.75 | $3.75 | — | |||
| meta-llama/CodeLlama-7b-Instruct-hfmeta-llama/CodeLlama-7b-Instruct-hf | Not documented | — | — | — | |||
| Qwen: Qwen Plus 0728qwen/qwen-plus-2025-07-28 | 1M | $0.26 | $0.78 | — | |||
| gemini-3.7-flashgemini/gemini-3.7-flash | 1.04858M | $0.75 | $3.75 | — | |||
| us-gov-west-1/anthropic.claude-3-haiku-20240307-v1:0bedrock/us-gov-west-1/anthropic.claude-3-haiku-20240307-v1:0 | 200K | $0.3 | $1.5 | — | |||
| us-east-1/1-month-commitment/anthropic.claude-v2:1bedrock/us-east-1/1-month-commitment/anthropic.claude-v2:1 | 100K | — | — | — | |||
| microsoft/UniRG-CXRmicrosoft/UniRG-CXR | Not documented | — | — | — | |||
| eu.anthropic.claude-opus-5bedrock_converse/eu.anthropic.claude-opus-5 | 1M | $5.5 | $27.5 | — | |||
| Qwen: Qwen3 Next 80B A3B Instructqwen/qwen3-next-80b-a3b-instruct | 262.144K | $0.09 | $1.1 | — | |||
| Morph: Morph V3 Largemorph/morph-v3-large | 262.144K | $0.9 | $1.9 | — | |||
| google/gemma-4-26B-A4B-it-assistantgoogle/gemma-4-26B-A4B-it-assistant | Not documented | — | — | — | |||
| Arcee AI: Coder Largearcee-ai/coder-large | 32.768K | $0.5 | $0.8 | — | |||
| @cf/qwen/qwen3-30b-a3b-fp8cloudflare/@cf/qwen/qwen3-30b-a3b-fp8 | 32.768K | $0.051 | $0.335 | — | |||
| google/gemma-4-E4B-it-qat-mobile-transformersgoogle/gemma-4-E4B-it-qat-mobile-transformers | Not documented | — | — | — | |||
| anthropic.claude-opus-4-8bedrock_converse/anthropic.claude-opus-4-8 | 1M | $5 | $25 | — | |||
| gemini-3.7-flashvertex_ai-language-models/gemini-3.7-flash | 1.04858M | $0.75 | $3.75 | — | |||
| Inference.net: Schematron V2 Smallinference-net/schematron-v2-small | 128K | $0.05 | $0.23 | — | |||
| daybreak-blue-latestopenai/daybreak-blue-latest | 1.05M | $5 | $30 | — | |||
| meta-llama/llama-prompt-guard-2-86mgroq/meta-llama/llama-prompt-guard-2-86m | 512 | $0.04 | $0.04 | — | |||
| qwen/qwen3.6-27bgroq/qwen/qwen3.6-27b | 131.072K | $0.6 | $3 | — | |||
| us-gov-west-1/anthropic.claude-sonnet-4-5-20250929-v1:0bedrock/us-gov-west-1/anthropic.claude-sonnet-4-5-20250929-v1:0 | 200K | $3.6 | $18 | — | |||
| google/gemma-4-E2B-it-qat-mobile-transformersgoogle/gemma-4-E2B-it-qat-mobile-transformers | Not documented | — | — | — | |||
| moonshotai/Kimi-K2-Instruct-0905baseten/moonshotai/kimi-k2-instruct-0905 | Not documented | $0.6 | $2.5 | — |