No provider description is available for this model yet.
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
No provider description is available for this model yet.
Qwen3-VL-8B-Thinking is the reasoning-optimized variant of the Qwen3-VL-8B multimodal model, designed for advanced visual and textual reasoning across complex scenes, documents, and temporal sequences. It integrates enhanced multimodal alignment and...
Gemma 3n E4B-it is optimized for efficient execution on mobile and low-resource devices, such as phones, laptops, and tablets. It supports multimodal inputs—including text, visual data, and audio—enabling diverse tasks...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Claude Haiku 4.5 is Anthropic’s fastest and most efficient model, delivering near-frontier intelligence at a fraction of the cost and latency of larger Claude models. Matching Claude Sonnet 4’s performance...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Qwen3-Next-80B-A3B-Instruct is an instruction-tuned chat model in the Qwen3-Next series optimized for fast, stable responses without “thinking” traces. It targets complex tasks across reasoning, code generation, knowledge QA, and multilingual...
Morph's high-accuracy apply model for complex code edits. ~4,500 tokens/sec with 98% accuracy for precise code transformations. The model requires the prompt to be in the following format: <instruction>{instruction}</instruction> <code>{initial_code}</code>...
Mistral Small 4 is the next major release in the Mistral Small family, unifying the capabilities of several flagship Mistral models into a single system. It combines strong reasoning from...
This model always redirects to the latest Grok model from xAI.
GPT-5.6 Terra Pro is the same underlying model as [GPT-5.6 Terra](https://openrouter.ai/openai/gpt-5.6-terra), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. Learn more in OpenAI's docs: https://developers.openai.com/api/docs/guides/reasoning#reasoning-mode
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
NVIDIA Nemotron Nano 2 VL is a 12-billion-parameter open multimodal reasoning model designed for video understanding and document intelligence. It introduces a hybrid Transformer-Mamba architecture, combining transformer-level accuracy with Mamba’s...
No provider description is available for this model yet.
Hunyuan-A13B is a 13B active parameter Mixture-of-Experts (MoE) language model developed by Tencent, with a total parameter count of 80B and support for reasoning via Chain-of-Thought. It offers competitive benchmark...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| us-gov-west-1/anthropic.claude-3-5-sonnet-20240620-v1:0bedrock/us-gov-west-1/anthropic.claude-3-5-sonnet-20240620-v1:0 | 200K | $3.6 | $18 | — | |||
| sa-east-1/qwen.qwen3-coder-nextbedrock/sa-east-1/qwen.qwen3-coder-next | 262.144K | $0.6 | $1.44 | — | |||
| Qwen: Qwen3 VL 8B Thinkingqwen/qwen3-vl-8b-thinking | 131.072K | $0.18 | $2.1 | — | |||
| Google: Gemma 3n 4Bgoogle/gemma-3n-e4b-it | 32.768K | $0.06 | $0.12 | — | |||
| kimi-k2.7-codedashscope/kimi-k2.7-code | 229.376K | $0.95 | $4 | — | |||
| glm-5.2dashscope/glm-5.2 | 1.04858M | $1.4 | $4.4 | — | |||
| sa-east-1/moonshotai.kimi-k2.5bedrock/sa-east-1/moonshotai.kimi-k2.5 | 262.144K | $0.72 | $3.6 | — | |||
| us.anthropic.claude-opus-5bedrock_converse/us.anthropic.claude-opus-5 | 1M | $5.5 | $27.5 | — | |||
| global.anthropic.claude-opus-5bedrock_converse/global.anthropic.claude-opus-5 | 1M | $5 | $25 | — | |||
| gpt-5.6-cyberopenai/gpt-5.6-cyber | 400K | $12.5 | $75 | — | |||
| Anthropic: Claude Haiku 4.5anthropic/claude-haiku-4.5 | 200K | $1 | $5 | — | |||
| anthropic.claude-opus-5bedrock_converse/anthropic.claude-opus-5 | 1M | $5 | $25 | — | |||
| eu.anthropic.claude-fable-5bedrock_converse/eu.anthropic.claude-fable-5 | 1M | $11 | $55 | — | |||
| us.anthropic.claude-fable-5bedrock_converse/us.anthropic.claude-fable-5 | 1M | $11 | $55 | — | |||
| @cf/aisingapore/gemma-sea-lion-v4-27b-itcloudflare/@cf/aisingapore/gemma-sea-lion-v4-27b-it | 128K | $0.351 | $0.555 | — | |||
| sa-east-1/moonshotai.kimi-k2-thinkingbedrock/sa-east-1/moonshotai.kimi-k2-thinking | 262.144K | $0.73 | $3.03 | — | |||
| Qwen: Qwen3 Next 80B A3B Instructqwen/qwen3-next-80b-a3b-instruct | 262.144K | $0.09 | $1.1 | — | |||
| Morph: Morph V3 Largemorph/morph-v3-large | 262.144K | $0.9 | $1.9 | — | |||
| Mistral: Mistral Small 4mistralai/mistral-small-2603 | 262.144K | $0.15 | $0.6 | — | |||
| xAI: Grok Latest~x-ai/grok-latest | 500K | $2 | $6 | — | |||
| OpenAI: GPT-5.6 Terra Proopenai/gpt-5.6-terra-pro | 1.05M | $2 | $12 | — | |||
| global.anthropic.claude-fable-5bedrock_converse/global.anthropic.claude-fable-5 | 1M | $10 | $50 | — | |||
| global.openai.gpt-5.6-lunabedrock_converse/global.openai.gpt-5.6-luna | 1M | $0.2 | $1.2 | — | |||
| sa-east-1/minimax.minimax-m2.5bedrock/sa-east-1/minimax.minimax-m2.5 | 1M | $0.36 | $1.44 | — | |||
| us.openai.gpt-5.6-lunabedrock_converse/us.openai.gpt-5.6-luna | 1M | $0.22 | $1.32 | — | |||
| qwen3-vl-32b-instructdashscope/qwen3-vl-32b-instruct | 131.072K | $0.16 | $0.64 | — | |||
| glm-5.1dashscope/glm-5.1 | 202.745K | $1.4 | $4.4 | — | |||
| gpt-6-astraazure_ai/gpt-6-astra | 922K | $10 | $50 | — | |||
| qwen3-vl-235b-a22b-instructdashscope/qwen3-vl-235b-a22b-instruct | 131.072K | $0.4 | $1.6 | — | |||
| qwen3-next-80b-a3b-instructdashscope/qwen3-next-80b-a3b-instruct | 262.144K | $0.15 | $1.2 | — | |||
| deepseek-v4-prodashscope/deepseek-v4-pro | 1M | $2.4 | $4.8 | — | |||
| qwen3-max-2026-01-23dashscope/qwen3-max-2026-01-23 | 258.048K | — | — | — | |||
| qwen3-maxdashscope/qwen3-max | 258.048K | — | — | — | |||
| us-gov-west-1/anthropic.claude-3-7-sonnet-20250219-v1:0bedrock/us-gov-west-1/anthropic.claude-3-7-sonnet-20250219-v1:0 | 200K | $3.6 | $18 | — | |||
| qwen3-max-previewdashscope/qwen3-max-preview | 258.048K | — | — | — | |||
| qwen3-coder-plus-2025-07-22dashscope/qwen3-coder-plus-2025-07-22 | 997.952K | — | — | — | |||
| global.openai.gpt-5.6-terrabedrock_converse/global.openai.gpt-5.6-terra | 1M | $2 | $12 | — | |||
| MiniMaxAI/MiniMax-M2.5friendliai/minimaxai/minimax-m2.5 | 196.608K | $0.3 | $1.2 | — | |||
| NVIDIA: Nemotron Nano 12B 2 VL (free)nvidia/nemotron-nano-12b-v2-vl:free | 128K | Free | Free | — | |||
| qwen3-next-80b-a3b-thinkingdashscope/qwen3-next-80b-a3b-thinking | 262.144K | $0.15 | $1.2 | — | |||
| Tencent: Hunyuan A13B Instructtencent/hunyuan-a13b-instruct | 131.072K | $0.14 | $0.57 | — | |||
| qwen3-vl-235b-a22b-thinkingdashscope/qwen3-vl-235b-a22b-thinking | 131.072K | $0.4 | $4 | — | |||
| qwen3-coder-plusdashscope/qwen3-coder-plus | 997.952K | — | — | — | |||
| qwen3-vl-32b-thinkingdashscope/qwen3-vl-32b-thinking | 131.072K | $0.16 | $2.87 | — | |||
| qwen3-vl-plusdashscope/qwen3-vl-plus | 260.096K | — | — | — | |||
| qwen3.5-plusdashscope/qwen3.5-plus | 991.808K | — | — | — | |||
| qwen3.7-maxdashscope/qwen3.7-max | 991.808K | $2.5 | $7.5 | — | |||
| qwen3.7-plusdashscope/qwen3.7-plus | 991.808K | — | — | — | |||
| qwq-plusdashscope/qwq-plus | 98.304K | $0.8 | $2.4 | — | |||
| qwen3-coder-flash-2025-07-28dashscope/qwen3-coder-flash-2025-07-28 | 997.952K | — | — | — |