No provider description is available for this model yet.
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
No provider description is available for this model yet.
Reka Flash 3 is a general-purpose, instruction-tuned large language model with 21 billion parameters, developed by Reka. It excels at general chat, coding tasks, instruction-following, and function calling. Featuring a...
No provider description is available for this model yet.
No provider description is available for this model yet.
Command A is an open-weights 111B parameter model with a 256k context window focused on delivering great performance across agentic, multilingual, and coding use cases. Compared to other leading proprietary...
Llama 4 Maverick 17B Instruct (128E) is a high-capacity multimodal language model from Meta, built on a mixture-of-experts (MoE) architecture with 128 experts and 17 billion active parameters per forward...
Ling-2.6-flash is an instant (instruct) model from inclusionAI with 104B total parameters and 7.4B active parameters, designed for real-world agents that require fast responses, strong execution, and high token efficiency....
No provider description is available for this model yet.
No provider description is available for this model yet.
DeepSeek V3, a 685B-parameter, mixture-of-experts model, is the latest iteration of the flagship chat model family from the DeepSeek team. It succeeds the [DeepSeek V3](/deepseek/deepseek-chat-v3) model and performs really well...
No provider description is available for this model yet.
No provider description is available for this model yet.
Llama 4 Scout 17B Instruct (16E) is a mixture-of-experts (MoE) language model developed by Meta, activating 17 billion parameters out of a total of 109B. It supports native multimodal input...
Skyfall 36B v2 is an enhanced iteration of Mistral Small 2501, specifically fine-tuned for improved creativity, nuanced writing, role-playing, and coherent storytelling.
No provider description is available for this model yet.
No provider description is available for this model yet.
Qwen3-32B is a dense 32.8B parameter causal language model from the Qwen3 series, optimized for both complex reasoning and efficient dialogue. It supports seamless switching between a "thinking" mode for...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Qwen3-8B is a dense 8.2B parameter causal language model from the Qwen3 series, designed for both reasoning-heavy tasks and efficient dialogue. It supports seamless switching between "thinking" mode for math,...
No provider description is available for this model yet.
Gemini 2.5 Pro is Google’s state-of-the-art AI model designed for advanced reasoning, coding, mathematics, and scientific tasks. It employs “thinking” capabilities, enabling it to reason through responses with enhanced accuracy...
No provider description is available for this model yet.
No provider description is available for this model yet.
Inkling Small is an open-weight multimodal mixture-of-experts model from Thinking Machines Lab, with 12B active parameters out of 276B total. It is positioned as the smaller, more efficient member of...
Claude Fable 5.1 improves on Claude Fable 5 across the board, with the biggest gains in agentic coding, long-running agentic workflows, and knowledge work: long code refactors, front-end and visual...
Qwen3.6-35B-A3B is an open-weight multimodal model from Alibaba Cloud with 35 billion total parameters and 3 billion active parameters per token. It uses a hybrid sparse mixture-of-experts architecture combining Gated...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Gemma 4 31B Instruct is Google DeepMind's 30.7B dense multimodal model supporting text and image input with text output. Features a 256K token context window, configurable thinking/reasoning mode, native function...
No provider description is available for this model yet.
No provider description is available for this model yet.
Qwen2.5 72B is the latest series of Qwen large language models. Qwen2.5 brings the following improvements upon Qwen2: - Significantly more knowledge and has greatly improved capabilities in coding and...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
GPT-6 Astra Pro is the same underlying model as [GPT-6 Astra](https://openrouter.ai/openai/gpt-6-astra), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. Learn more in OpenAI's docs: https://developers.openai.com/api/docs/guides/reasoning#reasoning-mode
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| us-gov.anthropic.claude-opus-5bedrock_converse/us-gov.anthropic.claude-opus-5 | 1M | $6 | $30 | — | |||
| mistralollama/mistral | 8.192K | — | — | — | |||
| Reka Flash 3rekaai/reka-flash-3 | 65.536K | $0.1 | $0.2 | — | |||
| us-east-1/deepseek.v3.2bedrock/us-east-1/deepseek.v3.2 | 163.84K | $0.62 | $1.85 | — | |||
| us-east-1/mistral.mixtral-8x7b-instruct-v0:1bedrock/us-east-1/mistral.mixtral-8x7b-instruct-v0:1 | 32K | $0.45 | $0.7 | — | |||
| Cohere: Command Acohere/command-a | 256K | $2.5 | $10 | — | |||
| Meta: Llama 4 Maverickmeta-llama/llama-4-maverick | 128K | $0.2 | $0.696 | — | |||
| inclusionAI: Ling-2.6-flashinclusionai/ling-2.6-flash | 262.144K | $0.01 | $0.03 | — | |||
| FW-GLM-5azure_ai/fw-glm-5 | 200K | $1.1 | $3.52 | — | |||
| FW-DeepSeek-V4-Proazure_ai/fw-deepseek-v4-pro | 1M | $1.925 | $3.828 | — | |||
| DeepSeek: DeepSeek V3 0324deepseek/deepseek-chat-v3-0324 | 163.84K | $0.25 | $1 | — | |||
| FW-DeepSeek-V3.2azure_ai/fw-deepseek-v3.2 | 163.84K | $0.62 | $1.85 | — | |||
| us-gov.anthropic.claude-3-haiku-20240307-v1:0bedrock_converse/us-gov.anthropic.claude-3-haiku-20240307-v1:0 | 200K | $0.3 | $1.5 | — | |||
| Meta: Llama 4 Scoutmeta-llama/llama-4-scout | 327.68K | $0.1 | $0.3 | — | |||
| TheDrummer: Skyfall 36B V2thedrummer/skyfall-36b-v2 | 32.768K | $0.55 | $0.8 | — | |||
| us-east-1/mistral.mistral-large-2402-v1:0bedrock/us-east-1/mistral.mistral-large-2402-v1:0 | 32K | $8 | $24 | — | |||
| us-east-1/mistral.mistral-7b-instruct-v0:2bedrock/us-east-1/mistral.mistral-7b-instruct-v0:2 | 32K | $0.15 | $0.2 | — | |||
| Qwen: Qwen3 32Bqwen/qwen3-32b | 40.96K | $0.08 | $0.28 | — | |||
| @cf/meta/llama-3.2-3b-instructcloudflare/@cf/meta/llama-3.2-3b-instruct | 80K | $0.051 | $0.335 | — | |||
| apac.amazon.nova-micro-v1:0bedrock_converse/apac.amazon.nova-micro-v1:0 | 128K | $0.037 | $0.148 | — | |||
| llama3:8bollama/llama3:8b | 8.192K | — | — | — | |||
| apac.amazon.nova-lite-v1:0bedrock_converse/apac.amazon.nova-lite-v1:0 | 300K | $0.063 | $0.252 | — | |||
| llama3:70bollama/llama3:70b | 8.192K | — | — | — | |||
| Qwen: Qwen3 8Bqwen/qwen3-8b | 131.072K | $0.117 | $0.455 | — | |||
| apac.amazon.nova-pro-v1:0bedrock_converse/apac.amazon.nova-pro-v1:0 | 300K | $0.84 | $3.36 | — | |||
| Google: Gemini 2.5 Pro Preview 06-05google/gemini-2.5-pro-preview | 1.04858M | $1.25 | $10 | — | |||
| apac.anthropic.claude-3-5-sonnet-20240620-v1:0bedrock/apac.anthropic.claude-3-5-sonnet-20240620-v1:0 | 200K | $3 | $15 | — | |||
| apac.anthropic.claude-3-5-sonnet-20241022-v2:0bedrock/apac.anthropic.claude-3-5-sonnet-20241022-v2:0 | 200K | $3 | $15 | — | |||
| Thinking Machines: Inkling Small (free)thinkingmachines/inkling-small:free | 1.04858M | Free | Free | — | |||
| Anthropic: Claude Fable 5.1anthropic/claude-fable-5.1 | 1M | $10 | $50 | — | |||
| Qwen: Qwen3.6 35B A3Bqwen/qwen3.6-35b-a3b | 262.144K | $0.1 | $0.9 | — | |||
| apac.anthropic.claude-3-haiku-20240307-v1:0bedrock/apac.anthropic.claude-3-haiku-20240307-v1:0 | 200K | $0.25 | $1.25 | — | |||
| @cf/meta/llama-guard-3-8bcloudflare/@cf/meta/llama-guard-3-8b | 131.072K | $0.484 | $0.03 | — | |||
| apac.anthropic.claude-haiku-4-5-20251001-v1:0bedrock_converse/apac.anthropic.claude-haiku-4-5-20251001-v1:0 | 200K | $1.1 | $5.5 | — | |||
| Google: Gemma 4 31B (batch)google/gemma-4-31b-it:batch | 262.144K | $0.39 | $0.97 | — | |||
| apac.anthropic.claude-3-sonnet-20240229-v1:0bedrock/apac.anthropic.claude-3-sonnet-20240229-v1:0 | 200K | $3 | $15 | — | |||
| apac.anthropic.claude-sonnet-4-20250514-v1:0bedrock_converse/apac.anthropic.claude-sonnet-4-20250514-v1:0 | 1M | $3 | $15 | — | |||
| Qwen2.5 72B Instructqwen/qwen-2.5-72b-instruct | 32.768K | $0.36 | $0.4 | — | |||
| au.anthropic.claude-sonnet-4-5-20250929-v1:0bedrock_converse/au.anthropic.claude-sonnet-4-5-20250929-v1:0 | 200K | $3.3 | $16.5 | — | |||
| command-r-plusazure/command-r-plus | 128K | $3 | $15 | — | |||
| claude-haiku-4-5azure_ai/claude-haiku-4-5 | 200K | $1 | $5 | — | |||
| OpenAI: GPT-6 Astra Proopenai/gpt-6-astra-pro | 1.05M | $10 | $50 | — | |||
| claude-opus-4-5azure_ai/claude-opus-4-5 | 200K | $5 | $25 | — | |||
| @cf/mistral/mistral-7b-instruct-v0.2-loracloudflare/@cf/mistral/mistral-7b-instruct-v0.2-lora | 15K | — | — | — | |||
| claude-opus-4-6azure_ai/claude-opus-4-6 | 1M | $5 | $25 | — | |||
| claude-opus-4-7azure_ai/claude-opus-4-7 | 1M | $5 | $25 | — | |||
| mistralai/Mixtral-8x7B-Instruct-v0.1anyscale/mistralai/mixtral-8x7b-instruct-v0.1 | 16.384K | $0.15 | $0.15 | — | |||
| claude-fable-5azure_ai/claude-fable-5 | 1M | $10 | $50 | — | |||
| claude-opus-5azure_ai/claude-opus-5 | 1M | $5 | $25 | — | |||
| mistralai/Mixtral-8x22B-Instruct-v0.1anyscale/mistralai/mixtral-8x22b-instruct-v0.1 | 65.536K | $0.9 | $0.9 | — |