No provider description is available for this model yet.
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
Command A is an open-weights 111B parameter model with a 256k context window focused on delivering great performance across agentic, multilingual, and coding use cases. Compared to other leading proprietary...
No provider description is available for this model yet.
GPT-5 Chat is designed for advanced, natural, multimodal, and context-aware conversations for enterprise applications.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
DeepSeek V3, a 685B-parameter, mixture-of-experts model, is the latest iteration of the flagship chat model family from the DeepSeek team. It succeeds the [DeepSeek V3](/deepseek/deepseek-chat-v3) model and performs really well...
Llama 4 Maverick 17B Instruct (128E) is a high-capacity multimodal language model from Meta, built on a mixture-of-experts (MoE) architecture with 128 experts and 17 billion active parameters per forward...
Ling-2.6-flash is an instant (instruct) model from inclusionAI with 104B total parameters and 7.4B active parameters, designed for real-world agents that require fast responses, strong execution, and high token efficiency....
No provider description is available for this model yet.
GPT-4o mini Search Preview is a specialized model for web search in Chat Completions. It is trained to understand and execute web search queries.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Llama 4 Scout 17B Instruct (16E) is a mixture-of-experts (MoE) language model developed by Meta, activating 17 billion parameters out of a total of 109B. It supports native multimodal input...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Kimi K2 Instruct is a large-scale Mixture-of-Experts (MoE) language model developed by Moonshot AI, featuring 1 trillion total parameters with 32 billion active per forward pass. It is optimized for...
No provider description is available for this model yet.
No provider description is available for this model yet.
Qwen3-8B is a dense 8.2B parameter causal language model from the Qwen3 series, designed for both reasoning-heavy tasks and efficient dialogue. It supports seamless switching between "thinking" mode for math,...
Gemini 2.5 Pro is Google’s state-of-the-art AI model designed for advanced reasoning, coding, mathematics, and scientific tasks. It employs “thinking” capabilities, enabling it to reason through responses with enhanced accuracy...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
May 28th update to the [original DeepSeek R1](/deepseek/deepseek-r1) Performance on par with [OpenAI o1](/openai/o1), but open-sourced and with fully open reasoning tokens. It's 671B parameters in size, with 37B active...
No provider description is available for this model yet.
Qwen3 Coder Plus is Alibaba's proprietary version of the Open Source Qwen3 Coder 480B A35B. It is a powerful coding agent model specializing in autonomous programming via tool calling and...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Qwen3-235B-A22B-Instruct-2507 is a multilingual, instruction-tuned mixture-of-experts language model based on the Qwen3-235B architecture, with 22B active parameters per forward pass. It is optimized for general-purpose text generation, including instruction following,...
Morph's high-accuracy apply model for complex code edits. ~4,500 tokens/sec with 98% accuracy for precise code transformations. The model requires the prompt to be in the following format: <instruction>{instruction}</instruction> <code>{initial_code}</code>...
Qwen3.6 27B is a dense 27-billion-parameter language model from the Qwen Team at Alibaba, released in April 2026. It features hybrid multimodal capabilities — accepting text, image, and video inputs...
Fast-mode variant of [Opus 4.8](/anthropic/claude-opus-4.8) - identical capabilities with higher output speed at 2x pricing relative to regular Opus 4.8. Learn more in Anthropic's docs: https://platform.claude.com/docs/en/build-with-claude/fast-mode Note: As of September...
GPT-5.6 Luna Pro is the same underlying model as [GPT-5.6 Luna](https://openrouter.ai/openai/gpt-5.6-luna), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. Learn more in OpenAI's docs: https://developers.openai.com/api/docs/guides/reasoning#reasoning-mode
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
GPT-4o Search Previewis a specialized model for web search in Chat Completions. It is trained to understand and execute web search queries.
No provider description is available for this model yet.
Qwen3-Coder-30B-A3B-Instruct is a 30.5B parameter Mixture-of-Experts (MoE) model with 128 experts (8 active per forward pass), designed for advanced code generation, repository-scale understanding, and agentic tool use. Built on the...
Mistral Large 3 2512 is Mistral’s most capable model to date, featuring a sparse mixture-of-experts architecture with 41B active parameters (675B total), and released under the Apache 2.0 license.
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| FW-GLM-5.2-Fastazure_ai/fw-glm-5.2-fast | 1.04858M | $2.1 | $6.6 | — | |||
| Cohere: Command Acohere/command-a | 256K | $2.5 | $10 | — | |||
| Qwen/Qwen3-Coder-Next-FP8together_ai/qwen/qwen3-coder-next-fp8 | 262.144K | $0.5 | $1.2 | — | |||
| OpenAI: GPT-5 Chatopenai/gpt-5-chat | 128K | $1.25 | $10 | — | |||
| us-gov.nvidia.nemotron-nano-12b-v2bedrock_converse/us-gov.nvidia.nemotron-nano-12b-v2 | 128K | $0.24 | $0.72 | — | |||
| deepseek-ai/DeepSeek-R1-0528together_ai/deepseek-ai/deepseek-r1-0528 | 163.84K | $3 | $7 | — | |||
| zai-org/GLM-5.1together_ai/zai-org/glm-5.1 | 202.752K | $1.4 | $4.4 | — | |||
| us-gov.nvidia.nemotron-nano-3-30bbedrock_converse/us-gov.nvidia.nemotron-nano-3-30b | 262.144K | $0.072 | $0.288 | — | |||
| DeepSeek: DeepSeek V3 0324deepseek/deepseek-chat-v3-0324 | 163.84K | $0.25 | $1 | — | |||
| Meta: Llama 4 Maverickmeta-llama/llama-4-maverick | 128K | $0.2 | $0.696 | — | |||
| inclusionAI: Ling-2.6-flashinclusionai/ling-2.6-flash | 262.144K | $0.01 | $0.03 | — | |||
| zai-org/GLM-5together_ai/zai-org/glm-5 | 202.752K | $1 | $3.2 | — | |||
| OpenAI: GPT-4o-mini Search Previewopenai/gpt-4o-mini-search-preview | 128K | $0.15 | $0.6 | — | |||
| MiniMaxAI/MiniMax-M2.7together_ai/minimaxai/minimax-m2.7 | 196.608K | $0.3 | $1.2 | — | |||
| moonshotai/Kimi-K2.5-fp4together_ai/moonshotai/kimi-k2.5-fp4 | 262.144K | $0.5 | $2.8 | — | |||
| FW-GLM-5.2azure_ai/fw-glm-5.2 | 1.04858M | $1.54 | $4.84 | — | |||
| Meta: Llama 4 Scoutmeta-llama/llama-4-scout | 327.68K | $0.1 | $0.3 | — | |||
| moonshotai/Kimi-K2.6together_ai/moonshotai/kimi-k2.6 | 262.144K | $1.2 | $4.5 | — | |||
| deepseek-v4p1-flashfireworks_ai/deepseek-v4p1-flash | 1.04858M | $0.22 | $0.66 | — | |||
| FW-GLM-5.1azure_ai/fw-glm-5.1 | 202.8K | $1.54 | $4.84 | — | |||
| MoonshotAI: Kimi K2 0711moonshotai/kimi-k2 | 131.072K | $0.57 | $2.3 | — | |||
| accounts/fireworks/models/deepseek-v4p1-flashfireworks_ai/accounts/fireworks/models/deepseek-v4p1-flash | 1.04858M | $0.22 | $0.66 | — | |||
| accounts/fireworks/models/kimi-k3fireworks_ai/accounts/fireworks/models/kimi-k3 | 1.04858M | $3 | $15 | — | |||
| Qwen: Qwen3 8Bqwen/qwen3-8b | 131.072K | $0.117 | $0.455 | — | |||
| Google: Gemini 2.5 Pro Preview 06-05google/gemini-2.5-pro-preview | 1.04858M | $1.25 | $10 | — | |||
| deepseek-flashdeepseek/deepseek-flash | 1M | $0.3 | $1.2 | — | |||
| openai/gpt-5.6-sol-proopenrouter/openai/gpt-5.6-sol-pro | 1.05M | $2 | $10 | — | |||
| accounts/fireworks/models/deepseek-v4-flash-0731fireworks_ai/accounts/fireworks/models/deepseek-v4-flash-0731 | 1.04858M | $0.14 | $0.28 | — | |||
| deepseek/deepseek-v4.1-flashopenrouter/deepseek/deepseek-v4.1-flash | 1.04858M | $0.15 | $0.6 | — | |||
| anthropic.claude-sonnet-4-5-20250929-v1:0bedrock_converse/anthropic.claude-sonnet-4-5-20250929-v1:0 | 200K | $3 | $15 | — | |||
| us-gov.anthropic.claude-fable-5-1bedrock_converse/us-gov.anthropic.claude-fable-5-1 | 1M | $12 | $60 | — | |||
| DeepSeek: R1 0528deepseek/deepseek-r1-0528 | 163.84K | $0.5 | $2.15 | — | |||
| anthropic.claude-sonnet-4-20250514-v1:0bedrock_converse/anthropic.claude-sonnet-4-20250514-v1:0 | 1M | $3 | $15 | — | |||
| Qwen: Qwen3 Coder Plusqwen/qwen3-coder-plus | 1M | $0.65 | $3.25 | — | |||
| us-gov.anthropic.claude-opus-5bedrock_converse/us-gov.anthropic.claude-opus-5 | 1M | $6 | $30 | — | |||
| jp.anthropic.claude-sonnet-4-6bedrock_converse/jp.anthropic.claude-sonnet-4-6 | 1M | $3.3 | $16.5 | — | |||
| us-east-1/deepseek.v3.2bedrock/us-east-1/deepseek.v3.2 | 163.84K | $0.62 | $1.85 | — | |||
| Qwen: Qwen3 235B A22B Instruct 2507qwen/qwen3-235b-a22b-2507 | 262.144K | $0.087 | $0.35 | — | |||
| Morph: Morph V3 Largemorph/morph-v3-large | 262.144K | $0.9 | $1.9 | — | |||
| Qwen: Qwen3.6 27Bqwen/qwen3.6-27b | 262.144K | $0.3 | $2 | — | |||
| Anthropic: Claude Opus 4.8 (Fast)anthropic/claude-opus-4.8-fast | 1M | $10 | $50 | — | |||
| OpenAI: GPT-5.6 Luna Proopenai/gpt-5.6-luna-pro | 1.05M | $0.2 | $1.2 | — | |||
| au.anthropic.claude-sonnet-4-6bedrock_converse/au.anthropic.claude-sonnet-4-6 | 1M | $3.3 | $16.5 | — | |||
| eu.anthropic.claude-sonnet-4-6bedrock_converse/eu.anthropic.claude-sonnet-4-6 | 1M | $3.3 | $16.5 | — | |||
| FW-GLM-5azure_ai/fw-glm-5 | 200K | $1.1 | $3.52 | — | |||
| @cf/openai/gpt-oss-120bcloudflare/@cf/openai/gpt-oss-120b | 128K | $0.35 | $0.75 | — | |||
| OpenAI: GPT-4o Search Previewopenai/gpt-4o-search-preview | 128K | $2.5 | $10 | — | |||
| FW-DeepSeek-V4-Proazure_ai/fw-deepseek-v4-pro | 1M | $1.925 | $3.828 | — | |||
| Qwen: Qwen3 Coder 30B A3B Instructqwen/qwen3-coder-30b-a3b-instruct | 262.144K | $0.07 | $0.28 | — | |||
| Mistral: Mistral Large 3 2512 (batch)mistralai/mistral-large-2512:batch | 262.144K | $0.25 | $0.75 | — |