Muse Spark 1.2 contributor tier is a reasoning model from Meta designed for developers who want to start building at an even lower cost. It’s meaningfully cheaper than Muse Spark...
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
No provider description is available for this model yet.
Aion-2.0 is a variant of DeepSeek V3.2 optimized for immersive roleplaying and storytelling. It is particularly strong at introducing tension, crises, and conflict into stories, making narratives feel more engaging....
This model always redirects to the latest Grok model from xAI.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
The Qwen3.5 Series 35B-A3B is a native vision-language model designed with a hybrid architecture that integrates linear attention mechanisms and a sparse mixture-of-experts model, achieving higher inference efficiency. Its overall...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
UI-TARS-1.5 is a multimodal vision-language agent optimized for GUI-based environments, including desktop interfaces, web browsers, mobile systems, and games. Built by ByteDance, it builds upon the UI-TARS framework with reinforcement...
The Qwen3.5 122B-A10B native vision-language model is built on a hybrid architecture that integrates a linear attention mechanism with a sparse mixture-of-experts model, achieving higher inference efficiency. In terms of...
KAT-Coder-Pro V2.5 is a flagship-level Agentic Coding model that can directly hand over an entire issue or an entire business workflow to it, allowing it to autonomously locate and make...
*Ling-3.0-flash* is a *124B-parameter Mixture-of-Experts (MoE) model*, with approximately *5.1B parameters activated per token*. The model is designed with *token efficiency and production-scale agentic inference* as key priorities, enabling developers...
No provider description is available for this model yet.
Qwen3-Max-Thinking is the flagship reasoning model in the Qwen3 series, designed for high-stakes cognitive tasks that require deep, multi-step reasoning. By significantly scaling model capacity and reinforcement learning compute, it...
Sonnet 4.6 is Anthropic's most capable Sonnet-class model yet, with frontier performance across coding, agents, and professional work. It excels at iterative development, complex codebase navigation, end-to-end project management with...
Mistral Small 4 is the next major release in the Mistral Small family, unifying the capabilities of several flagship Mistral models into a single system. It combines strong reasoning from...
This is Mistral AI's flagship model, Mistral Large 2 (version `mistral-large-2407`). It's a proprietary weights-available model and excels at reasoning, code, JSON, chat, and more. Read the launch announcement [here](https://mistral.ai/news/mistral-large-2407/)....
The Qwen3.5 27B native vision-language Dense model incorporates a linear attention mechanism, delivering fast response times while balancing inference speed and performance. Its overall capabilities are comparable to those of...
GPT-5.4 mini brings the core capabilities of GPT-5.4 to a faster, more efficient model optimized for high-throughput workloads. It supports text and image inputs with strong performance across reasoning, coding,...
GPT-5.5 is OpenAI’s frontier model designed for complex professional workloads, building on GPT-5.4 with stronger reasoning, higher reliability, and improved token efficiency on hard tasks. It features a 1M+ token...
Claude 3 Haiku is Anthropic's fastest and most compact model for near-instant responsiveness. Quick and accurate targeted performance. See the launch announcement and benchmark results [here](https://www.anthropic.com/news/claude-3-haiku) #multimodal
MiniMax-M3 is a multimodal foundation model from MiniMax. It supports text, image, and video inputs with text output, a 1M-token context window, and is suited for long-horizon agentic work, coding,...
No provider description is available for this model yet.
GPT-4o mini is OpenAI's newest model after [GPT-4 Omni](/models/openai/gpt-4o), supporting both text and image inputs with text outputs. As their most advanced small model, it is many multiples more affordable...
Qwen3.5-9B is a multimodal foundation model from the Qwen3.5 family, designed to deliver strong reasoning, coding, and visual understanding in an efficient 9B-parameter architecture. It uses a unified vision-language design...
Auto Router (Beta) is a task-aware router from OpenRouter. It classifies each request, then routes it the [most popular model](/rankings#task-spend) for that task based on aggregate spend, filtered by your...
No provider description is available for this model yet.
No provider description is available for this model yet.
A 12B parameter model with a 128k token context length built by Mistral in collaboration with NVIDIA. The model is multilingual, supporting English, French, German, Spanish, Italian, Portuguese, Chinese, Japanese,...
No provider description is available for this model yet.
No provider description is available for this model yet.
Meta's latest class of model (Llama 3.1) launched with a variety of sizes & flavors. This 8B instruct-tuned version is fast and efficient. It has demonstrated strong performance compared to...
NVIDIA Nemotron 3 Super is a 120B-parameter open hybrid MoE model, activating just 12B parameters for maximum compute efficiency and accuracy in complex multi-agent applications. Built on a hybrid Mamba-Transformer...
No provider description is available for this model yet.
No provider description is available for this model yet.
Meta's latest class of model (Llama 3.1) launched with a variety of sizes & flavors. This 70B instruct-tuned version is optimized for high quality dialogue usecases. It has demonstrated strong...
No provider description is available for this model yet.
No provider description is available for this model yet.
Hermes 3 is a generalist language model with many improvements over Hermes 2, including advanced agentic capabilities, much better roleplaying, reasoning, multi-turn conversation, long context coherence, and improvements across the...
KAT-Coder-Pro V2 is the latest high-performance model in KwaiKAT’s KAT-Coder series, designed for complex enterprise-grade software engineering and SaaS integration. It builds on the agentic coding strengths of earlier versions,...
KAT-Coder-Air V2.5 is a flagship-level Agentic Coding model that can directly hand over an entire issue or an entire business workflow to it, allowing it to autonomously locate and make...
Rocinante 12B is designed for engaging storytelling and rich prose. Early testers have reported: - Expanded vocabulary with unique and expressive word choices - Enhanced creativity for vivid narratives -...
Hermes 4 70B is a hybrid reasoning model from Nous Research, built on Meta-Llama-3.1-70B. It introduces the same hybrid mode as the larger 405B release, allowing the model to either...
This model always redirects to the latest model in the DeepSeek V4 Flash family.
No provider description is available for this model yet.
No provider description is available for this model yet.
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| Meta: Muse Spark 1.2 Contributormeta/muse-spark-1.2-contributor | 1.04858M | $0.1 | $0.2 | — | |||
| sa-east-1/deepseek.v3.2bedrock/sa-east-1/deepseek.v3.2 | 163.84K | $0.74 | $2.22 | — | |||
| AionLabs: Aion-2.0aion-labs/aion-2.0 | 131.072K | $0.8 | $1.6 | — | |||
| xAI: Grok Latest~x-ai/grok-latest | 500K | $2 | $6 | — | |||
| eu-central-1/minimax.minimax-m2.1bedrock/eu-central-1/minimax.minimax-m2.1 | 196K | $0.36 | $1.44 | — | |||
| @cf/zai-org/glm-5.2cloudflare/@cf/zai-org/glm-5.2 | 262.144K | $1.4 | $4.4 | — | |||
| eu-central-1/anthropic.claude-v2:1bedrock/eu-central-1/anthropic.claude-v2:1 | 100K | $8 | $24 | — | |||
| Qwen: Qwen3.5-35B-A3Bqwen/qwen3.5-35b-a3b | 256K | $0.312 | $1.25 | — | |||
| eu-central-1/anthropic.claude-v1bedrock/eu-central-1/anthropic.claude-v1 | 100K | $8 | $24 | — | |||
| eu-central-1/anthropic.claude-instant-v1bedrock/eu-central-1/anthropic.claude-instant-v1 | 100K | $2.48 | $8.38 | — | |||
| eu-central-1/6-month-commitment/anthropic.claude-v2:1bedrock/eu-central-1/6-month-commitment/anthropic.claude-v2:1 | 100K | — | — | — | |||
| eu-central-1/6-month-commitment/anthropic.claude-v1bedrock/eu-central-1/6-month-commitment/anthropic.claude-v1 | 100K | — | — | — | |||
| ByteDance: UI-TARS 7B bytedance/ui-tars-1.5-7b | 128K | $0.1 | $0.2 | — | |||
| Qwen: Qwen3.5-122B-A10Bqwen/qwen3.5-122b-a10b | 262.144K | $0.26 | $2.08 | — | |||
| Kwaipilot: KAT-Coder-Pro V2.5kwaipilot/kat-coder-pro-v2.5 | 262.144K | $0.74 | $2.96 | — | |||
| Ling-3.0-flash (free)inclusionai/ling-3.0-flash:free | 262.144K | Free | Free | — | |||
| eu-central-1/6-month-commitment/anthropic.claude-instant-v1bedrock/eu-central-1/6-month-commitment/anthropic.claude-instant-v1 | 100K | — | — | — | |||
| Qwen: Qwen3 Max Thinkingqwen/qwen3-max-thinking | 262.144K | $0.78 | $3.9 | — | |||
| Anthropic: Claude Sonnet 4.6 (batch)anthropic/claude-sonnet-4.6:batch | 1M | $1.5 | $7.5 | — | |||
| Mistral: Mistral Small 4 (batch)mistralai/mistral-small-2603:batch | 262.144K | $0.075 | $0.3 | — | |||
| Mistral Largemistralai/mistral-large | 128K | $2 | $6 | — | |||
| Qwen: Qwen3.5-27Bqwen/qwen3.5-27b | 262.144K | $0.195 | $1.56 | — | |||
| OpenAI: GPT-5.4 Mini (batch)openai/gpt-5.4-mini:batch | 400K | $0.375 | $2.25 | — | |||
| OpenAI: GPT-5.5 (batch)openai/gpt-5.5:batch | 1.05M | $2.5 | $15 | — | |||
| Anthropic: Claude 3 Haikuanthropic/claude-3-haiku | 200K | $0.25 | $1.25 | — | |||
| MiniMax: MiniMax M3minimax/minimax-m3 | 524.288K | $0.3 | $1.2 | — | |||
| eu-central-1/1-month-commitment/anthropic.claude-v2:1bedrock/eu-central-1/1-month-commitment/anthropic.claude-v2:1 | 100K | — | — | — | |||
| OpenAI: GPT-4o-mini (2024-07-18)openai/gpt-4o-mini-2024-07-18 | 128K | $0.15 | $0.6 | — | |||
| Qwen: Qwen3.5-9Bqwen/qwen3.5-9b | 262.144K | $0.1 | $0.15 | — | |||
| Auto Router (Beta)openrouter/auto-beta | 2M | — | — | — | |||
| eu-central-1/1-month-commitment/anthropic.claude-v1bedrock/eu-central-1/1-month-commitment/anthropic.claude-v1 | 100K | — | — | — | |||
| eu-central-1/1-month-commitment/anthropic.claude-instant-v1bedrock/eu-central-1/1-month-commitment/anthropic.claude-instant-v1 | 100K | — | — | — | |||
| Mistral: Mistral Nemomistralai/mistral-nemo | 131.072K | $0.019 | $0.03 | — | |||
| eu-north-1/moonshotai.kimi-k2.5bedrock/eu-north-1/moonshotai.kimi-k2.5 | 262.144K | $0.72 | $3.6 | — | |||
| eu-north-1/minimax.minimax-m2.5bedrock/eu-north-1/minimax.minimax-m2.5 | 1M | $0.36 | $1.44 | — | |||
| Meta: Llama 3.1 8B Instructmeta-llama/llama-3.1-8b-instruct | 131.072K | $0.05 | $0.08 | — | |||
| NVIDIA: Nemotron 3 Super (free)nvidia/nemotron-3-super-120b-a12b:free | 262.144K | Free | Free | — | |||
| eu-north-1/minimax.minimax-m2.1bedrock/eu-north-1/minimax.minimax-m2.1 | 196K | $0.36 | $1.44 | — | |||
| eu-north-1/deepseek.v3.2bedrock/eu-north-1/deepseek.v3.2 | 163.84K | $0.74 | $2.22 | — | |||
| Meta: Llama 3.1 70B Instructmeta-llama/llama-3.1-70b-instruct | 131.072K | $0.72 | $0.72 | — | |||
| ap-southeast-3/qwen.qwen3-coder-nextbedrock/ap-southeast-3/qwen.qwen3-coder-next | 262.144K | $0.6 | $1.44 | — | |||
| ap-southeast-3/moonshotai.kimi-k2.5bedrock/ap-southeast-3/moonshotai.kimi-k2.5 | 262.144K | $0.72 | $3.6 | — | |||
| Nous: Hermes 3 405B Instructnousresearch/hermes-3-llama-3.1-405b | 131.072K | $1 | $1 | — | |||
| Kwaipilot: KAT-Coder-Pro V2kwaipilot/kat-coder-pro-v2 | 262.144K | $0.3 | $1.2 | — | |||
| Kwaipilot: KAT-Coder-Air V2.5kwaipilot/kat-coder-air-v2.5 | 256K | $0.15 | $0.6 | — | |||
| TheDrummer: Rocinante 12Bthedrummer/rocinante-12b | 65.536K | $0.25 | $0.5 | — | |||
| Nous: Hermes 4 70Bnousresearch/hermes-4-70b | 131.072K | $0.13 | $0.4 | — | |||
| DeepSeek V4 Flash Latest~deepseek/deepseek-v4-flash-latest | 1.04858M | $0.03 | $0.07 | — | |||
| ap-southeast-3/minimax.minimax-m2.5bedrock/ap-southeast-3/minimax.minimax-m2.5 | 1M | $0.36 | $1.44 | — | |||
| ap-southeast-3/minimax.minimax-m2.1bedrock/ap-southeast-3/minimax.minimax-m2.1 | 196K | $0.36 | $1.44 | — |