OpenAI o4-mini-high is the same model as [o4-mini](/openai/o4-mini) with reasoning_effort set to high. OpenAI o4-mini is a compact reasoning model in the o-series, optimized for fast, cost-efficient performance while retaining...
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
Qwen3-235B-A22B-Thinking-2507 is a high-performance, open-weight Mixture-of-Experts (MoE) language model optimized for complex reasoning tasks. It activates 22B of its 235B parameters per forward pass and natively supports up to 262,144...
No provider description is available for this model yet.
Laguna M.1 is the flagship coding agent model from [Poolside](https://poolside.ai/), optimized for complex software engineering tasks. Designed for agentic coding workflows, it supports tool calling and reasoning, with a 256K...
*Ling-3.0-flash* is a *124B-parameter Mixture-of-Experts (MoE) model*, with approximately *5.1B parameters activated per token*. The model is designed with *token efficiency and production-scale agentic inference* as key priorities, enabling developers...
No provider description is available for this model yet.
DeepSeek V4 Flash Vision Exp is an experimental vision-enabled version of [DeepSeek V4 Flash 0731](https://openrouter.ai/deepseek/deepseek-v4-flash-0731) from DeepSeek, adding image understanding while matching the base model on text capabilities including agents,...
Gemini 3.6 Flash is a high-efficiency model from Google for coding, agentic workflows, and web and app development. It is designed to produce polished outputs with fewer unnecessary edits and...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
As a 30B-class SOTA model, GLM-4.7-Flash offers a new option that balances performance and efficiency. It is further optimized for agentic coding use cases, strengthening coding capabilities, long-horizon task planning,...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Claude Opus 4.8 is Anthropic's most capable generally available model in the Opus family. It supports text, image, and file inputs with text output, with reasoning support and a 1M-token...
GPT-5.2 Pro is OpenAI’s most advanced model, offering major improvements in agentic coding and long context performance over GPT-5 Pro. It is optimized for complex tasks that require step-by-step reasoning,...
DeepSeek-V3.1 Terminus is an update to [DeepSeek V3.1](/deepseek/deepseek-chat-v3.1) that maintains the model's original capabilities while addressing issues reported by users, including language consistency and agent capabilities, further optimizing the model's...
Qwen3-Next-80B-A3B-Thinking is a reasoning-first chat model in the Qwen3-Next line that outputs structured “thinking” traces by default. It’s designed for hard multi-step problems; math proofs, code synthesis/debugging, logic, and agentic...
gpt-oss-20b is an open-weight 21B parameter model released by OpenAI under the Apache 2.0 license. It uses a Mixture-of-Experts (MoE) architecture with 3.6B active parameters per forward pass, optimized for...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
NVIDIA Nemotron 3 Ultra is an open frontier-reasoning and orchestration model from NVIDIA, with 55B active parameters out of 550B total (MoE). Built on a hybrid Transformer-Mamba mixture-of-experts architecture, it...
No provider description is available for this model yet.
Qwen3 Coder Plus is Alibaba's proprietary version of the Open Source Qwen3 Coder 480B A35B. It is a powerful coding agent model specializing in autonomous programming via tool calling and...
No provider description is available for this model yet.
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| OpenAI: o4 Mini Highopenai/o4-mini-high | 200K | $1.1 | $4.4 | — | |||
| Qwen: Qwen3 235B A22B Thinking 2507qwen/qwen3-235b-a22b-thinking-2507 | 131.072K | $0.23 | $2.3 | — | |||
| Qwen/Qwen3-Max-Thinkingdeepinfra/qwen/qwen3-max-thinking | 256K | $1.2 | $6 | — | |||
| Poolside: Laguna M.1 (free)poolside/laguna-m.1:free | 262.144K | Free | Free | — | |||
| inclusionAI: Ling 3.0 Flashinclusionai/ling-3.0-flash | 262.144K | $0.021 | $0.063 | — | |||
| ap-southeast-2/minimax.minimax-m2.5bedrock/ap-southeast-2/minimax.minimax-m2.5 | 1M | $0.309 | $1.236 | — | |||
| DeepSeek: DeepSeek V4 Flash Vision Exp (batch)deepseek/deepseek-v4-flash-vision-exp:batch | 1.04858M | $0.11 | $0.33 | — | |||
| Google: Gemini 3.6 Flash (batch)google/gemini-3.6-flash:batch | 1.04858M | $0.375 | $1.875 | — | |||
| openai/gpt-3.5-turbo-instructopenrouter/openai/gpt-3.5-turbo-instruct | 4.095K | $1.5 | $2 | — | |||
| ap-southeast-3/deepseek.v3.2bedrock/ap-southeast-3/deepseek.v3.2 | 163.84K | $0.74 | $2.22 | — | |||
| ap-southeast-3/minimax.minimax-m2.1bedrock/ap-southeast-3/minimax.minimax-m2.1 | 196K | $0.36 | $1.44 | — | |||
| ap-southeast-3/minimax.minimax-m2.5bedrock/ap-southeast-3/minimax.minimax-m2.5 | 1M | $0.36 | $1.44 | — | |||
| openai/gpt-4-turbo-previewopenrouter/openai/gpt-4-turbo-preview | 128K | $10 | $30 | — | |||
| ap-southeast-3/moonshotai.kimi-k2.5bedrock/ap-southeast-3/moonshotai.kimi-k2.5 | 262.144K | $0.72 | $3.6 | — | |||
| ap-southeast-3/qwen.qwen3-coder-nextbedrock/ap-southeast-3/qwen.qwen3-coder-next | 262.144K | $0.6 | $1.44 | — | |||
| ca-central-1/meta.llama3-70b-instruct-v1:0bedrock/ca-central-1/meta.llama3-70b-instruct-v1:0 | 8.192K | $3.05 | $4.03 | — | |||
| Z.ai: GLM 4.7 Flashz-ai/glm-4.7-flash | 131.072K | $0.061 | $0.4 | — | |||
| ca-central-1/meta.llama3-8b-instruct-v1:0bedrock/ca-central-1/meta.llama3-8b-instruct-v1:0 | 8.192K | $0.35 | $0.69 | — | |||
| eu-north-1/deepseek.v3.2bedrock/eu-north-1/deepseek.v3.2 | 163.84K | $0.74 | $2.22 | — | |||
| eu-north-1/minimax.minimax-m2.1bedrock/eu-north-1/minimax.minimax-m2.1 | 196K | $0.36 | $1.44 | — | |||
| eu-north-1/minimax.minimax-m2.5bedrock/eu-north-1/minimax.minimax-m2.5 | 1M | $0.36 | $1.44 | — | |||
| eu-north-1/moonshotai.kimi-k2.5bedrock/eu-north-1/moonshotai.kimi-k2.5 | 262.144K | $0.72 | $3.6 | — | |||
| eu-central-1/1-month-commitment/anthropic.claude-instant-v1bedrock/eu-central-1/1-month-commitment/anthropic.claude-instant-v1 | 100K | — | — | — | |||
| eu-central-1/1-month-commitment/anthropic.claude-v1bedrock/eu-central-1/1-month-commitment/anthropic.claude-v1 | 100K | — | — | — | |||
| eu-central-1/1-month-commitment/anthropic.claude-v2:1bedrock/eu-central-1/1-month-commitment/anthropic.claude-v2:1 | 100K | — | — | — | |||
| Anthropic: Claude Opus 4.8 (batch)anthropic/claude-opus-4.8:batch | 1M | $2.5 | $12.5 | — | |||
| OpenAI: GPT-5.2 Pro (batch)openai/gpt-5.2-pro:batch | 400K | $10.5 | $84 | — | |||
| DeepSeek: DeepSeek V3.1 Terminusdeepseek/deepseek-v3.1-terminus | 131.072K | $0.27 | $1 | — | |||
| Qwen: Qwen3 Next 80B A3B Thinkingqwen/qwen3-next-80b-a3b-thinking | 262.144K | $0.15 | $1.2 | — | |||
| OpenAI: gpt-oss-20b (batch)openai/gpt-oss-20b:batch | 131.072K | $0.05 | $0.2 | — | |||
| eu-central-1/6-month-commitment/anthropic.claude-instant-v1bedrock/eu-central-1/6-month-commitment/anthropic.claude-instant-v1 | 100K | — | — | — | |||
| openai/gpt-4-turboopenrouter/openai/gpt-4-turbo | 128K | $10 | $30 | — | |||
| eu-central-1/6-month-commitment/anthropic.claude-v1bedrock/eu-central-1/6-month-commitment/anthropic.claude-v1 | 100K | — | — | — | |||
| eu-central-1/6-month-commitment/anthropic.claude-v2:1bedrock/eu-central-1/6-month-commitment/anthropic.claude-v2:1 | 100K | — | — | — | |||
| eu-central-1/anthropic.claude-instant-v1bedrock/eu-central-1/anthropic.claude-instant-v1 | 100K | $2.48 | $8.38 | — | |||
| eu-central-1/anthropic.claude-v1bedrock/eu-central-1/anthropic.claude-v1 | 100K | $8 | $24 | — | |||
| eu-central-1/anthropic.claude-v2:1bedrock/eu-central-1/anthropic.claude-v2:1 | 100K | $8 | $24 | — | |||
| @cf/zai-org/glm-5.2cloudflare/@cf/zai-org/glm-5.2 | 262.144K | $1.4 | $4.4 | — | |||
| google/gemma-2-27b-itopenrouter/google/gemma-2-27b-it | 8.192K | $0.65 | $0.65 | — | |||
| eu-central-1/minimax.minimax-m2.1bedrock/eu-central-1/minimax.minimax-m2.1 | 196K | $0.36 | $1.44 | — | |||
| eu-central-1/minimax.minimax-m2.5bedrock/eu-central-1/minimax.minimax-m2.5 | 1M | $0.36 | $1.44 | — | |||
| eu-central-1/qwen.qwen3-coder-nextbedrock/eu-central-1/qwen.qwen3-coder-next | 262.144K | $0.6 | $1.44 | — | |||
| eu-west-1/meta.llama3-70b-instruct-v1:0bedrock/eu-west-1/meta.llama3-70b-instruct-v1:0 | 8.192K | $2.86 | $3.78 | — | |||
| eu-west-1/meta.llama3-8b-instruct-v1:0bedrock/eu-west-1/meta.llama3-8b-instruct-v1:0 | 8.192K | $0.32 | $0.65 | — | |||
| eu-west-1/minimax.minimax-m2.1bedrock/eu-west-1/minimax.minimax-m2.1 | 196K | $0.36 | $1.44 | — | |||
| eu-west-1/minimax.minimax-m2.5bedrock/eu-west-1/minimax.minimax-m2.5 | 1M | $0.36 | $1.44 | — | |||
| NVIDIA: Nemotron 3 Ultra (batch)nvidia/nemotron-3-ultra-550b-a55b:batch | 512.288K | $0.6 | $3.6 | — | |||
| grok-4.20xai/grok-4.20 | 1M | $1.25 | $2.5 | — | |||
| Qwen: Qwen3 Coder Plusqwen/qwen3-coder-plus | 1M | $0.65 | $3.25 | — | |||
| openai/gpt-4o-mini-2024-07-18openrouter/openai/gpt-4o-mini-2024-07-18 | 128K | $0.15 | $0.6 | — |