Qwen3-VL-235B-A22B Thinking is a multimodal model that unifies strong text generation with visual understanding across images and video. The Thinking model is optimized for multimodal reasoning in STEM and math....
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
No provider description is available for this model yet.
Gemini 3.5 Flash is Google's high-efficiency multimodal model, bringing near-Pro level coding and reasoning at Flash-tier cost and speed. It is highly optimized for coding proficiency and parallel agentic execution...
No provider description is available for this model yet.
Qwen3 Coder Plus is Alibaba's proprietary version of the Open Source Qwen3 Coder 480B A35B. It is a powerful coding agent model specializing in autonomous programming via tool calling and...
No provider description is available for this model yet.
Relace Apply 3 is a specialized code-patching LLM that merges AI-suggested edits straight into your source files. It can apply updates from GPT-4o, Claude, and others into your files at...
No provider description is available for this model yet.
GPT-4o ("o" for "omni") is OpenAI's latest AI model, supporting both text and image inputs with text outputs. It maintains the intelligence level of [GPT-4 Turbo](/models/openai/gpt-4-turbo) while being twice as...
*Ling-3.0-flash* is a *124B-parameter Mixture-of-Experts (MoE) model*, with approximately *5.1B parameters activated per token*. The model is designed with *token efficiency and production-scale agentic inference* as key priorities, enabling developers...
Gemini 3.8 Flash is Google's most intelligent Flash model with significant gains from 3.7 Flash across software engineering, agentic tasks, and multi-step reasoning.
DeepSeek-V3.1 Terminus is an update to [DeepSeek V3.1](/deepseek/deepseek-chat-v3.1) that maintains the model's original capabilities while addressing issues reported by users, including language consistency and agent capabilities, further optimizing the model's...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
GPT-6 Astra Pro is the same underlying model as [GPT-6 Astra](https://openrouter.ai/openai/gpt-6-astra), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. Learn more in OpenAI's docs: https://developers.openai.com/api/docs/guides/reasoning#reasoning-mode
OpenAI o4-mini is a compact reasoning model in the o-series, optimized for fast, cost-efficient performance while retaining strong multimodal and agentic capabilities. It supports tool use and demonstrates competitive reasoning...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
GPT-5.2 is the latest frontier-grade model in the GPT-5 series, offering stronger agentic and long context perfomance compared to GPT-5.1. It uses adaptive reasoning to allocate computation dynamically, responding quickly...
No provider description is available for this model yet.
Gemini 3.7 Flash is a multimodal model from Google for fast agentic workflows, coding, and complex multi-step reasoning. It is designed for tasks that require responsive performance and reliable multi-step...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
MiniMax-M2 is a compact, high-efficiency large language model optimized for end-to-end coding and agentic workflows. With 10 billion activated parameters (230 billion total), it delivers near-frontier intelligence across general reasoning,...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| Qwen: Qwen3 VL 235B A22B Thinkingqwen/qwen3-vl-235b-a22b-thinking | 131.072K | $0.4 | $4 | — | |||
| zai-org/glm-5-turbonovita/zai-org/glm-5-turbo | 202.8K | $1.2 | $4 | — | |||
| Google: Gemini 3.5 Flash (batch)google/gemini-3.5-flash:batch | 1.04858M | $0.75 | $4.5 | — | |||
| anthropic/claude-opus-4.6openrouter/anthropic/claude-opus-4.6 | 1M | $5 | $25 | — | |||
| Qwen: Qwen3 Coder Plusqwen/qwen3-coder-plus | 1M | $0.65 | $3.25 | — | |||
| google/gemma-4-31b-itnovita/google/gemma-4-31b-it | 262.144K | $0.14 | $0.4 | — | |||
| Relace: Relace Apply 3relace/relace-apply-3 | 256K | $0.85 | $1.25 | — | |||
| nvidia/NVIDIA-NemotronLabs-AI-for-Media-Sports-Tennisnvidia/NVIDIA-NemotronLabs-AI-for-Media-Sports-Tennis | Not documented | — | — | — | |||
| OpenAI: GPT-4o (batch)openai/gpt-4o:batch | 128K | $1.25 | $5 | — | |||
| inclusionAI: Ling 3.0 Flashinclusionai/ling-3.0-flash | 262.144K | $0.021 | $0.063 | — | |||
| Google: Gemini 3.8 Flash (batch)google/gemini-3.8-flash:batch | 1.04858M | $0.375 | $1.875 | — | |||
| DeepSeek: DeepSeek V3.1 Terminusdeepseek/deepseek-v3.1-terminus | 131.072K | $0.27 | $1 | — | |||
| anthropic/claude-opus-4.5openrouter/anthropic/claude-opus-4.5 | 200K | $5 | $25 | — | |||
| eu.anthropic.claude-3-5-sonnet-20240620-v1:0bedrock/eu.anthropic.claude-3-5-sonnet-20240620-v1:0 | 200K | $3 | $15 | — | |||
| Llama-4-Scout-17B-16E-Instructazure_ai/llama-4-scout-17b-16e-instruct | 10M | $0.2 | $0.78 | — | |||
| OpenAI: GPT-6 Astra Pro (batch)openai/gpt-6-astra-pro:batch | 1.05M | $5 | $25 | — | |||
| OpenAI: o4 Mini (batch)openai/o4-mini:batch | 200K | $0.55 | $2.2 | — | |||
| google/gemma-4-26b-a4b-itnovita/google/gemma-4-26b-a4b-it | 262.144K | $0.13 | $0.4 | — | |||
| anthropic/claude-sonnet-4.6openrouter/anthropic/claude-sonnet-4.6 | 1M | $3 | $15 | — | |||
| zai-org/glm-5v-turbonovita/zai-org/glm-5v-turbo | 204.8K | $1.2 | $4 | — | |||
| global.twelvelabs.pegasus-1-2-v1:0bedrock/global.twelvelabs.pegasus-1-2-v1:0 | Not documented | — | $7.5 | — | |||
| gpt-chat-latestazure_ai/gpt-chat-latest | 272K | $5 | $30 | — | |||
| model-routerazure_ai/model-router | 200K | $0.14 | — | — | |||
| cohere-command-aazure_ai/cohere-command-a | 131.072K | $2.5 | $10 | — | |||
| xai/grok-4.3vertex_ai/xai/grok-4.3 | 200K | $1.25 | $2.5 | — | |||
| grok-4-20-reasoningazure_ai/grok-4-20-reasoning | 262K | $1.25 | $2.5 | — | |||
| xai/grok-4.6vertex_ai/xai/grok-4.6 | 524.288K | $2 | $6 | — | |||
| grok-4-20-non-reasoningazure_ai/grok-4-20-non-reasoning | 262K | $1.25 | $2.5 | — | |||
| us.openai.gpt-6-astrabedrock_converse/us.openai.gpt-6-astra | 1.05M | $11 | $55 | — | |||
| global.openai.gpt-6-astrabedrock_converse/global.openai.gpt-6-astra | 1.05M | $10 | $50 | — | |||
| lyria-3.5gemini/lyria-3.5 | 1.04858M | — | — | — | |||
| OpenAI: GPT-5.2 (batch)openai/gpt-5.2:batch | 400K | $0.875 | $7 | — | |||
| anthropic/claude-sonnet-4openrouter/anthropic/claude-sonnet-4 | 1M | $3 | $15 | — | |||
| Google: Gemini 3.7 Flash (batch)google/gemini-3.7-flash:batch | 1.04858M | $0.375 | $1.875 | — | |||
| gpt-6-astraazure/gpt-6-astra | 922K | $10 | $50 | — | |||
| eu.anthropic.claude-haiku-4-5-20251001-v1:0bedrock_converse/eu.anthropic.claude-haiku-4-5-20251001-v1:0 | 200K | $1.1 | $5.5 | — | |||
| us/gpt-6-astraazure/us/gpt-6-astra | 922K | $11 | $55 | — | |||
| Codestral-2501azure_ai/codestral-2501 | 256K | $0.3 | $0.9 | — | |||
| MiniMax: MiniMax M2minimax/minimax-m2 | 204.8K | $0.255 | $1.02 | — | |||
| minimax/minimax-m2.7-highspeednovita/minimax/minimax-m2.7-highspeed | 204.8K | $0.6 | $2.4 | — | |||
| FW-Nemotron-Lightning-3.5-30B-A3Bazure_ai/fw-nemotron-lightning-3.5-30b-a3b | 262.144K | $0.06 | $0.22 | — | |||
| MAI-Thinking-1azure_ai/mai-thinking-1 | 256K | $2 | $8 | — | |||
| anthropic/claude-opus-4.1openrouter/anthropic/claude-opus-4.1 | 200K | $15 | $75 | — | |||
| grok-4.6azure_ai/grok-4.6 | 200K | $2 | $6 | — | |||
| databricks-claude-fable-5-1databricks/databricks-claude-fable-5-1 | 1M | $10 | $50 | — | |||
| databricks-gemini-3-1-flash-imagedatabricks/databricks-gemini-3-1-flash-image | 131.072K | — | — | — | |||
| zai-org/glm-5.1novita/zai-org/glm-5.1 | 204.8K | $1.38 | $4.4 | — | |||
| databricks-gemini-3-pro-imagedatabricks/databricks-gemini-3-pro-image | 65.536K | — | — | — | |||
| databricks-gemini-3-8-flashdatabricks/databricks-gemini-3-8-flash | 1.04858M | — | — | — | |||
| moonshotai/kimi-k2.6novita/moonshotai/kimi-k2.6 | 262.144K | $0.8 | $3.4 | — |