2,930 models

No provider description is available for this model yet.

qwen_ai_platform/qwen3-max 258.048K context Input not listed Output not listed

No provider description is available for this model yet.

qwen_ai_platform/qwen3-max-2026-01-23 258.048K context Input not listed Output not listed

GLM 5.2 is a large-scale reasoning model from Z.ai. It supports text input and output with a 1M-token context window, and is suited for long-horizon agent workflows, project-level software engineering,...

z-ai/glm-5.2:batch 1.04858M context $0.7/M input $2.2/M output

GLM-5.3 is a large-scale reasoning model from Z.ai, built for complex software engineering and long-horizon agent tasks. It supports text input and output with a 1M-token context window, and improves...

z-ai/glm-5.3 1.04858M context $1.4/M input $4.4/M output

No provider description is available for this model yet.

qwen_ai_platform/qwen3-vl-plus 260.096K context Input not listed Output not listed

No provider description is available for this model yet.

qwen_ai_platform/qwen3.5-plus 991.808K context Input not listed Output not listed

No provider description is available for this model yet.

qwen_ai_platform/qwen3.7-max 991.808K context $2.5/M input $7.5/M output

No provider description is available for this model yet.

qwen_ai_platform/qwen3.7-plus 991.808K context Input not listed Output not listed

No provider description is available for this model yet.

qwen_ai_platform/qwen3.8-max 991.808K context $2/M input $6/M output

No provider description is available for this model yet.

friendliai/zai-org/glm-5.3-flash 1.04858M context $0.15/M input $0.5/M output

No provider description is available for this model yet.

friendliai/zai-org/glm-5.3 1.04858M context $1.26/M input $3.96/M output

No provider description is available for this model yet.

gigachat/gigachat-2 128K context Input not listed Output not listed

No provider description is available for this model yet.

zai/glm-5.2 1M context $1.4/M input $4.4/M output

No provider description is available for this model yet.

together_ai/qwen/qwen3.8-flash 1M context $0.15/M input $0.47/M output

No provider description is available for this model yet.

cerebras/gemma-4-31b 131.072K context $0.99/M input $1.49/M output

MiniMax-M3 is a multimodal foundation model from MiniMax. It supports text, image, and video inputs with text output, a 1M-token context window, and is suited for long-horizon agentic work, coding,...

minimax/minimax-m3:free 1.04858M context Free input Free output

Sonnet 5 is Anthropic's most capable Sonnet-class model, with frontier performance across coding, agents, and professional work. It supports adaptive thinking with selectable reasoning effort levels (low, medium, high, max,...

anthropic/claude-sonnet-5:batch 1M context $1/M input $5/M output

Muse Glimmer 30B is a dense, open-weight multimodal model from Meta Superintelligence Labs, distilled from Muse Spark and optimized for autonomous agents on consumer hardware. It is suited for long-horizon...

meta/muse-glimmer-30b:batch 131.072K context $0.175/M input $0.75/M output

A balanced model in the Ministral 3 family, Ministral 3 8B is a powerful, efficient tiny language model with vision capabilities.

mistralai/ministral-8b-2512:batch 262.144K context $0.075/M input $0.075/M output

Qwen3-VL-32B-Instruct is a large-scale multimodal vision-language model designed for high-precision understanding and reasoning across text, images, and video. With 32 billion parameters, it combines deep visual perception with advanced text...

qwen/qwen3-vl-32b-instruct 131.072K context $0.104/M input $0.416/M output

MiniMax-M3 is a multimodal foundation model from MiniMax. It supports text, image, and video inputs with text output, a 1M-token context window, and is suited for long-horizon agentic work, coding,...

minimax/minimax-m3:batch 524.288K context $0.3/M input $1.2/M output

GPT-6 Astra is OpenAI's flagship model for demanding end-to-end work. It is suited for advanced analysis, software engineering, deep research, scientific work, and document creation, with particular strengths in long-horizon...

openai/gpt-6-astra:batch 1.05M context $5/M input $25/M output

Gemini 3.8 Flash is Google's most intelligent Flash model with significant gains from 3.7 Flash across software engineering, agentic tasks, and multi-step reasoning.

google/gemini-3.8-flash:batch 1.04858M context $0.375/M input $1.875/M output

DeepSeek V4 Flash Vision Exp is an experimental vision-enabled version of [DeepSeek V4 Flash 0731](https://openrouter.ai/deepseek/deepseek-v4-flash-0731) from DeepSeek, adding image understanding while matching the base model on text capabilities including agents,...

deepseek/deepseek-v4-flash-vision-exp:batch 1.04858M context $0.11/M input $0.33/M output

GPT-5.6 Luna is a fast, cost-efficient model in OpenAI's GPT-5.6 series. It is suited for high-volume, latency-sensitive tasks such as chat, classification, and lightweight agentic workflows, providing capable reasoning for...

openai/gpt-5.6-luna:batch 1.05M context $0.1/M input $0.6/M output

GLM-5.3 is a large-scale reasoning model from Z.ai, built for complex software engineering and long-horizon agent tasks. It supports text input and output with a 1M-token context window, and improves...

z-ai/glm-5.3:batch 1.04858M context $0.7/M input $2.2/M output

Grok 4.6 is SpaceXAI's smartest model with frontier performance on coding, knowledge work, and STEM.

x-ai/grok-4.6 500K context $2/M input $6/M output

Qwen3.8 27B is an open-weight dense vision-language model from Qwen. It is suited for coding, professional workflows, research, multimodal interaction, and long-running agent tasks, with flexible thinking that can be...

qwen/qwen3.8-27b 262.144K context $0.214/M input $2.55/M output

GPT-5.6 Sol Pro is the same underlying model as [GPT-5.6 Sol](https://openrouter.ai/openai/gpt-5.6-sol), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. Learn more in OpenAI's docs: https://developers.openai.com/api/docs/guides/reasoning#reasoning-mode

openai/gpt-5.6-sol-pro:batch 1.05M context $1/M input $5/M output

Qwen3.8 Max 0902 is an updated snapshot of Qwen3.8 Max from Alibaba's Qwen team. It is a 2.4-trillion-parameter mixture-of-experts model that accepts text, image, and video input and returns text,...

qwen/qwen3.8-max-0902 1M context $2/M input $6/M output

Claude Opus 4.5 is Anthropic’s frontier reasoning model optimized for complex software engineering, agentic workflows, and long-horizon computer use. It offers strong multimodal capabilities, competitive performance across real-world coding and...

anthropic/claude-opus-4.5:batch 200K context $2.5/M input $12.5/M output

GPT-5.2 Pro is OpenAI’s most advanced model, offering major improvements in agentic coding and long context performance over GPT-5 Pro. It is optimized for complex tasks that require step-by-step reasoning,...

openai/gpt-5.2-pro:batch 400K context $10.5/M input $84/M output

GLM-5.3-Flash is a native multimodal model from Z.ai. It is suited for efficient coding and long-horizon agent tasks. Its hybrid sparse and linear attention architecture maintains accurate long-context behavior while...

z-ai/glm-5.3-flash 1.04858M context $0.1/M input $0.333/M output

No provider description is available for this model yet.

azure/gpt-chat-latest 272K context $5/M input $30/M output

No provider description is available for this model yet.

azure/chat-latest 272K context $5/M input $30/M output

No provider description is available for this model yet.

azure/us/gpt-chat-latest 272K context $5.5/M input $33/M output

DeepSeek V4 Flash 0731 is a sparse mixture-of-experts model from DeepSeek, with 13B active parameters out of 284B total. This re-post-trained revision is suited for coding, reasoning, and agent workflows....

deepseek/deepseek-v4-flash-0731:batch 1.04858M context $0.11/M input $0.33/M output

Mercury 2 is an extremely fast reasoning LLM, and the first reasoning diffusion LLM (dLLM). Instead of generating tokens sequentially, Mercury 2 produces and refines multiple tokens in parallel, achieving...

inception/mercury-2 128K context $0.25/M input $0.75/M output

Opus 4.6 is Anthropic’s strongest model for coding and long-running professional tasks. It is built for agents that operate across entire workflows rather than single prompts, making it especially effective...

anthropic/claude-opus-4.6:batch 1M context $2.5/M input $12.5/M output

Uncensored and creative writing model based on Mistral Small 3.2 24B with good recall, prompt adherence, and intelligence.

thedrummer/cydonia-24b-v4.1 131.072K context $0.3/M input $0.5/M output