1,384 models

No provider description is available for this model yet.

openrouter/openai/o1-pro 200K context $150/M input $600/M output

No provider description is available for this model yet.

openrouter/google/gemma-3-4b-it 131.072K context $0.05/M input $0.1/M output

No provider description is available for this model yet.

openrouter/google/gemma-3-12b-it 131.072K context $0.05/M input $0.15/M output

No provider description is available for this model yet.

openrouter/google/gemma-3-27b-it 131.072K context $0.08/M input $0.45/M output

No provider description is available for this model yet.

novita/google/gemma-4-31b-it 262.144K context $0.14/M input $0.4/M output

No provider description is available for this model yet.

openrouter/minimax/minimax-01 1.00019M context $0.2/M input $1.1/M output

No provider description is available for this model yet.

novita/moonshotai/kimi-k3 1.04858M context $3/M input $15/M output

No provider description is available for this model yet.

openrouter/openai/gpt-4-turbo 128K context $10/M input $30/M output

Claude Opus 4.1 is an updated version of Anthropic’s flagship model, offering improved performance in coding, reasoning, and agentic tasks. It achieves 74.5% on SWE-bench Verified and shows notable gains...

anthropic/claude-opus-4.1 200K context $15/M input $75/M output

OpenAI o4-mini-high is the same model as [o4-mini](/openai/o4-mini) with reasoning_effort set to high. OpenAI o4-mini is a compact reasoning model in the o-series, optimized for fast, cost-efficient performance while retaining...

openai/o4-mini-high 200K context $1.1/M input $4.4/M output

No provider description is available for this model yet.

openrouter/openai/gpt-5.1-codex 400K context $1.25/M input $10/M output

No provider description is available for this model yet.

openrouter/z-ai/glm-4.6v 131.072K context $0.3/M input $0.9/M output

No provider description is available for this model yet.

openrouter/openai/gpt-5.4-pro 1.05M context $30/M input $180/M output

No provider description is available for this model yet.

openrouter/qwen/qwen3.5-9b 262.144K context $0.1/M input $0.15/M output

No provider description is available for this model yet.

qwen_ai_platform/kimi-k2.7-code 229.376K context $0.95/M input $4/M output

No provider description is available for this model yet.

xai/grok-4.20-reasoning 1M context $1.25/M input $2.5/M output

No provider description is available for this model yet.

openrouter/z-ai/glm-5v-turbo 202.752K context $1.2/M input $4/M output

Sonnet 5 is Anthropic's most capable Sonnet-class model, with frontier performance across coding, agents, and professional work. It supports adaptive thinking with selectable reasoning effort levels (low, medium, high, max,...

anthropic/claude-sonnet-5:batch 1M context $1/M input $5/M output

No provider description is available for this model yet.

openrouter/google/gemma-4-31b-it 262.144K context $0.09/M input $0.34/M output

No provider description is available for this model yet.

xai/grok-4.20 1M context $1.25/M input $2.5/M output

No provider description is available for this model yet.

openrouter/moonshotai/kimi-k2.6 262.144K context $0.95/M input $4/M output

No provider description is available for this model yet.

openrouter/openai/gpt-5.5-pro 1.05M context $30/M input $180/M output

No provider description is available for this model yet.

qwencloud/qwen3.8-max 991.808K context $2/M input $6/M output

No provider description is available for this model yet.

gemini/gemini-omni-1.1-flash 131.072K context $1.5/M input $9/M output

Seed 1.6 Flash is an ultra-fast multimodal deep thinking model by ByteDance Seed, supporting both text and visual understanding. It features a 256k context window and can generate outputs of...

bytedance-seed/seed-1.6-flash 262.144K context $0.075/M input $0.3/M output

No provider description is available for this model yet.

novita/google/gemma-4-26b-a4b-it 262.144K context $0.13/M input $0.4/M output

GPT-5.6 Luna Pro is the same underlying model as [GPT-5.6 Luna](https://openrouter.ai/openai/gpt-5.6-luna), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. Learn more in OpenAI's docs: https://developers.openai.com/api/docs/guides/reasoning#reasoning-mode

openai/gpt-5.6-luna-pro 1.05M context $0.2/M input $1.2/M output

No provider description is available for this model yet.

together_ai/zai-org/glm-5.3-flash 1.04858M context $0.15/M input $0.5/M output

Fast-mode variant of [Opus 5](/anthropic/claude-opus-5) - identical capabilities with higher output speed at 2x pricing relative to regular Opus 5. Learn more in Anthropic's docs: https://platform.claude.com/docs/en/build-with-claude/fast-mode Note: As of September...

anthropic/claude-opus-5-fast 1M context $10/M input $50/M output

GPT-6 Astra is OpenAI's flagship model for demanding end-to-end work. It is suited for advanced analysis, software engineering, deep research, scientific work, and document creation, with particular strengths in long-horizon...

openai/gpt-6-astra:batch 1.05M context $5/M input $25/M output

GPT-5.6 Sol Pro is the same underlying model as [GPT-5.6 Sol](https://openrouter.ai/openai/gpt-5.6-sol), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. Learn more in OpenAI's docs: https://developers.openai.com/api/docs/guides/reasoning#reasoning-mode

openai/gpt-5.6-sol-pro 1.05M context $2/M input $10/M output

No provider description is available for this model yet.

openrouter/qwen/qwen3.6-27b 262.144K context $0.3/M input $2/M output

Fast-mode variant of [Opus 4.8](/anthropic/claude-opus-4.8) - identical capabilities with higher output speed at 2x pricing relative to regular Opus 4.8. Learn more in Anthropic's docs: https://platform.claude.com/docs/en/build-with-claude/fast-mode Note: As of September...

anthropic/claude-opus-4.8-fast 1M context $10/M input $50/M output

No provider description is available for this model yet.

openrouter/qwen/qwen3.6-35b-a3b 262.144K context $0.1/M input $0.9/M output

Qwen3.8 Flash is a multimodal reasoning model from Alibaba. It is suited for coding assistance, agentic workflows, visual understanding, document and codebase analysis, desktop interaction, chart analysis, and long-video analysis.

qwen/qwen3.8-flash 1M context $0.15/M input $0.47/M output