1,436 models

Sonnet 4.6 is Anthropic's most capable Sonnet-class model yet, with frontier performance across coding, agents, and professional work. It excels at iterative development, complex codebase navigation, end-to-end project management with...

anthropic/claude-sonnet-4.6 1M context $3/M input $15/M output

No provider description is available for this model yet.

azure/gpt-5.4-nano 272K context $0.2/M input $1.25/M output

No provider description is available for this model yet.

azure/gpt-5.4-nano-2026-03-17 272K context $0.2/M input $1.25/M output

Qwen3.8 Max 0902 is an updated snapshot of Qwen3.8 Max from Alibaba's Qwen team. It is a 2.4-trillion-parameter mixture-of-experts model that accepts text, image, and video input and returns text,...

qwen/qwen3.8-max-0902 1M context $2/M input $6/M output

No provider description is available for this model yet.

azure/o1 200K context $15/M input $60/M output

No provider description is available for this model yet.

azure/o1-2024-12-17 200K context $15/M input $60/M output

GPT-6 Astra Pro is the same underlying model as [GPT-6 Astra](https://openrouter.ai/openai/gpt-6-astra), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. Learn more in OpenAI's docs: https://developers.openai.com/api/docs/guides/reasoning#reasoning-mode

openai/gpt-6-astra-pro 1.05M context $10/M input $50/M output

Qwen3.8 27B is an open-weight dense vision-language model from Qwen. It is suited for coding, professional workflows, research, multimodal interaction, and long-running agent tasks, with flexible thinking that can be...

qwen/qwen3.8-27b 262.144K context $0.214/M input $2.55/M output

No provider description is available for this model yet.

azure/o3 200K context $2/M input $8/M output

Opus 4.6 is Anthropic’s strongest model for coding and long-running professional tasks. It is built for agents that operate across entire workflows rather than single prompts, making it especially effective...

anthropic/claude-opus-4.6:batch 1M context $2.5/M input $12.5/M output

No provider description is available for this model yet.

azure/o3-2025-04-16 200K context $2/M input $8/M output

No provider description is available for this model yet.

azure/o4-mini 200K context $1.1/M input $4.4/M output

No provider description is available for this model yet.

azure/o4-mini-2025-04-16 200K context $1.1/M input $4.4/M output

No provider description is available for this model yet.

azure/us/gpt-4.1-2025-04-14 1.04758M context $2.2/M input $8.8/M output

No provider description is available for this model yet.

azure/us/gpt-4.1-mini-2025-04-14 1.04758M context $0.44/M input $1.76/M output

No provider description is available for this model yet.

azure/us/gpt-4.1-nano-2025-04-14 1.04758M context $0.11/M input $0.44/M output

No provider description is available for this model yet.

azure/us/gpt-4o-2024-08-06 128K context $2.75/M input $11/M output

No provider description is available for this model yet.

azure/us/gpt-4o-2024-11-20 128K context $2.75/M input $11/M output

GPT-5.2 is the latest frontier-grade model in the GPT-5 series, offering stronger agentic and long context perfomance compared to GPT-5.1. It uses adaptive reasoning to allocate computation dynamically, responding quickly...

openai/gpt-5.2:batch 400K context $0.875/M input $7/M output

No provider description is available for this model yet.

azure/us/gpt-4o-mini-2024-07-18 128K context $0.165/M input $0.66/M output

No provider description is available for this model yet.

azure/us/gpt-5-2025-08-07 272K context $1.375/M input $11/M output

No provider description is available for this model yet.

azure/us/gpt-5-mini-2025-08-07 272K context $0.275/M input $2.2/M output

No provider description is available for this model yet.

azure/us/gpt-5-nano-2025-08-07 272K context $0.055/M input $0.44/M output

No provider description is available for this model yet.

azure/us/gpt-5.1 272K context $1.38/M input $11/M output

No provider description is available for this model yet.

azure/us/gpt-5.1-chat 128K context $1.38/M input $11/M output

No provider description is available for this model yet.

azure/us/o1-2024-12-17 200K context $16.5/M input $66/M output

Inkling Small is an open-weight multimodal mixture-of-experts model from Thinking Machines Lab, with 12B active parameters out of 276B total. It is positioned as the smaller, more efficient member of...

thinkingmachines/inkling-small:free 1.04858M context Free input Free output

Inkling Small is an open-weight multimodal mixture-of-experts model from Thinking Machines Lab, with 12B active parameters out of 276B total. It is positioned as the smaller, more efficient member of...

thinkingmachines/inkling-small:batch 524.288K context $0.5/M input $1.2/M output

No provider description is available for this model yet.

azure/us/o3-2025-04-16 200K context $2.2/M input $8.8/M output

No provider description is available for this model yet.

azure/us/o4-mini-2025-04-16 200K context $1.21/M input $4.84/M output

Mistral Large 3 2512 is Mistral’s most capable model to date, featuring a sparse mixture-of-experts architecture with 41B active parameters (675B total), and released under the Apache 2.0 license.

mistralai/mistral-large-2512:batch 262.144K context $0.25/M input $0.75/M output

Claude Opus 4.5 is Anthropic’s frontier reasoning model optimized for complex software engineering, agentic workflows, and long-horizon computer use. It offers strong multimodal capabilities, competitive performance across real-world coding and...

anthropic/claude-opus-4.5:batch 200K context $2.5/M input $12.5/M output

No provider description is available for this model yet.

azure_ai/phi-3.5-vision-instruct 128K context $0.13/M input $0.52/M output

No provider description is available for this model yet.

azure_ai/phi-4-multimodal-instruct 131.072K context $0.08/M input $0.32/M output

Mistral Medium 3 is a high-performance enterprise-grade language model designed to deliver frontier-level capabilities at significantly reduced operational cost. It balances state-of-the-art reasoning and multimodal performance with 8× lower cost...

mistralai/mistral-medium-3 131.072K context $0.4/M input $2/M output

Ling 3.0 Flash VL builds on Ling 3.0 Flash (124B total / 5.5B active MoE from InclusionAI), further strengthening its language capabilities while adding native visual perception and advanced visual...

inclusionai/ling-3.0-flash-vl:free 262.144K context Free input Free output

Claude Opus 4.1 is an updated version of Anthropic’s flagship model, offering improved performance in coding, reasoning, and agentic tasks. It achieves 74.5% on SWE-bench Verified and shows notable gains...

anthropic/claude-opus-4.1 200K context $15/M input $75/M output

No provider description is available for this model yet.

azure_ai/kimi-k2.5 262.144K context $0.6/M input $3/M output

No provider description is available for this model yet.

azure_ai/kimi-k2.6 262.144K context $0.95/M input $4/M output

No provider description is available for this model yet.

azure_ai/mistral-large-3 256K context $0.5/M input $1.5/M output

No provider description is available for this model yet.

azure_ai/mistral-small-2503 128K context $0.1/M input $0.3/M output

Gemini 3.6 Flash is a high-efficiency model from Google for coding, agentic workflows, and web and app development. It is designed to produce polished outputs with fewer unnecessary edits and...

google/gemini-3.6-flash:batch 1.04858M context $0.375/M input $1.875/M output

Inkling is an open-weight multimodal mixture-of-experts model from Thinking Machines Lab, with 41B active parameters out of 975B total. It is designed for general-purpose reasoning, coding, agentic and tool-use systems,...

thinkingmachines/inkling:batch 524.288K context $1/M input $4.05/M output

No provider description is available for this model yet.

bedrock/moonshotai.kimi-k2.5 262.144K context $0.6/M input $3.03/M output

No provider description is available for this model yet.

anthropic/claude-4-opus-20250514 200K context $15/M input $75/M output