Sonnet 4.6 is Anthropic's most capable Sonnet-class model yet, with frontier performance across coding, agents, and professional work. It excels at iterative development, complex codebase navigation, end-to-end project management with...
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
No provider description is available for this model yet.
No provider description is available for this model yet.
Qwen3.8 Max 0902 is an updated snapshot of Qwen3.8 Max from Alibaba's Qwen team. It is a 2.4-trillion-parameter mixture-of-experts model that accepts text, image, and video input and returns text,...
No provider description is available for this model yet.
No provider description is available for this model yet.
GPT-6 Astra Pro is the same underlying model as [GPT-6 Astra](https://openrouter.ai/openai/gpt-6-astra), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. Learn more in OpenAI's docs: https://developers.openai.com/api/docs/guides/reasoning#reasoning-mode
Qwen3.8 27B is an open-weight dense vision-language model from Qwen. It is suited for coding, professional workflows, research, multimodal interaction, and long-running agent tasks, with flexible thinking that can be...
No provider description is available for this model yet.
Opus 4.6 is Anthropic’s strongest model for coding and long-running professional tasks. It is built for agents that operate across entire workflows rather than single prompts, making it especially effective...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
GPT-5.2 is the latest frontier-grade model in the GPT-5 series, offering stronger agentic and long context perfomance compared to GPT-5.1. It uses adaptive reasoning to allocate computation dynamically, responding quickly...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Inkling Small is an open-weight multimodal mixture-of-experts model from Thinking Machines Lab, with 12B active parameters out of 276B total. It is positioned as the smaller, more efficient member of...
Inkling Small is an open-weight multimodal mixture-of-experts model from Thinking Machines Lab, with 12B active parameters out of 276B total. It is positioned as the smaller, more efficient member of...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Mistral Large 3 2512 is Mistral’s most capable model to date, featuring a sparse mixture-of-experts architecture with 41B active parameters (675B total), and released under the Apache 2.0 license.
Claude Opus 4.5 is Anthropic’s frontier reasoning model optimized for complex software engineering, agentic workflows, and long-horizon computer use. It offers strong multimodal capabilities, competitive performance across real-world coding and...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Mistral Medium 3 is a high-performance enterprise-grade language model designed to deliver frontier-level capabilities at significantly reduced operational cost. It balances state-of-the-art reasoning and multimodal performance with 8× lower cost...
Ling 3.0 Flash VL builds on Ling 3.0 Flash (124B total / 5.5B active MoE from InclusionAI), further strengthening its language capabilities while adding native visual perception and advanced visual...
Claude Opus 4.1 is an updated version of Anthropic’s flagship model, offering improved performance in coding, reasoning, and agentic tasks. It achieves 74.5% on SWE-bench Verified and shows notable gains...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Gemini 3.6 Flash is a high-efficiency model from Google for coding, agentic workflows, and web and app development. It is designed to produce polished outputs with fewer unnecessary edits and...
Inkling is an open-weight multimodal mixture-of-experts model from Thinking Machines Lab, with 41B active parameters out of 975B total. It is designed for general-purpose reasoning, coding, agentic and tool-use systems,...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| Anthropic: Claude Sonnet 4.6anthropic/claude-sonnet-4.6 | 1M | $3 | $15 | — | |||
| gpt-5.4-nanoazure/gpt-5.4-nano | 272K | $0.2 | $1.25 | — | |||
| gpt-5.4-nano-2026-03-17azure/gpt-5.4-nano-2026-03-17 | 272K | $0.2 | $1.25 | — | |||
| Qwen: Qwen3.8 Max (0902)qwen/qwen3.8-max-0902 | 1M | $2 | $6 | — | |||
| o1azure/o1 | 200K | $15 | $60 | — | |||
| o1-2024-12-17azure/o1-2024-12-17 | 200K | $15 | $60 | — | |||
| OpenAI: GPT-6 Astra Proopenai/gpt-6-astra-pro | 1.05M | $10 | $50 | — | |||
| Qwen: Qwen3.8 27Bqwen/qwen3.8-27b | 262.144K | $0.214 | $2.55 | — | |||
| o3azure/o3 | 200K | $2 | $8 | — | |||
| Anthropic: Claude Opus 4.6 (batch)anthropic/claude-opus-4.6:batch | 1M | $2.5 | $12.5 | — | |||
| o3-2025-04-16azure/o3-2025-04-16 | 200K | $2 | $8 | — | |||
| o4-miniazure/o4-mini | 200K | $1.1 | $4.4 | — | |||
| o4-mini-2025-04-16azure/o4-mini-2025-04-16 | 200K | $1.1 | $4.4 | — | |||
| us/gpt-4.1-2025-04-14azure/us/gpt-4.1-2025-04-14 | 1.04758M | $2.2 | $8.8 | — | |||
| us/gpt-4.1-mini-2025-04-14azure/us/gpt-4.1-mini-2025-04-14 | 1.04758M | $0.44 | $1.76 | — | |||
| us/gpt-4.1-nano-2025-04-14azure/us/gpt-4.1-nano-2025-04-14 | 1.04758M | $0.11 | $0.44 | — | |||
| us/gpt-4o-2024-08-06azure/us/gpt-4o-2024-08-06 | 128K | $2.75 | $11 | — | |||
| us/gpt-4o-2024-11-20azure/us/gpt-4o-2024-11-20 | 128K | $2.75 | $11 | — | |||
| OpenAI: GPT-5.2 (batch)openai/gpt-5.2:batch | 400K | $0.875 | $7 | — | |||
| us/gpt-4o-mini-2024-07-18azure/us/gpt-4o-mini-2024-07-18 | 128K | $0.165 | $0.66 | — | |||
| us/gpt-5-2025-08-07azure/us/gpt-5-2025-08-07 | 272K | $1.375 | $11 | — | |||
| us/gpt-5-mini-2025-08-07azure/us/gpt-5-mini-2025-08-07 | 272K | $0.275 | $2.2 | — | |||
| us/gpt-5-nano-2025-08-07azure/us/gpt-5-nano-2025-08-07 | 272K | $0.055 | $0.44 | — | |||
| us/gpt-5.1azure/us/gpt-5.1 | 272K | $1.38 | $11 | — | |||
| us/gpt-5.1-chatazure/us/gpt-5.1-chat | 128K | $1.38 | $11 | — | |||
| us/o1-2024-12-17azure/us/o1-2024-12-17 | 200K | $16.5 | $66 | — | |||
| Thinking Machines: Inkling Small (free)thinkingmachines/inkling-small:free | 1.04858M | Free | Free | — | |||
| Thinking Machines: Inkling Small (batch)thinkingmachines/inkling-small:batch | 524.288K | $0.5 | $1.2 | — | |||
| us/o3-2025-04-16azure/us/o3-2025-04-16 | 200K | $2.2 | $8.8 | — | |||
| us/o4-mini-2025-04-16azure/us/o4-mini-2025-04-16 | 200K | $1.21 | $4.84 | — | |||
| Llama-3.2-11B-Vision-Instructazure_ai/llama-3.2-11b-vision-instruct | 128K | $0.37 | $0.37 | — | |||
| Mistral: Mistral Large 3 2512 (batch)mistralai/mistral-large-2512:batch | 262.144K | $0.25 | $0.75 | — | |||
| Anthropic: Claude Opus 4.5 (batch)anthropic/claude-opus-4.5:batch | 200K | $2.5 | $12.5 | — | |||
| Llama-3.2-90B-Vision-Instructazure_ai/llama-3.2-90b-vision-instruct | 128K | $2.04 | $2.04 | — | |||
| Llama-4-Maverick-17B-128E-Instruct-FP8azure_ai/llama-4-maverick-17b-128e-instruct-fp8 | 1M | $1.41 | $0.35 | — | |||
| Llama-4-Scout-17B-16E-Instructazure_ai/llama-4-scout-17b-16e-instruct | 10M | $0.2 | $0.78 | — | |||
| Phi-3.5-vision-instructazure_ai/phi-3.5-vision-instruct | 128K | $0.13 | $0.52 | — | |||
| Phi-4-multimodal-instructazure_ai/phi-4-multimodal-instruct | 131.072K | $0.08 | $0.32 | — | |||
| Mistral: Mistral Medium 3mistralai/mistral-medium-3 | 131.072K | $0.4 | $2 | — | |||
| inclusionAI: Ling 3.0 Flash VL (free)inclusionai/ling-3.0-flash-vl:free | 262.144K | Free | Free | — | |||
| Anthropic: Claude Opus 4.1anthropic/claude-opus-4.1 | 200K | $15 | $75 | — | |||
| kimi-k2.5azure_ai/kimi-k2.5 | 262.144K | $0.6 | $3 | — | |||
| kimi-k2.6azure_ai/kimi-k2.6 | 262.144K | $0.95 | $4 | — | |||
| mistral-large-3azure_ai/mistral-large-3 | 256K | $0.5 | $1.5 | — | |||
| mistral-small-2503azure_ai/mistral-small-2503 | 128K | $0.1 | $0.3 | — | |||
| Google: Gemini 3.6 Flash (batch)google/gemini-3.6-flash:batch | 1.04858M | $0.375 | $1.875 | — | |||
| Thinking Machines: Inkling (batch)thinkingmachines/inkling:batch | 524.288K | $1 | $4.05 | — | |||
| ap-northeast-1/moonshotai.kimi-k2.5bedrock/ap-northeast-1/moonshotai.kimi-k2.5 | 262.144K | $0.72 | $3.6 | — | |||
| moonshotai.kimi-k2.5bedrock/moonshotai.kimi-k2.5 | 262.144K | $0.6 | $3.03 | — | |||
| claude-4-opus-20250514anthropic/claude-4-opus-20250514 | 200K | $15 | $75 | — |