No provider description is available for this model yet.
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
No provider description is available for this model yet.
No provider description is available for this model yet.
Qwen3.8 Max 0902 is an updated snapshot of Qwen3.8 Max from Alibaba's Qwen team. It is a 2.4-trillion-parameter mixture-of-experts model that accepts text, image, and video input and returns text,...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
GPT-6 Astra Pro is the same underlying model as [GPT-6 Astra](https://openrouter.ai/openai/gpt-6-astra), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. Learn more in OpenAI's docs: https://developers.openai.com/api/docs/guides/reasoning#reasoning-mode
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Qwen3.8 27B is an open-weight dense vision-language model from Qwen. It is suited for coding, professional workflows, research, multimodal interaction, and long-running agent tasks, with flexible thinking that can be...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
GPT-4.1 Mini is a mid-sized model delivering performance competitive with GPT-4o at substantially lower latency and cost. It retains a 1 million token context window and scores 45.1% on hard...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Kimi K3 is a 2.8T parameter open-weight multimodal reasoning model from Moonshot AI. It is suited for complex coding, knowledge work, and long-horizon agentic workflows, and is particularly strong at...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Qwen Plus 0728, based on the Qwen3 foundation model, is a 1 million context hybrid reasoning model with a balanced performance, speed, and cost combination.
No provider description is available for this model yet.
No provider description is available for this model yet.
DeepSeek V4 Flash Vision Exp is an experimental vision-enabled version of [DeepSeek V4 Flash 0731](https://openrouter.ai/deepseek/deepseek-v4-flash-0731) from DeepSeek, adding image understanding while matching the base model on text capabilities including agents,...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Inkling Small is an open-weight multimodal mixture-of-experts model from Thinking Machines Lab, with 12B active parameters out of 276B total. It is positioned as the smaller, more efficient member of...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| gpt-5.4-nanoazure/gpt-5.4-nano | 272K | $0.2 | $1.25 | — | |||
| gpt-5.4-nano-2026-03-17azure/gpt-5.4-nano-2026-03-17 | 272K | $0.2 | $1.25 | — | |||
| mistral-large-2402azure/mistral-large-2402 | 32K | $8 | $24 | — | |||
| Qwen: Qwen3.8 Max (0902)qwen/qwen3.8-max-0902 | 1M | $2 | $6 | — | |||
| mistral-large-latestazure/mistral-large-latest | 32K | $8 | $24 | — | |||
| o1azure/o1 | 200K | $15 | $60 | — | |||
| o1-2024-12-17azure/o1-2024-12-17 | 200K | $15 | $60 | — | |||
| OpenAI: GPT-6 Astra Proopenai/gpt-6-astra-pro | 1.05M | $10 | $50 | — | |||
| o1-miniazure/o1-mini | 128K | $1.21 | $4.84 | — | |||
| o1-mini-2024-09-12azure/o1-mini-2024-09-12 | 128K | $1.1 | $4.4 | — | |||
| o1-previewazure/o1-preview | 128K | $15 | $60 | — | |||
| Qwen: Qwen3.8 27Bqwen/qwen3.8-27b | 262.144K | $0.214 | $2.55 | — | |||
| o1-preview-2024-09-12azure/o1-preview-2024-09-12 | 128K | $15 | $60 | — | |||
| llama3.1-8bcerebras/llama3.1-8b | 128K | $0.1 | $0.1 | — | |||
| o3azure/o3 | 200K | $2 | $8 | — | |||
| o3-2025-04-16azure/o3-2025-04-16 | 200K | $2 | $8 | — | |||
| o3-miniazure/o3-mini | 200K | $1.1 | $4.4 | — | |||
| OpenAI: GPT-4.1 Mini (batch)openai/gpt-4.1-mini:batch | 1.04758M | $0.2 | $0.8 | — | |||
| o3-mini-2025-01-31azure/o3-mini-2025-01-31 | 200K | $1.1 | $4.4 | — | |||
| gpt-oss-120bcerebras/gpt-oss-120b | 131.072K | $0.35 | $0.75 | — | |||
| o4-miniazure/o4-mini | 200K | $1.1 | $4.4 | — | |||
| o4-mini-2025-04-16azure/o4-mini-2025-04-16 | 200K | $1.1 | $4.4 | — | |||
| us/gpt-4.1-2025-04-14azure/us/gpt-4.1-2025-04-14 | 1.04758M | $2.2 | $8.8 | — | |||
| us/gpt-4.1-mini-2025-04-14azure/us/gpt-4.1-mini-2025-04-14 | 1.04758M | $0.44 | $1.76 | — | |||
| us/gpt-4.1-nano-2025-04-14azure/us/gpt-4.1-nano-2025-04-14 | 1.04758M | $0.11 | $0.44 | — | |||
| us/gpt-4o-2024-08-06azure/us/gpt-4o-2024-08-06 | 128K | $2.75 | $11 | — | |||
| us/gpt-4o-2024-11-20azure/us/gpt-4o-2024-11-20 | 128K | $2.75 | $11 | — | |||
| MoonshotAI: Kimi K3 (batch)moonshotai/kimi-k3:batch | 1.04858M | $3 | $15 | — | |||
| us/gpt-4o-mini-2024-07-18azure/us/gpt-4o-mini-2024-07-18 | 128K | $0.165 | $0.66 | — | |||
| us/gpt-5-2025-08-07azure/us/gpt-5-2025-08-07 | 272K | $1.375 | $11 | — | |||
| @cf/meta/llama-3.3-70b-instruct-fp8-fastcloudflare/@cf/meta/llama-3.3-70b-instruct-fp8-fast | 24K | $0.293 | $2.253 | — | |||
| us/gpt-5-mini-2025-08-07azure/us/gpt-5-mini-2025-08-07 | 272K | $0.275 | $2.2 | — | |||
| us/gpt-5-nano-2025-08-07azure/us/gpt-5-nano-2025-08-07 | 272K | $0.055 | $0.44 | — | |||
| Qwen: Qwen Plus 0728qwen/qwen-plus-2025-07-28 | 1M | $0.26 | $0.78 | — | |||
| us/gpt-5.1azure/us/gpt-5.1 | 272K | $1.38 | $11 | — | |||
| @cf/ibm-granite/granite-4.0-h-microcloudflare/@cf/ibm-granite/granite-4.0-h-micro | 131K | $0.017 | $0.112 | — | |||
| DeepSeek: DeepSeek V4 Flash Vision Exp (batch)deepseek/deepseek-v4-flash-vision-exp:batch | 1.04858M | $0.11 | $0.33 | — | |||
| us/gpt-5.1-chatazure/us/gpt-5.1-chat | 128K | $1.38 | $11 | — | |||
| us/o1-2024-12-17azure/us/o1-2024-12-17 | 200K | $16.5 | $66 | — | |||
| us/o1-mini-2024-09-12azure/us/o1-mini-2024-09-12 | 128K | $1.21 | $4.84 | — | |||
| @cf/qwen/qwen2.5-coder-32b-instructcloudflare/@cf/qwen/qwen2.5-coder-32b-instruct | 32.768K | $0.66 | $1 | — | |||
| Thinking Machines: Inkling Small (free)thinkingmachines/inkling-small:free | 1.04858M | Free | Free | — | |||
| us/o1-preview-2024-09-12azure/us/o1-preview-2024-09-12 | 128K | $16.5 | $66 | — | |||
| qwen-3-32bcerebras/qwen-3-32b | 128K | $0.4 | $0.8 | — | |||
| us/o3-2025-04-16azure/us/o3-2025-04-16 | 200K | $2.2 | $8.8 | — | |||
| us/o3-mini-2025-01-31azure/us/o3-mini-2025-01-31 | 200K | $1.21 | $4.84 | — | |||
| us/o4-mini-2025-04-16azure/us/o4-mini-2025-04-16 | 200K | $1.21 | $4.84 | — | |||
| Llama-3.2-11B-Vision-Instructazure_ai/llama-3.2-11b-vision-instruct | 128K | $0.37 | $0.37 | — | |||
| stabilityai/codellama13b_instruct_260k_synthesisstabilityai/codellama13b_instruct_260k_synthesis | Not documented | — | — | — | |||
| Llama-3.2-90B-Vision-Instructazure_ai/llama-3.2-90b-vision-instruct | 128K | $2.04 | $2.04 | — |