3,462 models

DeepSeek V4 Pro 0813 is a large-scale mixture-of-experts model from DeepSeek. This is the GA release of DeepSeek V4 Pro.

deepseek/deepseek-v4-pro-0813:batch 1.04858M context $0.66/M input $1.98/M output

Qwen3.8 Max 0902 is an updated snapshot of Qwen3.8 Max from Alibaba's Qwen team. It is a 2.4-trillion-parameter mixture-of-experts model that accepts text, image, and video input and returns text,...

qwen/qwen3.8-max-0902 1M context $2/M input $6/M output

No provider description is available for this model yet.

azure/gpt-5.4-nano-2026-03-17 272K context $0.2/M input $1.25/M output

No provider description is available for this model yet.

azure/mistral-large-2402 32K context $8/M input $24/M output

No provider description is available for this model yet.

azure/mistral-large-latest 32K context $8/M input $24/M output

No provider description is available for this model yet.

azure/o1 200K context $15/M input $60/M output

No provider description is available for this model yet.

azure/o1-2024-12-17 200K context $15/M input $60/M output

GPT-6 Astra Pro is the same underlying model as [GPT-6 Astra](https://openrouter.ai/openai/gpt-6-astra), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. Learn more in OpenAI's docs: https://developers.openai.com/api/docs/guides/reasoning#reasoning-mode

openai/gpt-6-astra-pro 1.05M context $10/M input $50/M output

No provider description is available for this model yet.

azure/o1-mini 128K context $1.21/M input $4.84/M output

No provider description is available for this model yet.

azure/o1-mini-2024-09-12 128K context $1.1/M input $4.4/M output

Qwen3.8 27B is an open-weight dense vision-language model from Qwen. It is suited for coding, professional workflows, research, multimodal interaction, and long-running agent tasks, with flexible thinking that can be...

qwen/qwen3.8-27b 262.144K context $0.214/M input $2.55/M output

No provider description is available for this model yet.

azure/o1-preview 128K context $15/M input $60/M output

No provider description is available for this model yet.

azure/o1-preview-2024-09-12 128K context $15/M input $60/M output

No provider description is available for this model yet.

cerebras/llama3.1-8b 128K context $0.1/M input $0.1/M output

No provider description is available for this model yet.

azure/o3 200K context $2/M input $8/M output

GPT-4.1 Mini is a mid-sized model delivering performance competitive with GPT-4o at substantially lower latency and cost. It retains a 1 million token context window and scores 45.1% on hard...

openai/gpt-4.1-mini:batch 1.04758M context $0.2/M input $0.8/M output

No provider description is available for this model yet.

azure/o3-2025-04-16 200K context $2/M input $8/M output

No provider description is available for this model yet.

azure/o3-mini 200K context $1.1/M input $4.4/M output

No provider description is available for this model yet.

azure/o3-mini-2025-01-31 200K context $1.1/M input $4.4/M output

No provider description is available for this model yet.

cerebras/gpt-oss-120b 131.072K context $0.35/M input $0.75/M output

No provider description is available for this model yet.

azure/o4-mini 200K context $1.1/M input $4.4/M output

No provider description is available for this model yet.

azure/o4-mini-2025-04-16 200K context $1.1/M input $4.4/M output

No provider description is available for this model yet.

azure/us/gpt-4.1-2025-04-14 1.04758M context $2.2/M input $8.8/M output

No provider description is available for this model yet.

azure/us/gpt-4.1-mini-2025-04-14 1.04758M context $0.44/M input $1.76/M output

No provider description is available for this model yet.

azure/us/gpt-4.1-nano-2025-04-14 1.04758M context $0.11/M input $0.44/M output

No provider description is available for this model yet.

azure/us/gpt-4o-2024-08-06 128K context $2.75/M input $11/M output

No provider description is available for this model yet.

azure/us/gpt-4o-2024-11-20 128K context $2.75/M input $11/M output

Kimi K3 is a 2.8T parameter open-weight multimodal reasoning model from Moonshot AI. It is suited for complex coding, knowledge work, and long-horizon agentic workflows, and is particularly strong at...

moonshotai/kimi-k3:batch 1.04858M context $3/M input $15/M output

No provider description is available for this model yet.

azure/us/gpt-4o-mini-2024-07-18 128K context $0.165/M input $0.66/M output

No provider description is available for this model yet.

azure/us/gpt-5-2025-08-07 272K context $1.375/M input $11/M output

No provider description is available for this model yet.

azure/us/gpt-5-mini-2025-08-07 272K context $0.275/M input $2.2/M output

No provider description is available for this model yet.

azure/us/gpt-5-nano-2025-08-07 272K context $0.055/M input $0.44/M output

No provider description is available for this model yet.

azure/us/gpt-5.1 272K context $1.38/M input $11/M output

No provider description is available for this model yet.

azure/eu/o1-mini-2024-09-12 128K context $1.21/M input $4.84/M output

No provider description is available for this model yet.

azure/eu/o1-preview-2024-09-12 128K context $16.5/M input $66/M output

No provider description is available for this model yet.

azure/us-gov/gpt-5.1 272K context $1.719/M input $13.75/M output

No provider description is available for this model yet.

azure/us-gov/o3-mini 200K context $1.513/M input $6.05/M output

Schematron V2 Small is a 3B-parameter HTML-to-JSON extraction model from Inference.net. It prioritizes extraction quality for complex schemas and long pages. Extraction instructions must be supplied through a JSON schema...

inference-net/schematron-v2-small 128K context $0.05/M input $0.23/M output

No provider description is available for this model yet.

azure/eu/o3-mini-2025-01-31 200K context $1.21/M input $4.84/M output

Ling 3.0 Flash Fin is a finance-focused mixture-of-experts model from InclusionAI, built on Ling 3.0 Flash with 5.1B active parameters out of 124B total. It is designed for real-world investment...

inclusionai/ling-3.0-flash-fin:free 262.144K context Free input Free output

No provider description is available for this model yet.

azure/global-standard/gpt-4o-mini 128K context $0.15/M input $0.6/M output

No provider description is available for this model yet.

azure/global/gpt-4o-2024-08-06 128K context $2.5/M input $10/M output

No provider description is available for this model yet.

azure/global/gpt-4o-2024-11-20 128K context $2.5/M input $10/M output

No provider description is available for this model yet.

azure/global/gpt-5.1 272K context $1.25/M input $10/M output

No provider description is available for this model yet.

azure/global/gpt-5.1-chat 128K context $1.25/M input $10/M output