3,223 models

Nex-N2.5 is an agentic model built to turn goals into working, verified outcomes. Its core strength is agentic coding within a visual feedback loop: it can explore codebases, implement multi-file...

nex-agi/nex-n2.5-mini:free 262.144K context Free input Free output

GPT-6 Astra Pro is the same underlying model as [GPT-6 Astra](https://openrouter.ai/openai/gpt-6-astra), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. Learn more in OpenAI's docs: https://developers.openai.com/api/docs/guides/reasoning#reasoning-mode

openai/gpt-6-astra-pro:batch 1.05M context $5/M input $25/M output

No provider description is available for this model yet.

azure/us/gpt-5.5-2026-04-23 1.05M context $5.5/M input $33/M output

No provider description is available for this model yet.

azure/eu/gpt-5.5-2026-04-23 1.05M context $5.5/M input $33/M output

Mercury 2.5 is the fastest reasoning LLM, and the latest diffusion LLM (dLLM) from Inception. Instead of generating tokens sequentially, Mercury 2.5 produces and refines multiple tokens in parallel, achieving...

inception/mercury-2.5-preview 260K context $0.2/M input $0.75/M output

No provider description is available for this model yet.

azure/gpt-5.4-mini 272K context $0.75/M input $4.5/M output

No provider description is available for this model yet.

cerebras/qwen-3.8-27b 65.536K context $0.99/M input $1.49/M output

Devstral 2 is a state-of-the-art open-source model by Mistral AI specializing in agentic coding. It is a 123B-parameter dense transformer model supporting a 256K context window. Devstral 2 supports exploring...

mistralai/devstral-2512 262.144K context $0.4/M input $2/M output

Qwen3, the latest generation in the Qwen large language model series, features both dense and mixture-of-experts (MoE) architectures to excel in reasoning, multilingual support, and advanced agent tasks. Its unique...

qwen/qwen3-30b-a3b 40.96K context $0.12/M input $0.5/M output

DeepSeek V4 Pro 0813 is a large-scale mixture-of-experts model from DeepSeek. This is the GA release of DeepSeek V4 Pro.

deepseek/deepseek-v4-pro-0813:batch 1.04858M context $0.66/M input $1.98/M output

Hunyuan-A13B is a 13B active parameter Mixture-of-Experts (MoE) language model developed by Tencent, with a total parameter count of 80B and support for reasoning via Chain-of-Thought. It offers competitive benchmark...

tencent/hunyuan-a13b-instruct 131.072K context $0.14/M input $0.57/M output

Qwen3.6-35B-A3B is an open-weight multimodal model from Alibaba Cloud with 35 billion total parameters and 3 billion active parameters per token. It uses a hybrid sparse mixture-of-experts architecture combining Gated...

qwen/qwen3.6-35b-a3b 262.144K context $0.1/M input $0.9/M output

GPT-5.6 Sol Pro is the same underlying model as [GPT-5.6 Sol](https://openrouter.ai/openai/gpt-5.6-sol), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. Learn more in OpenAI's docs: https://developers.openai.com/api/docs/guides/reasoning#reasoning-mode

openai/gpt-5.6-sol-pro 1.05M context $2/M input $10/M output

GPT-5.6 Luna Pro is the same underlying model as [GPT-5.6 Luna](https://openrouter.ai/openai/gpt-5.6-luna), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. Learn more in OpenAI's docs: https://developers.openai.com/api/docs/guides/reasoning#reasoning-mode

openai/gpt-5.6-luna-pro 1.05M context $0.2/M input $1.2/M output

NVIDIA-Nemotron-Nano-9B-v2 is a large language model (LLM) trained from scratch by NVIDIA, and designed as a unified model for both reasoning and non-reasoning tasks. It responds to user queries and...

nvidia/nemotron-nano-9b-v2:free 128K context Free input Free output

No provider description is available for this model yet.

azure/mistral-large-2402 32K context $8/M input $24/M output

No provider description is available for this model yet.

azure/mistral-large-latest 32K context $8/M input $24/M output

No provider description is available for this model yet.

azure/o1 200K context $15/M input $60/M output

No provider description is available for this model yet.

azure/o1-2024-12-17 200K context $15/M input $60/M output

No provider description is available for this model yet.

azure/o1-mini-2024-09-12 128K context $1.1/M input $4.4/M output

Qwen3.8 27B is an open-weight dense vision-language model from Qwen. It is suited for coding, professional workflows, research, multimodal interaction, and long-running agent tasks, with flexible thinking that can be...

qwen/qwen3.8-27b 262.144K context $0.214/M input $2.55/M output

No provider description is available for this model yet.

azure/o1-preview 128K context $15/M input $60/M output

No provider description is available for this model yet.

azure/o1-preview-2024-09-12 128K context $15/M input $60/M output

No provider description is available for this model yet.

cerebras/llama3.1-8b 128K context $0.1/M input $0.1/M output

No provider description is available for this model yet.

cloudflare/@cf/openai/gpt-oss-120b 128K context $0.35/M input $0.75/M output

GPT-4o Search Previewis a specialized model for web search in Chat Completions. It is trained to understand and execute web search queries.

openai/gpt-4o-search-preview 128K context $2.5/M input $10/M output

No provider description is available for this model yet.

azure/o3-2025-04-16 200K context $2/M input $8/M output

No provider description is available for this model yet.

azure/o3-mini 200K context $1.1/M input $4.4/M output

No provider description is available for this model yet.

azure/o3-mini-2025-01-31 200K context $1.1/M input $4.4/M output

No provider description is available for this model yet.

cerebras/gpt-oss-120b 131.072K context $0.35/M input $0.75/M output

May 28th update to the [original DeepSeek R1](/deepseek/deepseek-r1) Performance on par with [OpenAI o1](/openai/o1), but open-sourced and with fully open reasoning tokens. It's 671B parameters in size, with 37B active...

deepseek/deepseek-r1-0528 163.84K context $0.5/M input $2.15/M output

No provider description is available for this model yet.

bedrock_mantle/xai.grok-4.6 500K context $2.2/M input $6.6/M output

No provider description is available for this model yet.

azure/us/gpt-4.1-2025-04-14 1.04758M context $2.2/M input $8.8/M output

No provider description is available for this model yet.

azure/us/gpt-4.1-mini-2025-04-14 1.04758M context $0.44/M input $1.76/M output

No provider description is available for this model yet.

azure/us/gpt-4.1-nano-2025-04-14 1.04758M context $0.11/M input $0.44/M output

No provider description is available for this model yet.

azure/us/gpt-4o-2024-08-06 128K context $2.75/M input $11/M output

No provider description is available for this model yet.

azure/us/gpt-4o-2024-11-20 128K context $2.75/M input $11/M output

Qwen Plus 0728, based on the Qwen3 foundation model, is a 1 million context hybrid reasoning model with a balanced performance, speed, and cost combination.

qwen/qwen-plus-2025-07-28 1M context $0.26/M input $0.78/M output

No provider description is available for this model yet.

azure/us/gpt-4o-mini-2024-07-18 128K context $0.165/M input $0.66/M output

No provider description is available for this model yet.

azure/us/gpt-5-2025-08-07 272K context $1.375/M input $11/M output

No provider description is available for this model yet.

azure/us/gpt-5-mini-2025-08-07 272K context $0.275/M input $2.2/M output

GPT-4o ("o" for "omni") is OpenAI's latest AI model, supporting both text and image inputs with text outputs. It maintains the intelligence level of [GPT-4 Turbo](/models/openai/gpt-4-turbo) while being twice as...

openai/gpt-4o:batch 128K context $1.25/M input $5/M output