2,852 models

No provider description is available for this model yet.

scaleway/glm-5.2 256K context $1.8/M input $5.5/M output

No provider description is available for this model yet.

scaleway/deepseek-v4-flash-0731 256K context $0.4/M input $0.8/M output

DeepSeek V4 Flash Vision Exp is an experimental vision-enabled version of [DeepSeek V4 Flash 0731](https://openrouter.ai/deepseek/deepseek-v4-flash-0731) from DeepSeek, adding image understanding while matching the base model on text capabilities including agents,...

deepseek/deepseek-v4-flash-vision-exp:batch 1.04858M context $0.11/M input $0.33/M output

No provider description is available for this model yet.

gemini/gemini-3.8-flash 1.04858M context $0.75/M input $3.75/M output

GPT-5.2 Pro is OpenAI’s most advanced model, offering major improvements in agentic coding and long context performance over GPT-5 Pro. It is optimized for complex tasks that require step-by-step reasoning,...

openai/gpt-5.2-pro:batch 400K context $10.5/M input $84/M output

Gemini 2.5 Flash-Lite is a lightweight reasoning model in the Gemini 2.5 family, optimized for ultra-low latency and cost efficiency. It offers improved throughput, faster token generation, and better performance...

google/gemini-2.5-flash-lite:batch 1.04858M context $0.05/M input $0.2/M output

No provider description is available for this model yet.

azure_ai/gpt-5.5 1.05M context $5/M input $30/M output

Llama 3.2 3B is a 3-billion-parameter multilingual large language model, optimized for advanced natural language processing tasks like dialogue generation, reasoning, and summarization. Designed with the latest transformer architecture, it...

meta-llama/llama-3.2-3b-instruct:free 131.072K context Free input Free output

No provider description is available for this model yet.

azure_ai/gpt-5.4-2026-03-05 1.05M context $2.5/M input $15/M output

Seed 2.1 Turbo is a multimodal model from ByteDance Seed for coding and long-horizon agent workflows. It is suited for end-to-end software delivery, multi-step task execution, and understanding visual and...

bytedance-seed/seed-2-1-turbo 262.144K context $0.5/M input $2.5/M output

Amazon Nova Micro 1.0 is a text-only model that delivers the lowest latency responses in the Amazon Nova family of models at a very low cost. With a context length...

amazon/nova-micro-v1 128K context $0.035/M input $0.14/M output

The Pareto Router maintains a tiered shortlist of strong coding models, ranked by [Artificial Analysis](https://artificialanalysis.ai/) coding percentiles. Set min_coding_score between 0 and 1 on the [pareto-router plugin](https://openrouter.ai/docs/guides/routing/routers/pareto-router#the-min_coding_score-parameter) to control how...

openrouter/pareto-code 2M context Input not listed Output not listed

KAT-Coder-Air V2.5 is a flagship-level Agentic Coding model that can directly hand over an entire issue or an entire business workflow to it, allowing it to autonomously locate and make...

kwaipilot/kat-coder-air-v2.5 256K context $0.15/M input $0.6/M output

No provider description is available for this model yet.

azure_ai/gpt-5.4-nano 272K context $0.2/M input $1.25/M output

No provider description is available for this model yet.

azure_ai/gpt-5.4-nano-2026-03-17 272K context $0.2/M input $1.25/M output

Qwen3.7 Flash is a vision-language reasoning model from Alibaba. It is suited for multimodal agents, visual coding, search, and computer interaction, with strengths in object recognition, spatial understanding, and real-world...

qwen/qwen3.7-flash 1M context $0.03/M input $0.13/M output

No provider description is available for this model yet.

azure/eu/gpt-4o-2024-08-06 128K context $2.75/M input $11/M output

No provider description is available for this model yet.

azure/eu/gpt-4o-2024-11-20 128K context $2.75/M input $11/M output

No provider description is available for this model yet.

azure/eu/gpt-4o-mini-2024-07-18 128K context $0.165/M input $0.66/M output

No provider description is available for this model yet.

azure/eu/gpt-5-2025-08-07 272K context $1.375/M input $11/M output

No provider description is available for this model yet.

azure/eu/gpt-5-mini-2025-08-07 272K context $0.275/M input $2.2/M output

No provider description is available for this model yet.

azure/eu/gpt-5.1 272K context $1.38/M input $11/M output

No provider description is available for this model yet.

azure/eu/gpt-5.1-chat 128K context $1.38/M input $11/M output

No provider description is available for this model yet.

azure/eu/gpt-5-nano-2025-08-07 272K context $0.055/M input $0.44/M output

No provider description is available for this model yet.

azure/eu/o1-2024-12-17 200K context $16.5/M input $66/M output

No provider description is available for this model yet.

azure/eu/o1-mini-2024-09-12 128K context $1.21/M input $4.84/M output

No provider description is available for this model yet.

azure/eu/o1-preview-2024-09-12 128K context $16.5/M input $66/M output

o3 is a well-rounded and powerful model across domains. It sets a new standard for math, science, coding, and visual reasoning tasks. It also excels at technical writing and instruction-following....

openai/o3:batch 200K context $1/M input $4/M output

NVIDIA-Nemotron-Nano-9B-v2 is a large language model (LLM) trained from scratch by NVIDIA, and designed as a unified model for both reasoning and non-reasoning tasks. It responds to user queries and...

nvidia/nemotron-nano-9b-v2:free 128K context Free input Free output

MoonshotAI: Kimi K2.7 Code is a coding-focused model in Moonshot AI's Kimi K2 family, built to complete end-to-end programming tasks reliably over long contexts. It uses a native multimodal mixture-of-experts...

moonshotai/kimi-k2.7-code:batch 262.144K context $0.95/M input $4/M output

No provider description is available for this model yet.

azure/us-gov/o3-mini 200K context $1.513/M input $6.05/M output

Schematron V2 Small is a 3B-parameter HTML-to-JSON extraction model from Inference.net. It prioritizes extraction quality for complex schemas and long pages. Extraction instructions must be supplied through a JSON schema...

inference-net/schematron-v2-small 128K context $0.05/M input $0.23/M output

No provider description is available for this model yet.

azure_ai/gpt-oss-120b 131.072K context $0.15/M input $0.6/M output