2,852 models

No provider description is available for this model yet.

gmi/openai/gpt-4o-mini 131.072K context $0.15/M input $0.6/M output

No provider description is available for this model yet.

azure/gpt-4.1-mini 1.04758M context $0.4/M input $1.6/M output

Schematron V2 Turbo is a 3B-parameter HTML-to-JSON extraction model from Inference.net. It prioritizes throughput for high-volume extraction workloads. Extraction instructions must be supplied through a JSON schema in response_format rather...

inference-net/schematron-v2-turbo 128K context $0.03/M input $0.15/M output

No provider description is available for this model yet.

azure/gpt-4.1-2025-04-14 1.04758M context $2/M input $8/M output

No provider description is available for this model yet.

azure/gpt-4.1 1.04758M context $2/M input $8/M output

Ling 3.0 Tiny is a mixture-of-experts model from InclusionAI, with 1.3B active parameters out of 7.9B total. It is designed for responsive agents, instruction following, and multi-turn conversations, with switchable...

inclusionai/ling-3.0-tiny:free 262.144K context Free input Free output

Laguna M.1 is the flagship coding agent model from [Poolside](https://poolside.ai/), optimized for complex software engineering tasks. Designed for agentic coding workflows, it supports tool calling and reasoning, with a 256K...

poolside/laguna-m.1:free 262.144K context Free input Free output

Qwen3-Next-80B-A3B-Instruct is an instruction-tuned chat model in the Qwen3-Next series optimized for fast, stable responses without “thinking” traces. It targets complex tasks across reasoning, code generation, knowledge QA, and multilingual...

qwen/qwen3-next-80b-a3b-instruct 262.144K context $0.09/M input $1.1/M output

No provider description is available for this model yet.

azure/gpt-4-turbo-vision-preview 128K context $10/M input $30/M output

No provider description is available for this model yet.

azure/gpt-4-turbo-2024-04-09 128K context $10/M input $30/M output

GPT-5.6 Sol Pro is the same underlying model as [GPT-5.6 Sol](https://openrouter.ai/openai/gpt-5.6-sol), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. Learn more in OpenAI's docs: https://developers.openai.com/api/docs/guides/reasoning#reasoning-mode

openai/gpt-5.6-sol-pro:batch 1.05M context $1/M input $5/M output

This model always redirects to the latest model in the GPT Sol family.

~openai/gpt-sol-latest 1.05M context $2/M input $10/M output

Qwen Plus 0728, based on the Qwen3 foundation model, is a 1 million context hybrid reasoning model with a balanced performance, speed, and cost combination.

qwen/qwen-plus-2025-07-28:thinking 1M context $0.26/M input $0.78/M output

No provider description is available for this model yet.

azure/gpt-4-turbo 128K context $10/M input $30/M output

No provider description is available for this model yet.

azure/gpt-4-1106-preview 128K context $10/M input $30/M output

No provider description is available for this model yet.

azure/gpt-4-0125-preview 128K context $10/M input $30/M output

Qwen3-235B-A22B is a 235B parameter mixture-of-experts (MoE) model developed by Qwen, activating 22B parameters per forward pass. It supports seamless switching between a "thinking" mode for complex reasoning, math, and...

qwen/qwen3-235b-a22b 131.072K context $0.455/M input $1.82/M output

No provider description is available for this model yet.

azure/global/gpt-5.1-chat 128K context $1.25/M input $10/M output

GLM-5V-Turbo is Z.ai’s first native multimodal agent foundation model, built for vision-based coding and agent-driven tasks. It natively handles image, video, and text inputs, excels at long-horizon planning, complex coding,...

z-ai/glm-5v-turbo 202.752K context $1.2/M input $4/M output

Mistral Small 3.1 24B Instruct is an upgraded variant of Mistral Small 3 (2501), featuring 24 billion parameters with advanced multimodal capabilities. It provides state-of-the-art performance in text-based reasoning and...

mistralai/mistral-small-3.1-24b-instruct 128K context $0.351/M input $0.555/M output

No provider description is available for this model yet.

azure/global/gpt-5.1 272K context $1.25/M input $10/M output

No provider description is available for this model yet.

openai/gpt-5-2025-08-07 272K context $1.25/M input $10/M output

No provider description is available for this model yet.

azure/global/gpt-4o-2024-11-20 128K context $2.5/M input $10/M output

No provider description is available for this model yet.

openai/gpt-5.4-nano-2026-03-17 272K context $0.2/M input $1.25/M output

No provider description is available for this model yet.

openai/gpt-5.4-mini-2026-03-17 272K context $0.75/M input $4.5/M output

No provider description is available for this model yet.

openai/gpt-5.4-2026-03-05 1.05M context $2.5/M input $15/M output

No provider description is available for this model yet.

openai/gpt-5.5-2026-04-23 1.05M context $5/M input $30/M output

No provider description is available for this model yet.

azure/global/gpt-4o-2024-08-06 128K context $2.5/M input $10/M output

Ox Alpha is a reasoning model designed for coding, sustained agentic work, and production workloads. It is suited for long-horizon software engineering, complex reasoning, and workflows that combine text with...

stealth/ox-alpha 1.04858M context Input not listed Output not listed

No provider description is available for this model yet.

openai/gpt-5-mini-2025-08-07 272K context $0.25/M input $2/M output

No provider description is available for this model yet.

openai/gpt-5-nano-2025-08-07 272K context $0.05/M input $0.4/M output

No provider description is available for this model yet.

openai/gpt-5.6 1.05M context $5/M input $30/M output