No provider description is available for this model yet.
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
No provider description is available for this model yet.
No provider description is available for this model yet.
Schematron V2 Turbo is a 3B-parameter HTML-to-JSON extraction model from Inference.net. It prioritizes throughput for high-volume extraction workloads. Extraction instructions must be supplied through a JSON schema in response_format rather...
No provider description is available for this model yet.
No provider description is available for this model yet.
Ling 3.0 Tiny is a mixture-of-experts model from InclusionAI, with 1.3B active parameters out of 7.9B total. It is designed for responsive agents, instruction following, and multi-turn conversations, with switchable...
Laguna M.1 is the flagship coding agent model from [Poolside](https://poolside.ai/), optimized for complex software engineering tasks. Designed for agentic coding workflows, it supports tool calling and reasoning, with a 256K...
Qwen3-Next-80B-A3B-Instruct is an instruction-tuned chat model in the Qwen3-Next series optimized for fast, stable responses without “thinking” traces. It targets complex tasks across reasoning, code generation, knowledge QA, and multilingual...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
GPT-5.6 Sol Pro is the same underlying model as [GPT-5.6 Sol](https://openrouter.ai/openai/gpt-5.6-sol), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. Learn more in OpenAI's docs: https://developers.openai.com/api/docs/guides/reasoning#reasoning-mode
This model always redirects to the latest model in the GPT Sol family.
Qwen Plus 0728, based on the Qwen3 foundation model, is a 1 million context hybrid reasoning model with a balanced performance, speed, and cost combination.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Qwen3-235B-A22B is a 235B parameter mixture-of-experts (MoE) model developed by Qwen, activating 22B parameters per forward pass. It supports seamless switching between a "thinking" mode for complex reasoning, math, and...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
GLM-5V-Turbo is Z.ai’s first native multimodal agent foundation model, built for vision-based coding and agent-driven tasks. It natively handles image, video, and text inputs, excels at long-horizon planning, complex coding,...
Mistral Small 3.1 24B Instruct is an upgraded variant of Mistral Small 3 (2501), featuring 24 billion parameters with advanced multimodal capabilities. It provides state-of-the-art performance in text-based reasoning and...
This model always redirects to the latest model in the Claude Sonnet family.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Ox Alpha is a reasoning model designed for coding, sustained agentic work, and production workloads. It is suited for long-horizon software engineering, complex reasoning, and workflows that combine text with...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| openai/gpt-4o-minigmi/openai/gpt-4o-mini | 131.072K | $0.15 | $0.6 | — | |||
| @cf/moonshotai/kimi-k2.6cloudflare/@cf/moonshotai/kimi-k2.6 | 262.144K | $0.95 | $4 | — | |||
| gpt-4.1-miniazure/gpt-4.1-mini | 1.04758M | $0.4 | $1.6 | — | |||
| Inference.net: Schematron V2 Turboinference-net/schematron-v2-turbo | 128K | $0.03 | $0.15 | — | |||
| gpt-4.1-2025-04-14azure/gpt-4.1-2025-04-14 | 1.04758M | $2 | $8 | — | |||
| gpt-4.1azure/gpt-4.1 | 1.04758M | $2 | $8 | — | |||
| inclusionAI: Ling 3.0 Tiny (free)inclusionai/ling-3.0-tiny:free | 262.144K | Free | Free | — | |||
| Poolside: Laguna M.1 (free)poolside/laguna-m.1:free | 262.144K | Free | Free | — | |||
| Qwen: Qwen3 Next 80B A3B Instructqwen/qwen3-next-80b-a3b-instruct | 262.144K | $0.09 | $1.1 | — | |||
| gpt-4-turbo-vision-previewazure/gpt-4-turbo-vision-preview | 128K | $10 | $30 | — | |||
| databricks-claude-opus-5databricks/databricks-claude-opus-5 | 1M | $5 | $25 | — | |||
| amazon.nova-pro-v1:0bedrock_converse/amazon.nova-pro-v1:0 | 300K | $0.8 | $3.2 | — | |||
| gpt-4-turbo-2024-04-09azure/gpt-4-turbo-2024-04-09 | 128K | $10 | $30 | — | |||
| OpenAI: GPT-5.6 Sol Pro (batch)openai/gpt-5.6-sol-pro:batch | 1.05M | $1 | $5 | — | |||
| OpenAI: GPT Sol Latest~openai/gpt-sol-latest | 1.05M | $2 | $10 | — | |||
| Qwen: Qwen Plus 0728 (thinking)qwen/qwen-plus-2025-07-28:thinking | 1M | $0.26 | $0.78 | — | |||
| gpt-4-turboazure/gpt-4-turbo | 128K | $10 | $30 | — | |||
| gpt-4-1106-previewazure/gpt-4-1106-preview | 128K | $10 | $30 | — | |||
| amazon.nova-micro-v1:0bedrock_converse/amazon.nova-micro-v1:0 | 128K | $0.035 | $0.14 | — | |||
| databricks-claude-opus-4-8databricks/databricks-claude-opus-4-8 | 1M | $5 | $25 | — | |||
| gpt-4-0125-previewazure/gpt-4-0125-preview | 128K | $10 | $30 | — | |||
| us.amazon.nova-2-pro-preview-20251202-v1:0bedrock_converse/us.amazon.nova-2-pro-preview-20251202-v1:0 | 1M | $2.188 | $17.5 | — | |||
| sa-east-1/minimax.minimax-m2.1bedrock/sa-east-1/minimax.minimax-m2.1 | 196K | $0.36 | $1.44 | — | |||
| ap-south-1/moonshotai.kimi-k2.5bedrock/ap-south-1/moonshotai.kimi-k2.5 | 262.144K | $0.72 | $3.6 | — | |||
| databricks-claude-opus-4-7databricks/databricks-claude-opus-4-7 | 1M | $5 | $25 | — | |||
| us.amazon.nova-2-lite-v1:0bedrock_converse/us.amazon.nova-2-lite-v1:0 | 1M | $0.33 | $2.75 | — | |||
| Qwen: Qwen3 235B A22Bqwen/qwen3-235b-a22b | 131.072K | $0.455 | $1.82 | — | |||
| eu.amazon.nova-2-pro-preview-20251202-v1:0bedrock_converse/eu.amazon.nova-2-pro-preview-20251202-v1:0 | 1M | $2.188 | $17.5 | — | |||
| deepseek-ai/DeepSeek-V3.2friendliai/deepseek-ai/deepseek-v3.2 | 163.84K | $0.5 | $1.5 | — | |||
| databricks-claude-fable-5databricks/databricks-claude-fable-5 | 1M | $10 | $50 | — | |||
| eu.amazon.nova-2-lite-v1:0bedrock_converse/eu.amazon.nova-2-lite-v1:0 | 1M | $0.33 | $2.75 | — | |||
| global/gpt-5.1-chatazure/global/gpt-5.1-chat | 128K | $1.25 | $10 | — | |||
| apac.amazon.nova-2-pro-preview-20251202-v1:0bedrock_converse/apac.amazon.nova-2-pro-preview-20251202-v1:0 | 1M | $2.188 | $17.5 | — | |||
| LGAI-EXAONE/K-EXAONE-2.0-750B-A37Bfriendliai/lgai-exaone/k-exaone-2.0-750b-a37b | 262.144K | $0.6 | $2.4 | — | |||
| Z.ai: GLM 5V Turboz-ai/glm-5v-turbo | 202.752K | $1.2 | $4 | — | |||
| Mistral: Mistral Small 3.1 24Bmistralai/mistral-small-3.1-24b-instruct | 128K | $0.351 | $0.555 | — | |||
| Anthropic: Claude Sonnet Latest~anthropic/claude-sonnet-latest | 1M | $2 | $10 | — | |||
| global/gpt-5.1azure/global/gpt-5.1 | 272K | $1.25 | $10 | — | |||
| gpt-5-2025-08-07openai/gpt-5-2025-08-07 | 272K | $1.25 | $10 | — | |||
| global/gpt-4o-2024-11-20azure/global/gpt-4o-2024-11-20 | 128K | $2.5 | $10 | — | |||
| apac.amazon.nova-2-lite-v1:0bedrock_converse/apac.amazon.nova-2-lite-v1:0 | 1M | $0.33 | $2.75 | — | |||
| gpt-5.4-nano-2026-03-17openai/gpt-5.4-nano-2026-03-17 | 272K | $0.2 | $1.25 | — | |||
| gpt-5.4-mini-2026-03-17openai/gpt-5.4-mini-2026-03-17 | 272K | $0.75 | $4.5 | — | |||
| gpt-5.4-2026-03-05openai/gpt-5.4-2026-03-05 | 1.05M | $2.5 | $15 | — | |||
| gpt-5.5-2026-04-23openai/gpt-5.5-2026-04-23 | 1.05M | $5 | $30 | — | |||
| global/gpt-4o-2024-08-06azure/global/gpt-4o-2024-08-06 | 128K | $2.5 | $10 | — | |||
| Ox Alphastealth/ox-alpha | 1.04858M | — | — | — | |||
| gpt-5-mini-2025-08-07openai/gpt-5-mini-2025-08-07 | 272K | $0.25 | $2 | — | |||
| gpt-5-nano-2025-08-07openai/gpt-5-nano-2025-08-07 | 272K | $0.05 | $0.4 | — | |||
| gpt-5.6openai/gpt-5.6 | 1.05M | $5 | $30 | — |