No provider description is available for this model yet.
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
GPT-5.6 Luna Pro is the same underlying model as [GPT-5.6 Luna](https://openrouter.ai/openai/gpt-5.6-luna), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. Learn more in OpenAI's docs: https://developers.openai.com/api/docs/guides/reasoning#reasoning-mode
No provider description is available for this model yet.
No provider description is available for this model yet.
This model always redirects to the latest model in the GPT Astra family.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Claude Opus 5 is Anthropic’s flagship model for demanding reasoning, coding, and long-horizon agentic work. It is particularly strong at end-to-end software tasks, code review and bug finding, visual analysis...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Schematron V2 Turbo is a 3B-parameter HTML-to-JSON extraction model from Inference.net. It prioritizes throughput for high-volume extraction workloads. Extraction instructions must be supplied through a JSON schema in response_format rather...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Qwen3-Next-80B-A3B-Thinking is a reasoning-first chat model in the Qwen3-Next line that outputs structured “thinking” traces by default. It’s designed for hard multi-step problems; math proofs, code synthesis/debugging, logic, and agentic...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Ling 3.0 Tiny is a mixture-of-experts model from InclusionAI, with 1.3B active parameters out of 7.9B total. It is designed for responsive agents, instruction following, and multi-turn conversations, with switchable...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Grok 4.6 is SpaceXAI's smartest model with frontier performance on coding, knowledge work, and STEM.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
This model always redirects to the latest model in the GPT Sol family.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| gemini-3.7-flashvertex_ai-language-models/gemini-3.7-flash | 1.04858M | $0.75 | $3.75 | — | |||
| OpenAI: GPT-5.6 Luna Pro (batch)openai/gpt-5.6-luna-pro:batch | 1.05M | $0.1 | $0.6 | — | |||
| anthropic.claude-opus-4-8bedrock_converse/anthropic.claude-opus-4-8 | 1M | $5 | $25 | — | |||
| meta-llama/Llama-2-13b-hfmeta-llama/Llama-2-13b-hf | Not documented | — | — | — | |||
| OpenAI: GPT Astra Latest~openai/gpt-astra-latest | 1.05M | $10 | $50 | — | |||
| jp.anthropic.claude-opus-5bedrock_converse/jp.anthropic.claude-opus-5 | 1M | $5.5 | $27.5 | — | |||
| @cf/qwen/qwen3-30b-a3b-fp8cloudflare/@cf/qwen/qwen3-30b-a3b-fp8 | 32.768K | $0.051 | $0.335 | — | |||
| eu.twelvelabs.pegasus-1-2-v1:0bedrock/eu.twelvelabs.pegasus-1-2-v1:0 | Not documented | — | $7.5 | — | |||
| Claude Opus 5 (batch)anthropic/claude-opus-5:batch | 1M | $2.5 | $12.5 | — | |||
| meta-llama/Llama-3.2-3B-Instruct-QLORA_INT4_EO8meta-llama/Llama-3.2-3B-Instruct-QLORA_INT4_EO8 | Not documented | — | — | — | |||
| google/gemma-4-26B-A4B-it-assistantgoogle/gemma-4-26B-A4B-it-assistant | Not documented | — | — | — | |||
| au.anthropic.claude-opus-5bedrock_converse/au.anthropic.claude-opus-5 | 1M | $5.5 | $27.5 | — | |||
| us.twelvelabs.pegasus-1-2-v1:0bedrock/us.twelvelabs.pegasus-1-2-v1:0 | Not documented | — | $7.5 | — | |||
| eu.anthropic.claude-opus-5bedrock_converse/eu.anthropic.claude-opus-5 | 1M | $5.5 | $27.5 | — | |||
| microsoft/UniRG-CXRmicrosoft/UniRG-CXR | Not documented | — | — | — | |||
| twelvelabs.pegasus-1-2-v1:0bedrock/twelvelabs.pegasus-1-2-v1:0 | Not documented | — | $7.5 | — | |||
| meta-llama/Llama-3.2-11B-Vision-Instructmeta-llama/Llama-3.2-11B-Vision-Instruct | Not documented | — | — | — | |||
| accounts/fireworks/models/qwen3p8-maxfireworks_ai/accounts/fireworks/models/qwen3p8-max | 262.144K | $2 | $6 | — | |||
| qwen3p8-maxfireworks_ai/qwen3p8-max | 262.144K | $2 | $6 | — | |||
| glm-5p2-fast-usfireworks_ai/glm-5p2-fast-us | 1.04858M | $2.1 | $6.6 | — | |||
| deepseek-v4-flash-0731fireworks_ai/deepseek-v4-flash-0731 | 1.04858M | $0.14 | $0.28 | — | |||
| us-gov.nvidia.nemotron-nano-12b-v2bedrock_converse/us-gov.nvidia.nemotron-nano-12b-v2 | 128K | $0.24 | $0.72 | — | |||
| swe-1.7-lightningcognition/swe-1.7-lightning | Not documented | $2.5 | $12.5 | — | |||
| us-gov-west-1/anthropic.claude-3-haiku-20240307-v1:0bedrock/us-gov-west-1/anthropic.claude-3-haiku-20240307-v1:0 | 200K | $0.3 | $1.5 | — | |||
| Inference.net: Schematron V2 Turboinference-net/schematron-v2-turbo | 128K | $0.03 | $0.15 | — | |||
| gemini-3.7-flashgemini/gemini-3.7-flash | 1.04858M | $0.75 | $3.75 | — | |||
| upstage/Solar-Open2-250Bupstage/Solar-Open2-250B | Not documented | — | — | — | |||
| meta-llama/Llama-2-70b-hfmeta-llama/Llama-2-70b-hf | Not documented | — | — | — | |||
| Qwen: Qwen3 Next 80B A3B Thinkingqwen/qwen3-next-80b-a3b-thinking | 262.144K | $0.15 | $1.2 | — | |||
| gemini-3.7-flashvertex_ai/gemini-3.7-flash | 1.04858M | $0.75 | $3.75 | — | |||
| nvidia/NVIDIA-Nemotron-3.5-Lightningdeepinfra/nvidia/nvidia-nemotron-3.5-lightning | 262.144K | $0.05 | $0.2 | — | |||
| qwen3.8-maxdashscope/qwen3.8-max | 991.808K | $2 | $6 | — | |||
| us-gov-west-1/anthropic.claude-3-5-sonnet-20240620-v1:0bedrock/us-gov-west-1/anthropic.claude-3-5-sonnet-20240620-v1:0 | 200K | $3.6 | $18 | — | |||
| inclusionAI: Ling 3.0 Tiny (free)inclusionai/ling-3.0-tiny:free | 262.144K | Free | Free | — | |||
| meta-llama/Llama-3.1-70B-Instructmeta-llama/Llama-3.1-70B-Instruct | Not documented | — | — | — | |||
| gpt-4.1-2025-04-14github_copilot/gpt-4.1-2025-04-14 | 128K | — | — | — | |||
| kimi-k2.7-codedashscope/kimi-k2.7-code | 229.376K | $0.95 | $4 | — | |||
| SpaceXAI: Grok 4.6x-ai/grok-4.6 | 500K | $2 | $6 | — | |||
| amazon.nova-pro-v1:0bedrock_converse/amazon.nova-pro-v1:0 | 300K | $0.8 | $3.2 | — | |||
| glm-5.2dashscope/glm-5.2 | 1.04858M | $1.4 | $4.4 | — | |||
| us.anthropic.claude-opus-5bedrock_converse/us.anthropic.claude-opus-5 | 1M | $5.5 | $27.5 | — | |||
| OpenAI: GPT Sol Latest~openai/gpt-sol-latest | 1.05M | $2 | $10 | — | |||
| gpt-4o-2024-08-06github_copilot/gpt-4o-2024-08-06 | 64K | — | — | — | |||
| global.anthropic.claude-opus-5bedrock_converse/global.anthropic.claude-opus-5 | 1M | $5 | $25 | — | |||
| anthropic.claude-opus-5bedrock_converse/anthropic.claude-opus-5 | 1M | $5 | $25 | — | |||
| eu.anthropic.claude-fable-5bedrock_converse/eu.anthropic.claude-fable-5 | 1M | $11 | $55 | — | |||
| us.anthropic.claude-fable-5bedrock_converse/us.anthropic.claude-fable-5 | 1M | $11 | $55 | — | |||
| gpt-4o-2024-05-13github_copilot/gpt-4o-2024-05-13 | 64K | — | — | — | |||
| gpt-4.1github_copilot/gpt-4.1 | 128K | — | — | — | |||
| kimi-k3-usfireworks_ai/kimi-k3-us | 1.04858M | $3.3 | $16.5 | — |