No provider description is available for this model yet.
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
No provider description is available for this model yet.
No provider description is available for this model yet.
Gemini 2.5 Flash is Google's state-of-the-art workhorse model, specifically designed for advanced reasoning, coding, mathematics, and scientific tasks. It includes built-in "thinking" capabilities, enabling it to provide responses with greater...
Ministral 8B is an 8B parameter model featuring a unique interleaved sliding-window attention pattern for faster, memory-efficient inference. Designed for edge use cases, it supports up to 128k context length...
GLM-5 is Z.ai’s flagship open-source foundation model engineered for complex systems design and long-horizon agent workflows. Built for expert developers, it delivers production-grade performance on large-scale programming tasks, rivaling leading...
Fugu Max is the cost-performance model in Sakana AI's Fugu family. Rather than a single monolithic model, Fugu is a learned multi-agent orchestration system: a language model trained to route...
Qwen3-Next-80B-A3B-Thinking is a reasoning-first chat model in the Qwen3-Next line that outputs structured “thinking” traces by default. It’s designed for hard multi-step problems; math proofs, code synthesis/debugging, logic, and agentic...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
GLM-4.5-Air is the lightweight variant of our latest flagship model family, also purpose-built for agent-centric applications. Like GLM-4.5, it adopts the Mixture-of-Experts (MoE) architecture but with a more compact parameter...
OpenAI o4-mini-high is the same model as [o4-mini](/openai/o4-mini) with reasoning_effort set to high. OpenAI o4-mini is a compact reasoning model in the o-series, optimized for fast, cost-efficient performance while retaining...
No provider description is available for this model yet.
Granite 4.2 8B is a dense reasoning model from IBM. It is suited for mathematics, code generation, multilingual dialogue, and agentic workflows that need multi-step reasoning. It supports full, low-effort,...
Claude Fable 5 is a Mythos-class model from Anthropic, built for autonomous knowledge work and coding. It supports text, image, and file inputs with text output, with reasoning support and...
Claude Opus 4.8 is Anthropic's most capable generally available model in the Opus family. It supports text, image, and file inputs with text output, with reasoning support and a 1M-token...
Qwen 3.6 Plus builds on a hybrid architecture that combines efficient linear attention with sparse mixture-of-experts routing, enabling strong scalability and high-performance inference. Compared to the 3.5 series, it delivers...
This model always redirects to the latest model in the GPT Astra family.
Qwen3.8 2.4T A95B is an open-weight sparse mixture-of-experts model from Qwen and the open-weight variant of [Qwen3.8 Max](/qwen/qwen3.8-max), with 95 billion active parameters out of 2.4 trillion total. It is...
GPT-5 Pro is OpenAI’s most advanced model, offering major improvements in reasoning, code quality, and user experience. It is optimized for complex tasks that require step-by-step reasoning, instruction following, and...
GPT-5.6 Terra Pro is the same underlying model as [GPT-5.6 Terra](https://openrouter.ai/openai/gpt-5.6-terra), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. Learn more in OpenAI's docs: https://developers.openai.com/api/docs/guides/reasoning#reasoning-mode
Qwen2.5-VL is proficient in recognizing common objects such as flowers, birds, fish, and insects. It is also highly capable of analyzing texts, charts, icons, graphics, and layouts within images.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
MiniMax-M1 is a large-scale, open-weight reasoning model designed for extended context and high-efficiency inference. It leverages a hybrid Mixture-of-Experts (MoE) architecture paired with a custom "lightning attention" mechanism, allowing it...
No provider description is available for this model yet.
Grok 4.20 is a reasoning model from SpaceXAI with industry-leading speed and agentic tool calling capabilities. It combines the lowest hallucination rate on the market with strict prompt adherance, delivering...
A recreation trial of the original MythoMax-L2-B13 but with updated models. #merge
Opus 4.7 is the next generation of Anthropic's Opus family, built for long-running, asynchronous agents. Building on the coding and agentic strengths of Opus 4.6, it delivers stronger performance on...
One of the highest performing and most popular fine-tunes of Llama 2 13B, with rich descriptions and roleplay. #merge
No provider description is available for this model yet.
Qwen3.6 Flash is a fast, efficient language model from Alibaba's Qwen 3.6 series. It supports text, image, and video input with a 1M token context window. Tiered pricing kicks in...
GPT-5.3 Chat is an update to ChatGPT's most-used model that makes everyday conversations smoother, more useful, and more directly helpful. It delivers more accurate answers with better contextualization and significantly...
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| nvidia/NVIDIA-Nemotron-Labs-Teacher-Competition-Codingnvidia/NVIDIA-Nemotron-Labs-Teacher-Competition-Coding | Not documented | — | — | — | |||
| openai/gpt-oss-20breplicate/openai/gpt-oss-20b | Not documented | $0.09 | $0.36 | — | |||
| nvidia/NVIDIA-Nemotron-Labs-Teacher-Chatnvidia/NVIDIA-Nemotron-Labs-Teacher-Chat | Not documented | — | — | — | |||
| Google: Gemini 2.5 Flash (batch)google/gemini-2.5-flash:batch | 1.04858M | $0.15 | $1.25 | — | |||
| Mistral: Ministral 8Bmistralai/ministral-8b | 128K | $0.11 | $0.11 | — | |||
| Z.ai: GLM 5z-ai/glm-5 | 198K | $0.6 | $1.92 | — | |||
| Sakana: Fugu Maxsakana/fugu-max | 1M | $2 | $6 | — | |||
| Qwen: Qwen3 Next 80B A3B Thinkingqwen/qwen3-next-80b-a3b-thinking | 262.144K | $0.15 | $1.2 | — | |||
| openai/gpt-5.6-solopenrouter/openai/gpt-5.6-sol | 1.05M | $2 | $10 | — | |||
| deepseek-ai/DeepSeek-V4-Pro-0813wandb/deepseek-ai/deepseek-v4-pro-0813 | Not documented | $1.31 | $3.96 | — | |||
| ibm-granite/granite-4.2-8bwandb/ibm-granite/granite-4.2-8b | Not documented | $0.1 | $0.15 | — | |||
| Z.ai: GLM 4.5 Airz-ai/glm-4.5-air | 131.072K | $0.13 | $0.85 | — | |||
| OpenAI: o4 Mini Highopenai/o4-mini-high | 200K | $1.1 | $4.4 | — | |||
| nvidia/NVIDIA-Nemotron-Labs-Teacher-General-Reasoningnvidia/NVIDIA-Nemotron-Labs-Teacher-General-Reasoning | Not documented | — | — | — | |||
| IBM: Granite 4.2 8Bibm-granite/granite-4.2-8b | 131.072K | $0.06 | $0.25 | — | |||
| Anthropic: Claude Fable 5 (batch)anthropic/claude-fable-5:batch | 1M | $5 | $25 | — | |||
| Anthropic: Claude Opus 4.8 (batch)anthropic/claude-opus-4.8:batch | 1M | $2.5 | $12.5 | — | |||
| Qwen: Qwen3.6 Plusqwen/qwen3.6-plus | 1M | $0.325 | $1.95 | — | |||
| OpenAI: GPT Astra Latest~openai/gpt-astra-latest | 1.05M | $10 | $50 | — | |||
| Qwen: Qwen3.8 2.4T A95Bqwen/qwen3.8-2.4t-a95b | 1M | $2 | $6 | — | |||
| OpenAI: GPT-5 Pro (batch)openai/gpt-5-pro:batch | 400K | $7.5 | $60 | — | |||
| OpenAI: GPT-5.6 Terra Pro (batch)openai/gpt-5.6-terra-pro:batch | 1.05M | $1 | $6 | — | |||
| Qwen: Qwen2.5 VL 72B Instructqwen/qwen2.5-vl-72b-instruct | 128K | $0.8 | $1 | — | |||
| databricks-glm-5-2databricks/databricks-glm-5-2 | 1M | $1.4 | $4.4 | — | |||
| databricks-kimi-k3databricks/databricks-kimi-k3 | 1M | $3 | $15 | — | |||
| microsoft/Dayhoff-170M-UR90-HL-36000microsoft/Dayhoff-170M-UR90-HL-36000 | Not documented | — | — | — | |||
| accounts/fireworks/models/deepseek-v4-pro-0813fireworks_ai/accounts/fireworks/models/deepseek-v4-pro-0813 | 1.04858M | $1.32 | $3.96 | — | |||
| ministral-14b-2512mistral/ministral-14b-2512 | 262.144K | $0.2 | $0.2 | — | |||
| ministral-14b-latestmistral/ministral-14b-latest | 262.144K | $0.2 | $0.2 | — | |||
| ministral-3b-2512mistral/ministral-3b-2512 | 131.072K | $0.1 | $0.1 | — | |||
| ministral-3b-latestmistral/ministral-3b-latest | 131.072K | $0.1 | $0.1 | — | |||
| mistral-medium-3mistral/mistral-medium-3 | 262.144K | $1.5 | $7.5 | — | |||
| voxtral-small-2507mistral/voxtral-small-2507 | 32.768K | $0.1 | $0.4 | — | |||
| zai-org/GLM-5.3-Flashtogether_ai/zai-org/glm-5.3-flash | 1.04858M | $0.15 | $0.5 | — | |||
| glm-5.3zai/glm-5.3 | 1M | $1.4 | $4.4 | — | |||
| minimax-m3tencent/minimax-m3 | 1M | $0.3 | $1.2 | — | |||
| zai-org/glm-5.3novita/zai-org/glm-5.3 | 1.04858M | $1.4 | $4.4 | — | |||
| deepseek/deepseek-v4-pro-0813novita/deepseek/deepseek-v4-pro-0813 | 1.04858M | $1.32 | $3.96 | — | |||
| moonshotai/kimi-k3novita/moonshotai/kimi-k3 | 1.04858M | $3 | $15 | — | |||
| tencent/hy3novita/tencent/hy3 | 262.144K | $0.14 | $0.58 | — | |||
| zai-org/glm-5.2novita/zai-org/glm-5.2 | 1.04858M | $1.4 | $4.4 | — | |||
| MiniMax: MiniMax M1minimax/minimax-m1 | 1M | $0.55 | $2.2 | — | |||
| moonshotai/kimi-k2.7-codenovita/moonshotai/kimi-k2.7-code | 262.144K | $0.95 | $4 | — | |||
| SpaceXAI: Grok 4.20x-ai/grok-4.20 | 2M | $1.25 | $2.5 | — | |||
| ReMM SLERP 13Bundi95/remm-slerp-l2-13b | 6.144K | $0.35 | $0.65 | — | |||
| Anthropic: Claude Opus 4.7anthropic/claude-opus-4.7 | 1M | $5 | $25 | — | |||
| MythoMax 13Bgryphe/mythomax-l2-13b | 4.096K | $0.06 | $0.06 | — | |||
| deepseek-ai/ESFT-gate-law-litedeepseek-ai/ESFT-gate-law-lite | Not documented | — | — | — | |||
| Qwen: Qwen3.6 Flashqwen/qwen3.6-flash | 1M | $0.188 | $1.125 | — | |||
| OpenAI: GPT-5.3 Chatopenai/gpt-5.3-chat | 128K | $1.75 | $14 | — |