Inkling Small is an open-weight multimodal mixture-of-experts model from Thinking Machines Lab, with 12B active parameters out of 276B total. It is positioned as the smaller, more efficient member of...
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
No provider description is available for this model yet.
Opus 4.6 is Anthropic’s strongest model for coding and long-running professional tasks. It is built for agents that operate across entire workflows rather than single prompts, making it especially effective...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Inkling Small is an open-weight multimodal mixture-of-experts model from Thinking Machines Lab, with 12B active parameters out of 276B total. It is positioned as the smaller, more efficient member of...
GPT-5.6 Luna is a fast, cost-efficient model in OpenAI's GPT-5.6 series. It is suited for high-volume, latency-sensitive tasks such as chat, classification, and lightweight agentic workflows, providing capable reasoning for...
No provider description is available for this model yet.
GLM 5.2 is a large-scale reasoning model from Z.ai. It supports text input and output with a 1M-token context window, and is suited for long-horizon agent workflows, project-level software engineering,...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
MiniMax-M3 is a multimodal foundation model from MiniMax. It supports text, image, and video inputs with text output, a 1M-token context window, and is suited for long-horizon agentic work, coding,...
No provider description is available for this model yet.
No provider description is available for this model yet.
GLM-5V-Turbo is Z.ai’s first native multimodal agent foundation model, built for vision-based coding and agent-driven tasks. It natively handles image, video, and text inputs, excels at long-horizon planning, complex coding,...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Mistral's cutting-edge language model for coding released end of July 2025. Codestral specializes in low-latency, high-frequency tasks such as fill-in-the-middle (FIM), code correction and test generation. [Blog Post](https://mistral.ai/news/codestral-25-08)
No provider description is available for this model yet.
Gemini 3.1 Pro Preview is Google’s frontier reasoning model, delivering enhanced software engineering performance, improved agentic reliability, and more efficient token usage across complex workflows. Building on the multimodal foundation...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
GLM-5.1 delivers a major leap in coding capability, with particularly significant gains in handling long-horizon tasks. Unlike previous models built around minute-level interactions, GLM-5.1 can work independently and continuously on...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Claude Opus 4.5 is Anthropic’s frontier reasoning model optimized for complex software engineering, agentic workflows, and long-horizon computer use. It offers strong multimodal capabilities, competitive performance across real-world coding and...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
The simplest way to get free inference. openrouter/free is a router that selects free models at random from the models available on OpenRouter. The router smartly filters for models that...
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| Thinking Machines: Inkling Small (free)thinkingmachines/inkling-small:free | 1.04858M | Free | Free | — | |||
| us/o1-preview-2024-09-12azure/us/o1-preview-2024-09-12 | 128K | $16.5 | $66 | — | |||
| Anthropic: Claude Opus 4.6anthropic/claude-opus-4.6 | 1M | $5 | $25 | — | |||
| us/o3-2025-04-16azure/us/o3-2025-04-16 | 200K | $2.2 | $8.8 | — | |||
| us/o3-mini-2025-01-31azure/us/o3-mini-2025-01-31 | 200K | $1.21 | $4.84 | — | |||
| us/o4-mini-2025-04-16azure/us/o4-mini-2025-04-16 | 200K | $1.21 | $4.84 | — | |||
| Llama-3.2-11B-Vision-Instructazure_ai/llama-3.2-11b-vision-instruct | 128K | $0.37 | $0.37 | — | |||
| Thinking Machines: Inkling Small (batch)thinkingmachines/inkling-small:batch | 524.288K | $0.5 | $1.2 | — | |||
| OpenAI: GPT-5.6 Luna (batch)openai/gpt-5.6-luna:batch | 1.05M | $0.1 | $0.6 | — | |||
| Llama-3.2-90B-Vision-Instructazure_ai/llama-3.2-90b-vision-instruct | 128K | $2.04 | $2.04 | — | |||
| Z.ai: GLM 5.2 (batch)z-ai/glm-5.2:batch | 1.04858M | $0.7 | $2.2 | — | |||
| Llama-3.3-70B-Instructazure_ai/llama-3.3-70b-instruct | 128K | $0.71 | $0.71 | — | |||
| Llama-4-Maverick-17B-128E-Instruct-FP8azure_ai/llama-4-maverick-17b-128e-instruct-fp8 | 1M | $1.41 | $0.35 | — | |||
| Llama-4-Scout-17B-16E-Instructazure_ai/llama-4-scout-17b-16e-instruct | 10M | $0.2 | $0.78 | — | |||
| minimax/minimax-m2.7openrouter/minimax/minimax-m2.7 | 204.8K | $0.3 | $1.2 | — | |||
| Meta-Llama-3.1-70B-Instructazure_ai/meta-llama-3.1-70b-instruct | 128K | $2.68 | $3.54 | — | |||
| Meta-Llama-3.1-8B-Instructazure_ai/meta-llama-3.1-8b-instruct | 128K | $0.3 | $0.61 | — | |||
| Phi-3-medium-128k-instructazure_ai/phi-3-medium-128k-instruct | 128K | $0.17 | $0.68 | — | |||
| MiniMax: MiniMax M3minimax/minimax-m3 | 524.288K | $0.3 | $1.2 | — | |||
| Phi-3-mini-128k-instructazure_ai/phi-3-mini-128k-instruct | 128K | $0.13 | $0.52 | — | |||
| Phi-3-small-128k-instructazure_ai/phi-3-small-128k-instruct | 128K | $0.15 | $0.6 | — | |||
| Z.ai: GLM 5V Turboz-ai/glm-5v-turbo | 202.752K | $1.2 | $4 | — | |||
| Phi-3.5-MoE-instructazure_ai/phi-3.5-moe-instruct | 128K | $0.16 | $0.64 | — | |||
| Phi-3.5-mini-instructazure_ai/phi-3.5-mini-instruct | 128K | $0.13 | $0.52 | — | |||
| Phi-3.5-vision-instructazure_ai/phi-3.5-vision-instruct | 128K | $0.13 | $0.52 | — | |||
| Phi-4-mini-instructazure_ai/phi-4-mini-instruct | 131.072K | $0.075 | $0.3 | — | |||
| grok-4.6azure_ai/grok-4.6 | 200K | $2 | $6 | — | |||
| Mistral: Codestral 2508mistralai/codestral-2508 | 256K | $0.3 | $0.9 | — | |||
| Phi-4-mini-reasoningazure_ai/phi-4-mini-reasoning | 131.072K | $0.08 | $0.32 | — | |||
| Google: Gemini 3.1 Pro Preview (batch)google/gemini-3.1-pro-preview:batch | 1.04858M | $1 | $6 | — | |||
| MAI-DS-R1azure_ai/mai-ds-r1 | 128K | $1.35 | $5.4 | — | |||
| deepseek-v3.2azure_ai/deepseek-v3.2 | 163.84K | $0.58 | $1.68 | — | |||
| deepseek-v3.2-specialeazure_ai/deepseek-v3.2-speciale | 163.84K | $0.58 | $1.68 | — | |||
| deepseek-r1azure_ai/deepseek-r1 | 128K | $1.35 | $5.4 | — | |||
| Z.ai: GLM 5.1z-ai/glm-5.1 | 200K | $0.966 | $3.036 | — | |||
| deepseek-v3azure_ai/deepseek-v3 | 128K | $1.14 | $4.56 | — | |||
| deepseek-v3-0324azure_ai/deepseek-v3-0324 | 128K | $1.14 | $4.56 | — | |||
| deepseek-v3.1azure_ai/deepseek-v3.1 | 131.072K | $1.23 | $4.94 | — | |||
| deepseek-v4-proazure_ai/deepseek-v4-pro | 1M | $1.74 | $3.48 | — | |||
| deepseek-v4-flashazure_ai/deepseek-v4-flash | 1M | $0.19 | $0.51 | — | |||
| Anthropic: Claude Opus 4.5 (batch)anthropic/claude-opus-4.5:batch | 200K | $2.5 | $12.5 | — | |||
| global/grok-3azure_ai/global/grok-3 | 131.072K | $3 | $15 | — | |||
| us-gov-west-1/nvidia.nemotron-nano-12b-v2bedrock/us-gov-west-1/nvidia.nemotron-nano-12b-v2 | 128K | $0.24 | $0.72 | — | |||
| global/grok-3-miniazure_ai/global/grok-3-mini | 131.072K | $0.25 | $1.27 | — | |||
| grok-3azure_ai/grok-3 | 131.072K | $3 | $15 | — | |||
| grok-3-miniazure_ai/grok-3-mini | 131.072K | $0.25 | $1.27 | — | |||
| us-gov-west-1/nvidia.nemotron-nano-3-30bbedrock/us-gov-west-1/nvidia.nemotron-nano-3-30b | 262.144K | $0.072 | $0.288 | — | |||
| grok-4azure_ai/grok-4 | 131.072K | $3 | $15 | — | |||
| grok-4-fast-non-reasoningazure_ai/grok-4-fast-non-reasoning | 131.072K | $0.2 | $0.5 | — | |||
| Free Models Routeropenrouter/free | 200K | Free | Free | — |