2,851 models

Inkling Small is an open-weight multimodal mixture-of-experts model from Thinking Machines Lab, with 12B active parameters out of 276B total. It is positioned as the smaller, more efficient member of...

thinkingmachines/inkling-small:free 1.04858M context Free input Free output

No provider description is available for this model yet.

azure/us/o1-preview-2024-09-12 128K context $16.5/M input $66/M output

Opus 4.6 is Anthropic’s strongest model for coding and long-running professional tasks. It is built for agents that operate across entire workflows rather than single prompts, making it especially effective...

anthropic/claude-opus-4.6 1M context $5/M input $25/M output

No provider description is available for this model yet.

azure/us/o3-2025-04-16 200K context $2.2/M input $8.8/M output

No provider description is available for this model yet.

azure/us/o3-mini-2025-01-31 200K context $1.21/M input $4.84/M output

No provider description is available for this model yet.

azure/us/o4-mini-2025-04-16 200K context $1.21/M input $4.84/M output

Inkling Small is an open-weight multimodal mixture-of-experts model from Thinking Machines Lab, with 12B active parameters out of 276B total. It is positioned as the smaller, more efficient member of...

thinkingmachines/inkling-small:batch 524.288K context $0.5/M input $1.2/M output

GPT-5.6 Luna is a fast, cost-efficient model in OpenAI's GPT-5.6 series. It is suited for high-volume, latency-sensitive tasks such as chat, classification, and lightweight agentic workflows, providing capable reasoning for...

openai/gpt-5.6-luna:batch 1.05M context $0.1/M input $0.6/M output

GLM 5.2 is a large-scale reasoning model from Z.ai. It supports text input and output with a 1M-token context window, and is suited for long-horizon agent workflows, project-level software engineering,...

z-ai/glm-5.2:batch 1.04858M context $0.7/M input $2.2/M output

No provider description is available for this model yet.

azure_ai/llama-3.3-70b-instruct 128K context $0.71/M input $0.71/M output

No provider description is available for this model yet.

openrouter/minimax/minimax-m2.7 204.8K context $0.3/M input $1.2/M output

MiniMax-M3 is a multimodal foundation model from MiniMax. It supports text, image, and video inputs with text output, a 1M-token context window, and is suited for long-horizon agentic work, coding,...

minimax/minimax-m3 524.288K context $0.3/M input $1.2/M output

No provider description is available for this model yet.

azure_ai/phi-3-mini-128k-instruct 128K context $0.13/M input $0.52/M output

GLM-5V-Turbo is Z.ai’s first native multimodal agent foundation model, built for vision-based coding and agent-driven tasks. It natively handles image, video, and text inputs, excels at long-horizon planning, complex coding,...

z-ai/glm-5v-turbo 202.752K context $1.2/M input $4/M output

No provider description is available for this model yet.

azure_ai/phi-3.5-moe-instruct 128K context $0.16/M input $0.64/M output

No provider description is available for this model yet.

azure_ai/phi-3.5-mini-instruct 128K context $0.13/M input $0.52/M output

No provider description is available for this model yet.

azure_ai/phi-3.5-vision-instruct 128K context $0.13/M input $0.52/M output

No provider description is available for this model yet.

azure_ai/phi-4-mini-instruct 131.072K context $0.075/M input $0.3/M output

No provider description is available for this model yet.

azure_ai/grok-4.6 200K context $2/M input $6/M output

Mistral's cutting-edge language model for coding released end of July 2025. Codestral specializes in low-latency, high-frequency tasks such as fill-in-the-middle (FIM), code correction and test generation. [Blog Post](https://mistral.ai/news/codestral-25-08)

mistralai/codestral-2508 256K context $0.3/M input $0.9/M output

No provider description is available for this model yet.

azure_ai/phi-4-mini-reasoning 131.072K context $0.08/M input $0.32/M output

Gemini 3.1 Pro Preview is Google’s frontier reasoning model, delivering enhanced software engineering performance, improved agentic reliability, and more efficient token usage across complex workflows. Building on the multimodal foundation...

google/gemini-3.1-pro-preview:batch 1.04858M context $1/M input $6/M output

No provider description is available for this model yet.

azure_ai/mai-ds-r1 128K context $1.35/M input $5.4/M output

No provider description is available for this model yet.

azure_ai/deepseek-v3.2 163.84K context $0.58/M input $1.68/M output

No provider description is available for this model yet.

azure_ai/deepseek-v3.2-speciale 163.84K context $0.58/M input $1.68/M output

No provider description is available for this model yet.

azure_ai/deepseek-r1 128K context $1.35/M input $5.4/M output

GLM-5.1 delivers a major leap in coding capability, with particularly significant gains in handling long-horizon tasks. Unlike previous models built around minute-level interactions, GLM-5.1 can work independently and continuously on...

z-ai/glm-5.1 200K context $0.966/M input $3.036/M output

No provider description is available for this model yet.

azure_ai/deepseek-v3 128K context $1.14/M input $4.56/M output

No provider description is available for this model yet.

azure_ai/deepseek-v3-0324 128K context $1.14/M input $4.56/M output

No provider description is available for this model yet.

azure_ai/deepseek-v3.1 131.072K context $1.23/M input $4.94/M output

No provider description is available for this model yet.

azure_ai/deepseek-v4-pro 1M context $1.74/M input $3.48/M output

No provider description is available for this model yet.

azure_ai/deepseek-v4-flash 1M context $0.19/M input $0.51/M output

Claude Opus 4.5 is Anthropic’s frontier reasoning model optimized for complex software engineering, agentic workflows, and long-horizon computer use. It offers strong multimodal capabilities, competitive performance across real-world coding and...

anthropic/claude-opus-4.5:batch 200K context $2.5/M input $12.5/M output

No provider description is available for this model yet.

azure_ai/global/grok-3 131.072K context $3/M input $15/M output

No provider description is available for this model yet.

azure_ai/global/grok-3-mini 131.072K context $0.25/M input $1.27/M output

No provider description is available for this model yet.

azure_ai/grok-3 131.072K context $3/M input $15/M output

No provider description is available for this model yet.

azure_ai/grok-3-mini 131.072K context $0.25/M input $1.27/M output

No provider description is available for this model yet.

azure_ai/grok-4 131.072K context $3/M input $15/M output

No provider description is available for this model yet.

azure_ai/grok-4-fast-non-reasoning 131.072K context $0.2/M input $0.5/M output

The simplest way to get free inference. openrouter/free is a router that selects free models at random from the models available on OpenRouter. The router smartly filters for models that...

openrouter/free 200K context Free input Free output