No provider description is available for this model yet.

cerebras/qwen-3-32b 128K context $0.4/M input $0.8/M output

Inkling Small is an open-weight multimodal mixture-of-experts model from Thinking Machines Lab, with 12B active parameters out of 276B total. It is positioned as the smaller, more efficient member of...

thinkingmachines/inkling-small:batch 524.288K context $0.5/M input $1.2/M output

No provider description is available for this model yet.

azure/us/o3-2025-04-16 200K context $2.2/M input $8.8/M output

No provider description is available for this model yet.

azure/us/o3-mini-2025-01-31 200K context $1.21/M input $4.84/M output

No provider description is available for this model yet.

azure/us/o4-mini-2025-04-16 200K context $1.21/M input $4.84/M output

No provider description is available for this model yet.

azure/gpt-5.1-chat 128K context $1.25/M input $10/M output

GPT-4o mini is OpenAI's newest model after [GPT-4 Omni](/models/openai/gpt-4o), supporting both text and image inputs with text outputs. As their most advanced small model, it is many multiples more affordable...

openai/gpt-4o-mini:batch 128K context $0.075/M input $0.3/M output

No provider description is available for this model yet.

azure_ai/meta-llama-3-70b-instruct 8.192K context $1.1/M input $0.37/M output

No provider description is available for this model yet.

azure_ai/phi-3-medium-4k-instruct 4.096K context $0.17/M input $0.68/M output

No provider description is available for this model yet.

azure_ai/phi-3-mini-128k-instruct 128K context $0.13/M input $0.52/M output

No provider description is available for this model yet.

azure_ai/phi-3-mini-4k-instruct 4.096K context $0.13/M input $0.52/M output

No provider description is available for this model yet.

replicate/meta/llama-2-70b-chat 4.096K context $0.65/M input $2.75/M output

No provider description is available for this model yet.

azure_ai/phi-3-small-8k-instruct 8.192K context $0.15/M input $0.6/M output

No provider description is available for this model yet.

azure_ai/phi-3.5-moe-instruct 128K context $0.16/M input $0.64/M output

No provider description is available for this model yet.

azure_ai/phi-3.5-mini-instruct 128K context $0.13/M input $0.52/M output

No provider description is available for this model yet.

azure_ai/phi-3.5-vision-instruct 128K context $0.13/M input $0.52/M output

No provider description is available for this model yet.

azure_ai/phi-4 16.384K context $0.125/M input $0.5/M output

No provider description is available for this model yet.

azure_ai/phi-4-mini-instruct 131.072K context $0.075/M input $0.3/M output

No provider description is available for this model yet.

azure_ai/phi-4-multimodal-instruct 131.072K context $0.08/M input $0.32/M output

No provider description is available for this model yet.

together_ai/moonshotai/kimi-k3 1.04858M context $3/M input $15/M output

No provider description is available for this model yet.

azure_ai/phi-4-mini-reasoning 131.072K context $0.08/M input $0.32/M output

Mistral Medium 3 is a high-performance enterprise-grade language model designed to deliver frontier-level capabilities at significantly reduced operational cost. It balances state-of-the-art reasoning and multimodal performance with 8× lower cost...

mistralai/mistral-medium-3 131.072K context $0.4/M input $2/M output

No provider description is available for this model yet.

azure_ai/phi-4-reasoning 32.768K context $0.125/M input $0.5/M output

No provider description is available for this model yet.

azure_ai/mai-ds-r1 128K context $1.35/M input $5.4/M output

No provider description is available for this model yet.

azure_ai/deepseek-v3.2 163.84K context $0.58/M input $1.68/M output

No provider description is available for this model yet.

azure_ai/deepseek-v3.2-speciale 163.84K context $0.58/M input $1.68/M output

No provider description is available for this model yet.

azure_ai/deepseek-r1 128K context $1.35/M input $5.4/M output

No provider description is available for this model yet.

azure/gpt-5.1 272K context $1.25/M input $10/M output

No provider description is available for this model yet.

azure_ai/deepseek-v3 128K context $1.14/M input $4.56/M output

No provider description is available for this model yet.

azure_ai/deepseek-v3-0324 128K context $1.14/M input $4.56/M output

No provider description is available for this model yet.

azure_ai/deepseek-v3.1 131.072K context $1.23/M input $4.94/M output

No provider description is available for this model yet.

azure_ai/deepseek-v4-pro 1M context $1.74/M input $3.48/M output

GPT-5 Pro is OpenAI’s most advanced model, offering major improvements in reasoning, code quality, and user experience. It is optimized for complex tasks that require step-by-step reasoning, instruction following, and...

openai/gpt-5-pro:batch 400K context $7.5/M input $60/M output

Claude Sonnet 4.5 is Anthropic’s most advanced Sonnet model to date, optimized for real-world agents and coding workflows. It delivers state-of-the-art performance on coding benchmarks such as SWE-bench Verified, with...

anthropic/claude-sonnet-4.5 1M context $3/M input $15/M output

No provider description is available for this model yet.

azure_ai/deepseek-v4-flash 1M context $0.19/M input $0.51/M output

Ling 3.0 Flash VL builds on Ling 3.0 Flash (124B total / 5.5B active MoE from InclusionAI), further strengthening its language capabilities while adding native visual perception and advanced visual...

inclusionai/ling-3.0-flash-vl:free 262.144K context Free input Free output

No provider description is available for this model yet.

azure_ai/global/grok-3 131.072K context $3/M input $15/M output

No provider description is available for this model yet.

azure_ai/global/grok-3-mini 131.072K context $0.25/M input $1.27/M output