No provider description is available for this model yet.
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
Inkling Small is an open-weight multimodal mixture-of-experts model from Thinking Machines Lab, with 12B active parameters out of 276B total. It is positioned as the smaller, more efficient member of...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
GPT-4o mini is OpenAI's newest model after [GPT-4 Omni](/models/openai/gpt-4o), supporting both text and image inputs with text outputs. As their most advanced small model, it is many multiples more affordable...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Mistral Medium 3 is a high-performance enterprise-grade language model designed to deliver frontier-level capabilities at significantly reduced operational cost. It balances state-of-the-art reasoning and multimodal performance with 8× lower cost...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
GPT-5 Pro is OpenAI’s most advanced model, offering major improvements in reasoning, code quality, and user experience. It is optimized for complex tasks that require step-by-step reasoning, instruction following, and...
Claude Sonnet 4.5 is Anthropic’s most advanced Sonnet model to date, optimized for real-world agents and coding workflows. It delivers state-of-the-art performance on coding benchmarks such as SWE-bench Verified, with...
No provider description is available for this model yet.
Ling 3.0 Flash VL builds on Ling 3.0 Flash (124B total / 5.5B active MoE from InclusionAI), further strengthening its language capabilities while adding native visual perception and advanced visual...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| qwen-3-32bcerebras/qwen-3-32b | 128K | $0.4 | $0.8 | — | |||
| Thinking Machines: Inkling Small (batch)thinkingmachines/inkling-small:batch | 524.288K | $0.5 | $1.2 | — | |||
| us/o3-2025-04-16azure/us/o3-2025-04-16 | 200K | $2.2 | $8.8 | — | |||
| us/o3-mini-2025-01-31azure/us/o3-mini-2025-01-31 | 200K | $1.21 | $4.84 | — | |||
| us/o4-mini-2025-04-16azure/us/o4-mini-2025-04-16 | 200K | $1.21 | $4.84 | — | |||
| Llama-3.2-11B-Vision-Instructazure_ai/llama-3.2-11b-vision-instruct | 128K | $0.37 | $0.37 | — | |||
| gpt-5.1-chatazure/gpt-5.1-chat | 128K | $1.25 | $10 | — | |||
| deepseek-ai/ESFT-token-translation-litedeepseek-ai/ESFT-token-translation-lite | Not documented | — | — | — | |||
| stabilityai/codellama13b_instruct_260k_synthesisstabilityai/codellama13b_instruct_260k_synthesis | Not documented | — | — | — | |||
| Llama-3.2-90B-Vision-Instructazure_ai/llama-3.2-90b-vision-instruct | 128K | $2.04 | $2.04 | — | |||
| OpenAI: GPT-4o-mini (batch)openai/gpt-4o-mini:batch | 128K | $0.075 | $0.3 | — | |||
| Llama-4-Maverick-17B-128E-Instruct-FP8azure_ai/llama-4-maverick-17b-128e-instruct-fp8 | 1M | $1.41 | $0.35 | — | |||
| Llama-4-Scout-17B-16E-Instructazure_ai/llama-4-scout-17b-16e-instruct | 10M | $0.2 | $0.78 | — | |||
| Meta-Llama-3-70B-Instructazure_ai/meta-llama-3-70b-instruct | 8.192K | $1.1 | $0.37 | — | |||
| Meta-Llama-3.1-405B-Instructazure_ai/meta-llama-3.1-405b-instruct | 128K | $5.33 | $16 | — | |||
| Meta-Llama-3.1-70B-Instructazure_ai/meta-llama-3.1-70b-instruct | 128K | $2.68 | $3.54 | — | |||
| Meta-Llama-3.1-8B-Instructazure_ai/meta-llama-3.1-8b-instruct | 128K | $0.3 | $0.61 | — | |||
| Phi-3-medium-128k-instructazure_ai/phi-3-medium-128k-instruct | 128K | $0.17 | $0.68 | — | |||
| Phi-3-medium-4k-instructazure_ai/phi-3-medium-4k-instruct | 4.096K | $0.17 | $0.68 | — | |||
| Phi-3-mini-128k-instructazure_ai/phi-3-mini-128k-instruct | 128K | $0.13 | $0.52 | — | |||
| Phi-3-mini-4k-instructazure_ai/phi-3-mini-4k-instruct | 4.096K | $0.13 | $0.52 | — | |||
| Phi-3-small-128k-instructazure_ai/phi-3-small-128k-instruct | 128K | $0.15 | $0.6 | — | |||
| meta/llama-2-70b-chatreplicate/meta/llama-2-70b-chat | 4.096K | $0.65 | $2.75 | — | |||
| Phi-3-small-8k-instructazure_ai/phi-3-small-8k-instruct | 8.192K | $0.15 | $0.6 | — | |||
| Phi-3.5-MoE-instructazure_ai/phi-3.5-moe-instruct | 128K | $0.16 | $0.64 | — | |||
| Phi-3.5-mini-instructazure_ai/phi-3.5-mini-instruct | 128K | $0.13 | $0.52 | — | |||
| Phi-3.5-vision-instructazure_ai/phi-3.5-vision-instruct | 128K | $0.13 | $0.52 | — | |||
| Phi-4azure_ai/phi-4 | 16.384K | $0.125 | $0.5 | — | |||
| Phi-4-mini-instructazure_ai/phi-4-mini-instruct | 131.072K | $0.075 | $0.3 | — | |||
| Phi-4-multimodal-instructazure_ai/phi-4-multimodal-instruct | 131.072K | $0.08 | $0.32 | — | |||
| moonshotai/Kimi-K3together_ai/moonshotai/kimi-k3 | 1.04858M | $3 | $15 | — | |||
| Phi-4-mini-reasoningazure_ai/phi-4-mini-reasoning | 131.072K | $0.08 | $0.32 | — | |||
| Mistral: Mistral Medium 3mistralai/mistral-medium-3 | 131.072K | $0.4 | $2 | — | |||
| Phi-4-reasoningazure_ai/phi-4-reasoning | 32.768K | $0.125 | $0.5 | — | |||
| MAI-DS-R1azure_ai/mai-ds-r1 | 128K | $1.35 | $5.4 | — | |||
| deepseek-v3.2azure_ai/deepseek-v3.2 | 163.84K | $0.58 | $1.68 | — | |||
| deepseek-v3.2-specialeazure_ai/deepseek-v3.2-speciale | 163.84K | $0.58 | $1.68 | — | |||
| deepseek-r1azure_ai/deepseek-r1 | 128K | $1.35 | $5.4 | — | |||
| gpt-5.1azure/gpt-5.1 | 272K | $1.25 | $10 | — | |||
| deepseek-v3azure_ai/deepseek-v3 | 128K | $1.14 | $4.56 | — | |||
| deepseek-v3-0324azure_ai/deepseek-v3-0324 | 128K | $1.14 | $4.56 | — | |||
| deepseek-v3.1azure_ai/deepseek-v3.1 | 131.072K | $1.23 | $4.94 | — | |||
| deepseek-v4-proazure_ai/deepseek-v4-pro | 1M | $1.74 | $3.48 | — | |||
| OpenAI: GPT-5 Pro (batch)openai/gpt-5-pro:batch | 400K | $7.5 | $60 | — | |||
| Anthropic: Claude Sonnet 4.5anthropic/claude-sonnet-4.5 | 1M | $3 | $15 | — | |||
| deepseek-v4-flashazure_ai/deepseek-v4-flash | 1M | $0.19 | $0.51 | — | |||
| inclusionAI: Ling 3.0 Flash VL (free)inclusionai/ling-3.0-flash-vl:free | 262.144K | Free | Free | — | |||
| global/grok-3azure_ai/global/grok-3 | 131.072K | $3 | $15 | — | |||
| global/grok-3-miniazure_ai/global/grok-3-mini | 131.072K | $0.25 | $1.27 | — | |||
| anthropic.claude-3-opus-20240229-v1:0bedrock/anthropic.claude-3-opus-20240229-v1:0 | 200K | $15 | $75 | — |