No provider description is available for this model yet.
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Qwen3-Next-80B-A3B-Instruct is an instruction-tuned chat model in the Qwen3-Next series optimized for fast, stable responses without “thinking” traces. It targets complex tasks across reasoning, code generation, knowledge QA, and multilingual...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Mistral Medium 3.5 is a dense 128B instruction-following model from Mistral AI. It supports text and image inputs with text output, and is designed for agentic workflows, coding, and complex...
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| databricks-mpt-30b-instructdatabricks/databricks-mpt-30b-instruct | 8.192K | $1 | $1 | — | |||
| databricks-mpt-7b-instructdatabricks/databricks-mpt-7b-instruct | 8.192K | $0.5 | — | — | |||
| NousResearch/Hermes-3-Llama-3.1-405Bdeepinfra/nousresearch/hermes-3-llama-3.1-405b | 131.072K | $1 | $1 | — | |||
| NousResearch/Hermes-3-Llama-3.1-70Bdeepinfra/nousresearch/hermes-3-llama-3.1-70b | 131.072K | $0.3 | $0.3 | — | |||
| Qwen/QwQ-32Bdeepinfra/qwen/qwq-32b | 131.072K | $0.15 | $0.4 | — | |||
| Qwen/Qwen2.5-72B-Instructdeepinfra/qwen/qwen2.5-72b-instruct | 32.768K | $0.12 | $0.39 | — | |||
| Qwen/Qwen2.5-7B-Instructdeepinfra/qwen/qwen2.5-7b-instruct | 32.768K | $0.04 | $0.1 | — | |||
| Qwen/Qwen2.5-VL-32B-Instructdeepinfra/qwen/qwen2.5-vl-32b-instruct | 128K | $0.2 | $0.6 | — | |||
| Qwen/Qwen3-14Bdeepinfra/qwen/qwen3-14b | 40.96K | $0.06 | $0.24 | — | |||
| Qwen/Qwen3-235B-A22Bdeepinfra/qwen/qwen3-235b-a22b | 40.96K | $0.18 | $0.54 | — | |||
| Qwen/Qwen3-235B-A22B-Instruct-2507deepinfra/qwen/qwen3-235b-a22b-instruct-2507 | 262.144K | $0.09 | $0.6 | — | |||
| Qwen/Qwen3-235B-A22B-Thinking-2507deepinfra/qwen/qwen3-235b-a22b-thinking-2507 | 262.144K | $0.3 | $2.9 | — | |||
| Qwen/Qwen3-30B-A3Bdeepinfra/qwen/qwen3-30b-a3b | 40.96K | $0.08 | $0.29 | — | |||
| Qwen/Qwen3-32Bdeepinfra/qwen/qwen3-32b | 40.96K | $0.1 | $0.28 | — | |||
| Qwen: Qwen3 Next 80B A3B Instruct (free)qwen/qwen3-next-80b-a3b-instruct:free | 262.144K | Free | Free | — | |||
| Qwen/Qwen3-Coder-480B-A35B-Instruct-Turbodeepinfra/qwen/qwen3-coder-480b-a35b-instruct-turbo | 262.144K | $0.29 | $1.2 | — | |||
| Qwen/Qwen3-Next-80B-A3B-Instructdeepinfra/qwen/qwen3-next-80b-a3b-instruct | 262.144K | $0.14 | $1.4 | — | |||
| Qwen/Qwen3-Next-80B-A3B-Thinkingdeepinfra/qwen/qwen3-next-80b-a3b-thinking | 262.144K | $0.14 | $1.4 | — | |||
| Sao10K/L3-8B-Lunaris-v1-Turbodeepinfra/sao10k/l3-8b-lunaris-v1-turbo | 8.192K | $0.04 | $0.05 | — | |||
| Sao10K/L3.1-70B-Euryale-v2.2deepinfra/sao10k/l3.1-70b-euryale-v2.2 | 131.072K | $0.65 | $0.75 | — | |||
| Sao10K/L3.3-70B-Euryale-v2.3deepinfra/sao10k/l3.3-70b-euryale-v2.3 | 131.072K | $0.65 | $0.75 | — | |||
| allenai/olmOCR-7B-0725-FP8deepinfra/allenai/olmocr-7b-0725-fp8 | 16.384K | $0.27 | $1.5 | — | |||
| anthropic/claude-3-7-sonnet-latestdeepinfra/anthropic/claude-3-7-sonnet-latest | 200K | $3.3 | $16.5 | — | |||
| anthropic/claude-4-opusdeepinfra/anthropic/claude-4-opus | 200K | $16.5 | $82.5 | — | |||
| anthropic/claude-4-sonnetdeepinfra/anthropic/claude-4-sonnet | 200K | $3.3 | $16.5 | — | |||
| deepseek-ai/DeepSeek-R1deepinfra/deepseek-ai/deepseek-r1 | 163.84K | $0.7 | $2.4 | — | |||
| deepseek-ai/DeepSeek-R1-0528deepinfra/deepseek-ai/deepseek-r1-0528 | 163.84K | $0.5 | $2.15 | — | |||
| deepseek-ai/DeepSeek-R1-0528-Turbodeepinfra/deepseek-ai/deepseek-r1-0528-turbo | 32.768K | $1 | $3 | — | |||
| deepseek-ai/DeepSeek-R1-Distill-Llama-70Bdeepinfra/deepseek-ai/deepseek-r1-distill-llama-70b | 131.072K | $0.2 | $0.6 | — | |||
| deepseek-ai/DeepSeek-R1-Distill-Qwen-32Bdeepinfra/deepseek-ai/deepseek-r1-distill-qwen-32b | 131.072K | $0.27 | $0.27 | — | |||
| deepseek-ai/DeepSeek-R1-Turbodeepinfra/deepseek-ai/deepseek-r1-turbo | 40.96K | $1 | $3 | — | |||
| deepseek-ai/DeepSeek-V3deepinfra/deepseek-ai/deepseek-v3 | 163.84K | $0.38 | $0.89 | — | |||
| deepseek-ai/DeepSeek-V3-0324deepinfra/deepseek-ai/deepseek-v3-0324 | 163.84K | $0.25 | $0.88 | — | |||
| deepseek-ai/DeepSeek-V3.1deepinfra/deepseek-ai/deepseek-v3.1 | 163.84K | $0.27 | $1 | — | |||
| deepseek-ai/DeepSeek-V3.1-Terminusdeepinfra/deepseek-ai/deepseek-v3.1-terminus | 163.84K | $0.27 | $1 | — | |||
| google/gemini-2.0-flash-001deepinfra/google/gemini-2.0-flash-001 | 1M | $0.1 | $0.4 | — | |||
| google/gemini-2.5-flashdeepinfra/google/gemini-2.5-flash | 1M | $0.3 | $2.5 | — | |||
| fireworks-ai-up-to-4bfireworks_ai/fireworks-ai-up-to-4b | Not documented | $0.2 | $0.2 | — | |||
| google/gemma-3-12b-itdeepinfra/google/gemma-3-12b-it | 131.072K | $0.05 | $0.1 | — | |||
| google/gemma-3-27b-itdeepinfra/google/gemma-3-27b-it | 131.072K | $0.09 | $0.16 | — | |||
| google/gemma-3-4b-itdeepinfra/google/gemma-3-4b-it | 131.072K | $0.04 | $0.08 | — | |||
| meta-llama/Llama-3.2-11B-Vision-Instructdeepinfra/meta-llama/llama-3.2-11b-vision-instruct | 131.072K | $0.049 | $0.049 | — | |||
| meta-llama/Llama-3.2-3B-Instructdeepinfra/meta-llama/llama-3.2-3b-instruct | 131.072K | $0.02 | $0.02 | — | |||
| meta-llama/Llama-3.3-70B-Instructdeepinfra/meta-llama/llama-3.3-70b-instruct | 131.072K | $0.23 | $0.4 | — | |||
| meta-llama/Llama-3.3-70B-Instruct-Turbodeepinfra/meta-llama/llama-3.3-70b-instruct-turbo | 131.072K | $0.13 | $0.39 | — | |||
| meta-llama/Llama-4-Maverick-17B-128E-Instruct-FP8deepinfra/meta-llama/llama-4-maverick-17b-128e-instruct-fp8 | 1.04858M | $0.15 | $0.6 | — | |||
| meta-llama/Llama-4-Scout-17B-16E-Instructdeepinfra/meta-llama/llama-4-scout-17b-16e-instruct | 327.68K | $0.08 | $0.3 | — | |||
| meta-llama/Llama-Guard-3-8Bdeepinfra/meta-llama/llama-guard-3-8b | 131.072K | $0.055 | $0.055 | — | |||
| meta-llama/Llama-Guard-4-12Bdeepinfra/meta-llama/llama-guard-4-12b | 163.84K | $0.18 | $0.18 | — | |||
| Mistral: Mistral Medium 3.5mistralai/mistral-medium-3-5 | 262.144K | $1.5 | $7.5 | — |