No provider description is available for this model yet.
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Mistral Large 3 2512 is Mistral’s most capable model to date, featuring a sparse mixture-of-experts architecture with 41B active parameters (675B total), and released under the Apache 2.0 license.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Qwen3.8 Max (0803) is the August 3, 2026 checkpoint of Qwen3.8 Max, the flagship model in Alibaba's Qwen3.8 series and the general-availability successor to the Qwen3.8 Max Preview. It is...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| meta-llama/Llama-3.2-3B-Instructdeepinfra/meta-llama/llama-3.2-3b-instruct | 131.072K | $0.02 | $0.02 | — | |||
| meta-llama/Llama-3.3-70B-Instructdeepinfra/meta-llama/llama-3.3-70b-instruct | 131.072K | $0.23 | $0.4 | — | |||
| meta-llama/Llama-3.3-70B-Instruct-Turbodeepinfra/meta-llama/llama-3.3-70b-instruct-turbo | 131.072K | $0.13 | $0.39 | — | |||
| meta-llama/Llama-4-Maverick-17B-128E-Instruct-FP8deepinfra/meta-llama/llama-4-maverick-17b-128e-instruct-fp8 | 1.04858M | $0.15 | $0.6 | — | |||
| meta-llama/Llama-4-Scout-17B-16E-Instructdeepinfra/meta-llama/llama-4-scout-17b-16e-instruct | 327.68K | $0.08 | $0.3 | — | |||
| meta-llama/Llama-Guard-3-8Bdeepinfra/meta-llama/llama-guard-3-8b | 131.072K | $0.055 | $0.055 | — | |||
| qwen3-next-80b-a3b-instructdashscope/qwen3-next-80b-a3b-instruct | 262.144K | $0.15 | $1.2 | — | |||
| meta-llama/Meta-Llama-3.1-70B-Instructdeepinfra/meta-llama/meta-llama-3.1-70b-instruct | 131.072K | $0.4 | $0.4 | — | |||
| meta-llama/Meta-Llama-3.1-70B-Instruct-Turbodeepinfra/meta-llama/meta-llama-3.1-70b-instruct-turbo | 131.072K | $0.1 | $0.28 | — | |||
| meta-llama/Meta-Llama-3.1-8B-Instructdeepinfra/meta-llama/meta-llama-3.1-8b-instruct | 131.072K | $0.03 | $0.05 | — | |||
| gpt-5.6-lunaazure/gpt-5.6-luna | 1.05M | $1 | $6 | — | |||
| global/gpt-4o-2024-11-20azure/global/gpt-4o-2024-11-20 | 128K | $2.5 | $10 | — | |||
| us-west-2/6-month-commitment/anthropic.claude-v2:1bedrock/us-west-2/6-month-commitment/anthropic.claude-v2:1 | 100K | — | — | — | |||
| gpt-oss-120bazure_ai/gpt-oss-120b | 131.072K | $0.15 | $0.6 | — | |||
| qwen3-max-2026-01-23dashscope/qwen3-max-2026-01-23 | 258.048K | — | — | — | |||
| qwen3-maxdashscope/qwen3-max | 258.048K | — | — | — | |||
| gpt-5.6-terraazure/gpt-5.6-terra | 1.05M | $2.5 | $15 | — | |||
| qwen3-max-previewdashscope/qwen3-max-preview | 258.048K | — | — | — | |||
| qwen3-coder-plus-2025-07-22dashscope/qwen3-coder-plus-2025-07-22 | 997.952K | — | — | — | |||
| Mistral: Mistral Large 3 2512 (batch)mistralai/mistral-large-2512:batch | 262.144K | $0.25 | $0.75 | — | |||
| global/gpt-4o-2024-08-06azure/global/gpt-4o-2024-08-06 | 128K | $2.5 | $10 | — | |||
| qwen3-coder-plusdashscope/qwen3-coder-plus | 997.952K | — | — | — | |||
| qwen3-coder-flash-2025-07-28dashscope/qwen3-coder-flash-2025-07-28 | 997.952K | — | — | — | |||
| gpt-5.6-solazure/gpt-5.6-sol | 1.05M | $5 | $30 | — | |||
| qwen3-coder-flashdashscope/qwen3-coder-flash | 997.952K | — | — | — | |||
| qwen3-30b-a3bdashscope/qwen3-30b-a3b | 129.024K | — | — | — | |||
| gpt-5.6azure/gpt-5.6 | 1.05M | $5 | $30 | — | |||
| global-standard/gpt-4o-miniazure/global-standard/gpt-4o-mini | 128K | $0.15 | $0.6 | — | |||
| us-west-2/6-month-commitment/anthropic.claude-v1bedrock/us-west-2/6-month-commitment/anthropic.claude-v1 | 100K | — | — | — | |||
| qwen-turbo-latestdashscope/qwen-turbo-latest | 1M | $0.05 | $0.2 | — | |||
| qwen-turbo-2025-04-28dashscope/qwen-turbo-2025-04-28 | 1M | $0.05 | $0.2 | — | |||
| eu/gpt-5.4-2026-03-05azure/eu/gpt-5.4-2026-03-05 | 1.05M | $2.75 | $16.5 | — | |||
| qwen-turbo-2024-11-01dashscope/qwen-turbo-2024-11-01 | 1M | $0.05 | $0.2 | — | |||
| google/gemini-2.5-prodeepinfra/google/gemini-2.5-pro | 1M | $1.25 | $10 | — | |||
| us/gpt-5.4-2026-03-05azure/us/gpt-5.4-2026-03-05 | 1.05M | $2.75 | $16.5 | — | |||
| qwen-turbodashscope/qwen-turbo | 129.024K | $0.05 | $0.2 | — | |||
| qwen-plus-latestdashscope/qwen-plus-latest | 997.952K | — | — | — | |||
| gpt-5.4-2026-03-05azure/gpt-5.4-2026-03-05 | 1.05M | $2.5 | $15 | — | |||
| qwen-plus-2025-09-11dashscope/qwen-plus-2025-09-11 | 997.952K | — | — | — | |||
| qwen-plus-2025-07-28dashscope/qwen-plus-2025-07-28 | 997.952K | — | — | — | |||
| Qwen: Qwen3.8 Max (0803)qwen/qwen3.8-max | 1M | $2 | $6 | — | |||
| global-standard/gpt-4o-2024-11-20azure/global-standard/gpt-4o-2024-11-20 | 128K | $2.5 | $10 | — | |||
| kimi-k2.7-codeazure_ai/kimi-k2.7-code | 262.144K | $0.95 | $4 | — | |||
| gemini-3.8-flashvertex_ai/gemini-3.8-flash | 1.04858M | $0.75 | $3.75 | — | |||
| claude-opus-5azure_ai/claude-opus-5 | 1M | $5 | $25 | — | |||
| qwen-plus-2025-07-14dashscope/qwen-plus-2025-07-14 | 129.024K | $0.4 | $1.2 | — | |||
| qwen-plus-2025-04-28dashscope/qwen-plus-2025-04-28 | 129.024K | $0.4 | $1.2 | — | |||
| zai-org/GLM-5.2together_ai/zai-org/glm-5.2 | 1.04858M | $1.4 | $4.4 | — | |||
| qwen-plus-2025-01-25dashscope/qwen-plus-2025-01-25 | 129.024K | $0.4 | $1.2 | — | |||
| qwen-plusdashscope/qwen-plus | 129.024K | $0.4 | $1.2 | — |