No provider description is available for this model yet.
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
GLM-4.5-Air is the lightweight variant of our latest flagship model family, also purpose-built for agent-centric applications. Like GLM-4.5, it adopts the Mixture-of-Experts (MoE) architecture but with a more compact parameter...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| jamba-1.5-miniai21/jamba-1.5-mini | 256K | $0.2 | $0.4 | — | |||
| jamba-1.5-mini@001ai21/jamba-1.5-mini@001 | 256K | $0.2 | $0.4 | — | |||
| jamba-large-1.6ai21/jamba-large-1.6 | 256K | $2 | $8 | — | |||
| jamba-mini-1.6ai21/jamba-mini-1.6 | 256K | $0.2 | $0.4 | — | |||
| Qwen/Qwen3-Coder-480B-A35B-Instruct-Turbodeepinfra/qwen/qwen3-coder-480b-a35b-instruct-turbo | 262.144K | $0.29 | $1.2 | — | |||
| jp.anthropic.claude-sonnet-4-5-20250929-v1:0bedrock_converse/jp.anthropic.claude-sonnet-4-5-20250929-v1:0 | 200K | $3.3 | $16.5 | — | |||
| jp.anthropic.claude-haiku-4-5-20251001-v1:0bedrock_converse/jp.anthropic.claude-haiku-4-5-20251001-v1:0 | 200K | $1.1 | $5.5 | — | |||
| deepseek-ai/DeepSeek-R1-0528crusoe/deepseek-ai/deepseek-r1-0528 | 163.84K | $3 | $7 | — | |||
| deepseek-ai/DeepSeek-V3-0324crusoe/deepseek-ai/deepseek-v3-0324 | 163.84K | $1.5 | $1.5 | — | |||
| google/gemma-3-12b-itcrusoe/google/gemma-3-12b-it | 131.072K | $0.1 | $0.1 | — | |||
| meta-llama/Llama-3.3-70B-Instructcrusoe/meta-llama/llama-3.3-70b-instruct | 131.072K | $0.2 | $0.2 | — | |||
| moonshotai/Kimi-K2-Thinkingcrusoe/moonshotai/kimi-k2-thinking | 262.144K | $2.5 | $2.5 | — | |||
| Qwen/Qwen3-Coder-480B-A35B-Instructdeepinfra/qwen/qwen3-coder-480b-a35b-instruct | 262.144K | $0.4 | $1.6 | — | |||
| Qwen/Qwen3-235B-A22B-Instruct-2507crusoe/qwen/qwen3-235b-a22b-instruct-2507 | 262.144K | $3 | $3 | — | |||
| deepseek-llama3.3-70blambda_ai/deepseek-llama3.3-70b | 131.072K | $0.2 | $0.6 | — | |||
| deepseek-r1-0528lambda_ai/deepseek-r1-0528 | 131.072K | $0.2 | $0.6 | — | |||
| deepseek-r1-671blambda_ai/deepseek-r1-671b | 131.072K | $0.8 | $0.8 | — | |||
| deepseek-v3-0324lambda_ai/deepseek-v3-0324 | 131.072K | $0.2 | $0.6 | — | |||
| hermes3-405blambda_ai/hermes3-405b | 131.072K | $0.8 | $0.8 | — | |||
| hermes3-70blambda_ai/hermes3-70b | 131.072K | $0.12 | $0.3 | — | |||
| gemini-3.7-flashvertex_ai/gemini-3.7-flash | 1.04858M | $0.75 | $3.75 | — | |||
| lfm-40blambda_ai/lfm-40b | 131.072K | $0.1 | $0.2 | — | |||
| lfm-7blambda_ai/lfm-7b | 131.072K | $0.025 | $0.04 | — | |||
| llama-4-maverick-17b-128e-instruct-fp8lambda_ai/llama-4-maverick-17b-128e-instruct-fp8 | 131.072K | $0.05 | $0.1 | — | |||
| Z.ai: GLM 4.5 Airz-ai/glm-4.5-air | 131.072K | $0.13 | $0.85 | — | |||
| llama3.1-70b-instruct-fp8lambda_ai/llama3.1-70b-instruct-fp8 | 131.072K | $0.12 | $0.3 | — | |||
| llama3.1-8b-instructlambda_ai/llama3.1-8b-instruct | 131.072K | $0.025 | $0.04 | — | |||
| llama3.1-nemotron-70b-instruct-fp8lambda_ai/llama3.1-nemotron-70b-instruct-fp8 | 131.072K | $0.12 | $0.3 | — | |||
| Qwen/Qwen3-32Bdeepinfra/qwen/qwen3-32b | 40.96K | $0.1 | $0.28 | — | |||
| llama3.2-3b-instructlambda_ai/llama3.2-3b-instruct | 131.072K | $0.015 | $0.025 | — | |||
| llama3.3-70b-instruct-fp8lambda_ai/llama3.3-70b-instruct-fp8 | 131.072K | $0.12 | $0.3 | — | |||
| qwen25-coder-32b-instructlambda_ai/qwen25-coder-32b-instruct | 131.072K | $0.05 | $0.1 | — | |||
| Qwen/Qwen3-30B-A3Bdeepinfra/qwen/qwen3-30b-a3b | 40.96K | $0.08 | $0.29 | — | |||
| medlm-mediumvertex_ai-language-models/medlm-medium | 32.768K | — | — | — | |||
| meta.llama3-1-405b-instruct-v1:0bedrock/meta.llama3-1-405b-instruct-v1:0 | 128K | $5.32 | $16 | — | |||
| meta.llama3-1-70b-instruct-v1:0bedrock/meta.llama3-1-70b-instruct-v1:0 | 128K | $0.99 | $0.99 | — | |||
| nvidia/NVIDIA-Nemotron-3.5-Lightningdeepinfra/nvidia/nvidia-nemotron-3.5-lightning | 262.144K | $0.05 | $0.2 | — | |||
| meta.llama3-2-11b-instruct-v1:0bedrock/meta.llama3-2-11b-instruct-v1:0 | 128K | $0.35 | $0.35 | — | |||
| meta.llama3-2-1b-instruct-v1:0bedrock/meta.llama3-2-1b-instruct-v1:0 | 128K | $0.1 | $0.1 | — | |||
| meta.llama3-2-3b-instruct-v1:0bedrock/meta.llama3-2-3b-instruct-v1:0 | 128K | $0.15 | $0.15 | — | |||
| meta.llama3-2-90b-instruct-v1:0bedrock/meta.llama3-2-90b-instruct-v1:0 | 128K | $2 | $2 | — | |||
| meta.llama3-3-70b-instruct-v1:0bedrock_converse/meta.llama3-3-70b-instruct-v1:0 | 128K | $0.72 | $0.72 | — | |||
| meta.llama4-maverick-17b-instruct-v1:0bedrock_converse/meta.llama4-maverick-17b-instruct-v1:0 | 128K | $0.24 | $0.97 | — | |||
| meta.llama4-scout-17b-instruct-v1:0bedrock_converse/meta.llama4-scout-17b-instruct-v1:0 | 128K | $0.17 | $0.66 | — | |||
| Qwen/Qwen3-235B-A22B-Thinking-2507deepinfra/qwen/qwen3-235b-a22b-thinking-2507 | 262.144K | $0.3 | $2.9 | — | |||
| Llama-3.3-8B-Instructmeta_llama/llama-3.3-8b-instruct | 128K | — | — | — | |||
| Qwen/Qwen3-235B-A22B-Instruct-2507deepinfra/qwen/qwen3-235b-a22b-instruct-2507 | 262.144K | $0.09 | $0.6 | — | |||
| Llama-4-Scout-17B-16E-Instruct-FP8meta_llama/llama-4-scout-17b-16e-instruct-fp8 | 10M | — | — | — | |||
| minimax.minimax-m2bedrock_converse/minimax.minimax-m2 | 128K | $0.3 | $1.2 | — | |||
| qwen3.8-maxdashscope/qwen3.8-max | 991.808K | $2 | $6 | — |