3,229 models

No provider description is available for this model yet.

ai21/jamba-1.5-mini 256K context $0.2/M input $0.4/M output

No provider description is available for this model yet.

ai21/jamba-1.5-mini@001 256K context $0.2/M input $0.4/M output

No provider description is available for this model yet.

ai21/jamba-large-1.6 256K context $2/M input $8/M output

No provider description is available for this model yet.

ai21/jamba-mini-1.6 256K context $0.2/M input $0.4/M output

No provider description is available for this model yet.

crusoe/google/gemma-3-12b-it 131.072K context $0.1/M input $0.1/M output

No provider description is available for this model yet.

crusoe/moonshotai/kimi-k2-thinking 262.144K context $2.5/M input $2.5/M output

No provider description is available for this model yet.

lambda_ai/deepseek-llama3.3-70b 131.072K context $0.2/M input $0.6/M output

No provider description is available for this model yet.

lambda_ai/deepseek-r1-0528 131.072K context $0.2/M input $0.6/M output

No provider description is available for this model yet.

lambda_ai/deepseek-r1-671b 131.072K context $0.8/M input $0.8/M output

No provider description is available for this model yet.

lambda_ai/deepseek-v3-0324 131.072K context $0.2/M input $0.6/M output

No provider description is available for this model yet.

lambda_ai/hermes3-405b 131.072K context $0.8/M input $0.8/M output

No provider description is available for this model yet.

lambda_ai/hermes3-70b 131.072K context $0.12/M input $0.3/M output

No provider description is available for this model yet.

vertex_ai/gemini-3.7-flash 1.04858M context $0.75/M input $3.75/M output

No provider description is available for this model yet.

lambda_ai/lfm-40b 131.072K context $0.1/M input $0.2/M output

No provider description is available for this model yet.

lambda_ai/lfm-7b 131.072K context $0.025/M input $0.04/M output

GLM-4.5-Air is the lightweight variant of our latest flagship model family, also purpose-built for agent-centric applications. Like GLM-4.5, it adopts the Mixture-of-Experts (MoE) architecture but with a more compact parameter...

z-ai/glm-4.5-air 131.072K context $0.13/M input $0.85/M output

No provider description is available for this model yet.

lambda_ai/llama3.1-70b-instruct-fp8 131.072K context $0.12/M input $0.3/M output

No provider description is available for this model yet.

lambda_ai/llama3.1-8b-instruct 131.072K context $0.025/M input $0.04/M output

No provider description is available for this model yet.

deepinfra/qwen/qwen3-32b 40.96K context $0.1/M input $0.28/M output

No provider description is available for this model yet.

lambda_ai/llama3.2-3b-instruct 131.072K context $0.015/M input $0.025/M output

No provider description is available for this model yet.

lambda_ai/llama3.3-70b-instruct-fp8 131.072K context $0.12/M input $0.3/M output

No provider description is available for this model yet.

lambda_ai/qwen25-coder-32b-instruct 131.072K context $0.05/M input $0.1/M output

No provider description is available for this model yet.

deepinfra/qwen/qwen3-30b-a3b 40.96K context $0.08/M input $0.29/M output

No provider description is available for this model yet.

meta_llama/llama-3.3-8b-instruct 128K context Input not listed Output not listed

No provider description is available for this model yet.

dashscope/qwen3.8-max 991.808K context $2/M input $6/M output