3,229 models

No provider description is available for this model yet.

novita/thudm/glm-4-32b-0414 32K context $0.55/M input $1.66/M output

No provider description is available for this model yet.

wandb/deepseek-ai/deepseek-v4-flash 1.04858M context $0.14/M input $0.28/M output

No provider description is available for this model yet.

wandb/deepseek-ai/deepseek-v4-pro 1.04858M context $1.15/M input $2.55/M output

No provider description is available for this model yet.

wandb/google/gemma-4-31b-it 262.144K context $0.1/M input $0.34/M output

No provider description is available for this model yet.

wandb/ibm-granite/granite-4.1-8b 131.072K context $0.05/M input $0.1/M output

No provider description is available for this model yet.

wandb/minimaxai/minimax-m3 262.144K context $0.23/M input $0.96/M output

No provider description is available for this model yet.

wandb/moonshotai/kimi-k2.7-code 262.144K context $0.71/M input $3.5/M output

No provider description is available for this model yet.

wandb/moonshotai/kimi-k2.6 262.144K context $0.65/M input $3.41/M output

GPT-5.4 mini brings the core capabilities of GPT-5.4 to a faster, more efficient model optimized for high-throughput workloads. It supports text and image inputs with strong performance across reasoning, coding,...

openai/gpt-5.4-mini:batch 400K context $0.375/M input $2.25/M output

No provider description is available for this model yet.

wandb/openpipe/qwen3-14b-instruct 32.768K context $0.05/M input $0.22/M output

No provider description is available for this model yet.

wandb/qwen/qwen3.8-27b 262.144K context $0.4/M input $3/M output

No provider description is available for this model yet.

wandb/qwen/qwen3.6-35b-a3b 262.144K context $0.25/M input $1.25/M output

No provider description is available for this model yet.

wandb/qwen/qwen3.6-27b 262.144K context $0.6/M input $3.6/M output

Ling 3.0 Flash Sante is a health and medicine-focused mixture-of-experts model from InclusionAI, built on Ling 3.0 Flash with 5.1B active parameters out of 124B total. It is designed for...

inclusionai/ling-3.0-flash-sante:free 262.144K context Free input Free output

No provider description is available for this model yet.

wandb/qwen/qwen3.5-35b-a3b 262.144K context $0.25/M input $1.25/M output

No provider description is available for this model yet.

wandb/zai-org/glm-5.2 262.144K context $0.76/M input $2.42/M output

No provider description is available for this model yet.

deepinfra/openai/gpt-oss-120b-turbo 131.072K context $0.15/M input $0.6/M output

No provider description is available for this model yet.

deepinfra/minimaxai/minimax-m2.7 196.608K context $0.25/M input $1/M output

No provider description is available for this model yet.

deepinfra/qwen/qwen3.8-27b 262.144K context $0.4/M input $3/M output

GPT-5-Codex is a specialized version of GPT-5 optimized for software engineering and coding workflows. It is designed for both interactive development sessions and long, independent execution of complex engineering tasks....

openai/gpt-5-codex:batch 400K context $0.625/M input $5/M output

No provider description is available for this model yet.

deepinfra/moonshotai/kimi-k2.5 262.144K context $0.45/M input $2.25/M output

No provider description is available for this model yet.

deepinfra/zai-org/glm-4.7-flash 202.752K context $0.06/M input $0.4/M output

No provider description is available for this model yet.

deepinfra/zai-org/glm-4.6 202.752K context $0.5/M input $2/M output

GPT-5.1 Chat (AKA Instant is the fast, lightweight member of the 5.1 family, optimized for low-latency chat while retaining strong general intelligence. It uses adaptive reasoning to selectively “think” on...

openai/gpt-5.1-chat 128K context $1.25/M input $10/M output

No provider description is available for this model yet.

deepinfra/xiaomimimo/mimo-v2.5 262.144K context $0.4/M input $2/M output

No provider description is available for this model yet.

deepinfra/qwen/qwen3-max 256K context $1.2/M input $6/M output

No provider description is available for this model yet.

databricks/databricks-glm-5-3-flash 1.04858M context Input not listed Output not listed

No provider description is available for this model yet.

gemini/gemma-4-26b-a4b-it 262.144K context Input not listed Output not listed

No provider description is available for this model yet.

gemini/gemma-4-31b-it 262.144K context Input not listed Output not listed

No provider description is available for this model yet.

moonshot/kimi-k2.7-code 262.144K context $0.95/M input $4/M output

No provider description is available for this model yet.

together_ai/zai-org/glm-5.3 1.04858M context $1.4/M input $4.4/M output

No provider description is available for this model yet.

zai/glm-5.3-flash 1.04858M context $0.15/M input $0.5/M output

No provider description is available for this model yet.

gemini/gemini-omni-1.1-flash 131.072K context $1.5/M input $9/M output

No provider description is available for this model yet.

xai/grok-4.20 1M context $1.25/M input $2.5/M output