4,182 models

No provider description is available for this model yet.

databricks/databricks-grok-4-6 500K context $2.5/M input $7.5/M output

No provider description is available for this model yet.

databricks/databricks-inkling 1M context $1/M input $4.05/M output

No provider description is available for this model yet.

nebius/deepseek-ai/deepseek-v4-pro 1.04858M context $1.75/M input $3.5/M output

No provider description is available for this model yet.

nebius/minimaxai/minimax-m2.5 196.608K context $0.3/M input $1.2/M output

No provider description is available for this model yet.

nebius/minimaxai/minimax-m3 1.04858M context $0.3/M input $1.2/M output

No provider description is available for this model yet.

nebius/moonshotai/kimi-k2.6 262.144K context $0.95/M input $4/M output

No provider description is available for this model yet.

nebius/moonshotai/kimi-k2.7-code 262.144K context $0.95/M input $4/M output

No provider description is available for this model yet.

nebius/moonshotai/kimi-k3 1.024M context $3/M input $15/M output

No provider description is available for this model yet.

nebius/nousresearch/hermes-4-405b 131.072K context $1/M input $3/M output

No provider description is available for this model yet.

nebius/nousresearch/hermes-4-70b 131.072K context $0.13/M input $0.4/M output

No provider description is available for this model yet.

deepinfra/qwen/qwen3-max 256K context $1.2/M input $6/M output

No provider description is available for this model yet.

deepinfra/xiaomimimo/mimo-v2.5 262.144K context $0.4/M input $2/M output

Qwen3-235B-A22B is a 235B parameter mixture-of-experts (MoE) model developed by Qwen, activating 22B parameters per forward pass. It supports seamless switching between a "thinking" mode for complex reasoning, math, and...

qwen/qwen3-235b-a22b 131.072K context $0.455/M input $1.82/M output

GPT-5.1 Chat (AKA Instant is the fast, lightweight member of the 5.1 family, optimized for low-latency chat while retaining strong general intelligence. It uses adaptive reasoning to selectively “think” on...

openai/gpt-5.1-chat 128K context $1.25/M input $10/M output

No provider description is available for this model yet.

novita/deepseek/deepseek-r1-0528 163.84K context $0.7/M input $2.5/M output

No provider description is available for this model yet.

novita/minimaxai/minimax-m1-80k 1M context $0.55/M input $2.2/M output

No provider description is available for this model yet.

xai/grok-vision-beta 8.192K context $5/M input $15/M output

No provider description is available for this model yet.

xai/grok-4 256K context $3/M input $15/M output

No provider description is available for this model yet.

xai/grok-3-beta 131.072K context $3/M input $15/M output

No provider description is available for this model yet.

xai/grok-2-latest 131.072K context $2/M input $10/M output

No provider description is available for this model yet.

deepinfra/zai-org/glm-4.6 202.752K context $0.5/M input $2/M output

No provider description is available for this model yet.

novita/mistralai/mistral-nemo 60.288K context $0.04/M input $0.17/M output

No provider description is available for this model yet.

deepinfra/zai-org/glm-4.7-flash 202.752K context $0.06/M input $0.4/M output

No provider description is available for this model yet.

deepinfra/moonshotai/kimi-k2.5 262.144K context $0.45/M input $2.25/M output

No provider description is available for this model yet.

novita/qwen/qwen-2.5-72b-instruct 32K context $0.38/M input $0.4/M output

GPT-5-Codex is a specialized version of GPT-5 optimized for software engineering and coding workflows. It is designed for both interactive development sessions and long, independent execution of complex engineering tasks....

openai/gpt-5-codex:batch 400K context $0.625/M input $5/M output

No provider description is available for this model yet.

deepinfra/qwen/qwen3.8-27b 262.144K context $0.4/M input $3/M output

No provider description is available for this model yet.

deepinfra/minimaxai/minimax-m2.7 196.608K context $0.25/M input $1/M output