735 models

No provider description is available for this model yet.

zai/glm-5.3 1M context $1.4/M input $4.4/M output

No provider description is available for this model yet.

tencent/minimax-m3 1M context $0.3/M input $1.2/M output

No provider description is available for this model yet.

novita/zai-org/glm-5.3 1.04858M context $1.4/M input $4.4/M output

No provider description is available for this model yet.

novita/moonshotai/kimi-k3 1.04858M context $3/M input $15/M output

No provider description is available for this model yet.

novita/zai-org/glm-5.2 1.04858M context $1.4/M input $4.4/M output

Gemini 2.5 Flash-Lite is a lightweight reasoning model in the Gemini 2.5 family, optimized for ultra-low latency and cost efficiency. It offers improved throughput, faster token generation, and better performance...

google/gemini-2.5-flash-lite:batch 1.04858M context $0.05/M input $0.2/M output

No provider description is available for this model yet.

novita/mindai/macaron-v1-venti 1.04858M context $1.5/M input $4.5/M output

No provider description is available for this model yet.

novita/minimax/minimax-m3 1M context $0.3/M input $1.2/M output

No provider description is available for this model yet.

novita/deepseek/deepseek-v4-flash 1.04858M context $0.14/M input $0.28/M output

No provider description is available for this model yet.

novita/deepseek/deepseek-v4-pro 1.04858M context $1.6/M input $3.2/M output

No provider description is available for this model yet.

novita/qwen/qwen3.8-max 1M context $2/M input $6/M output

No provider description is available for this model yet.

novita/xiaomimimo/mimo-v2.5 1.04858M context $0.168/M input $0.336/M output

No provider description is available for this model yet.

novita/qwen/qwen3.7-max 1M context $1.25/M input $3.75/M output

No provider description is available for this model yet.

novita/xiaomimimo/mimo-v2.5-pro 1.04858M context $0.522/M input $1.044/M output

GPT-4.1 Mini is a mid-sized model delivering performance competitive with GPT-4o at substantially lower latency and cost. It retains a 1 million token context window and scores 45.1% on hard...

openai/gpt-4.1-mini:batch 1.04758M context $0.2/M input $0.8/M output

No provider description is available for this model yet.

databricks/databricks-gemini-3-pro 1.04858M context $2.5/M input $15/M output

No provider description is available for this model yet.

wandb/deepseek-ai/deepseek-v4-flash 1.04858M context $0.14/M input $0.28/M output

No provider description is available for this model yet.

wandb/deepseek-ai/deepseek-v4-pro 1.04858M context $1.15/M input $2.55/M output

No provider description is available for this model yet.

databricks/databricks-glm-5-3-flash 1.04858M context Input not listed Output not listed

No provider description is available for this model yet.

together_ai/zai-org/glm-5.3 1.04858M context $1.4/M input $4.4/M output

No provider description is available for this model yet.

zai/glm-5.3-flash 1.04858M context $0.15/M input $0.5/M output

No provider description is available for this model yet.

xai/grok-4.20 1M context $1.25/M input $2.5/M output

No provider description is available for this model yet.

xai/grok-4.20-reasoning 1M context $1.25/M input $2.5/M output

No provider description is available for this model yet.

xai/grok-4.20-reasoning-latest 1M context $1.25/M input $2.5/M output

No provider description is available for this model yet.

xai/grok-4.20-non-reasoning 1M context $1.25/M input $2.5/M output

GLM 5.2 is a large-scale reasoning model from Z.ai. It supports text input and output with a 1M-token context window, and is suited for long-horizon agent workflows, project-level software engineering,...

z-ai/glm-5.2:batch 1.04858M context $0.7/M input $2.2/M output

Grok 4.3 is a reasoning model from SpaceXAI. It accepts text and image inputs with text output, and is suited for agentic workflows, instruction-following tasks, and applications requiring high factual...

x-ai/grok-4.3 1M context $1.25/M input $2.5/M output

No provider description is available for this model yet.

deepinfra/google/gemini-3.7-flash 1M context $0.75/M input $3.75/M output

No provider description is available for this model yet.

deepinfra/zai-org/glm-5.2 1.04858M context $0.75/M input $2.4/M output

No provider description is available for this model yet.

deepinfra/moonshotai/kimi-k3 1.04858M context $2.85/M input $14.25/M output