3,455 models

No provider description is available for this model yet.

bedrock/sa-east-1/deepseek.v3.2 163.84K context $0.74/M input $2.22/M output

Claude Fable 5.1 improves on Claude Fable 5 across the board, with the biggest gains in agentic coding, long-running agentic workflows, and knowledge work: long code refactors, front-end and visual...

anthropic/claude-fable-5.1:batch 1M context $5/M input $25/M output

Qwen3-Max is an updated release built on the Qwen3 series, offering major improvements in reasoning, instruction following, multilingual support, and long-tail knowledge coverage compared to the January 2025 version. It...

qwen/qwen3-max 262.144K context $0.78/M input $3.9/M output

No provider description is available for this model yet.

gmi/zai-org/glm-4.7-fp8 202.752K context $0.4/M input $2/M output

Mistral Medium 3 is a high-performance enterprise-grade language model designed to deliver frontier-level capabilities at significantly reduced operational cost. It balances state-of-the-art reasoning and multimodal performance with 8× lower cost...

mistralai/mistral-medium-3 131.072K context $0.4/M input $2/M output

No provider description is available for this model yet.

scx-ai/glm-5.2 1.04858M context $0.61/M input $1.98/M output

No provider description is available for this model yet.

scx-ai/qwen3.8-max 1M context $1.65/M input $4.99/M output

Inkling is an open-weight multimodal mixture-of-experts model from Thinking Machines Lab, with 41B active parameters out of 975B total. It is designed for general-purpose reasoning, coding, agentic and tool-use systems,...

thinkingmachines/inkling:batch 524.288K context $1/M input $4.05/M output

GPT-5.4 nano is the most lightweight and cost-efficient variant of the GPT-5.4 family, optimized for speed-critical and high-volume tasks. It supports text and image inputs and is designed for low-latency...

openai/gpt-5.4-nano:batch 400K context $0.1/M input $0.625/M output

o3 is a well-rounded and powerful model across domains. It sets a new standard for math, science, coding, and visual reasoning tasks. It also excels at technical writing and instruction-following....

openai/o3:batch 200K context $1/M input $4/M output

Qwen3-Coder-480B-A35B-Instruct is a Mixture-of-Experts (MoE) code generation model developed by the Qwen team. It is optimized for agentic coding tasks such as function calling, tool use, and long-context reasoning over...

qwen/qwen3-coder:free 262K context Free input Free output

No provider description is available for this model yet.

dashscope/deepseek-v4-flash 1M context $0.2/M input $0.4/M output

No provider description is available for this model yet.

dashscope/deepseek-v4-flash-0731 1M context $0.2/M input $0.4/M output

No provider description is available for this model yet.

dashscope/deepseek-v4-pro 1M context $2.4/M input $4.8/M output

No provider description is available for this model yet.

dashscope/glm-5.1 202.745K context $1.4/M input $4.4/M output

No provider description is available for this model yet.

dashscope/glm-5.2 1.04858M context $1.4/M input $4.4/M output

No provider description is available for this model yet.

dashscope/kimi-k2.7-code 229.376K context $0.95/M input $4/M output

No provider description is available for this model yet.

replicate/deepseek-ai/deepseek-r1 65.536K context $3.75/M input $10/M output

No provider description is available for this model yet.

dashscope/qwen3.8-max 991.808K context $2/M input $6/M output