1,829 models

No provider description is available for this model yet.

azure/gpt-5.1-2025-11-13 272K context $1.25/M input $10/M output

No provider description is available for this model yet.

azure/gpt-5 272K context $1.25/M input $10/M output

No provider description is available for this model yet.

azure/gpt-5-2025-08-07 272K context $1.25/M input $10/M output

No provider description is available for this model yet.

together_ai/google/gemma-4-31b-it 262.144K context $0.39/M input $0.97/M output

No provider description is available for this model yet.

azure/gpt-5-mini 272K context $0.25/M input $2/M output

No provider description is available for this model yet.

azure/gpt-5-mini-2025-08-07 272K context $0.25/M input $2/M output

No provider description is available for this model yet.

azure/gpt-5-nano 272K context $0.05/M input $0.4/M output

No provider description is available for this model yet.

azure/gpt-5-nano-2025-08-07 272K context $0.05/M input $0.4/M output

No provider description is available for this model yet.

azure/gpt-5.1 272K context $1.25/M input $10/M output

No provider description is available for this model yet.

together_ai/moonshotai/kimi-k3 1.04858M context $3/M input $15/M output

No provider description is available for this model yet.

azure/gpt-5.2 272K context $1.75/M input $14/M output

No provider description is available for this model yet.

azure/gpt-5.2-2025-12-11 272K context $1.75/M input $14/M output

GLM-5.1 delivers a major leap in coding capability, with particularly significant gains in handling long-horizon tasks. Unlike previous models built around minute-level interactions, GLM-5.1 can work independently and continuously on...

z-ai/glm-5.1 200K context $0.966/M input $3.036/M output

Jamba Large 1.7 is the latest model in the Jamba open family, offering improvements in grounding, instruction-following, and overall efficiency. Built on a hybrid SSM-Transformer architecture with a 256K context...

ai21/jamba-large-1.7 256K context $2/M input $8/M output

No provider description is available for this model yet.

azure/gpt-5.4 1.05M context $2.5/M input $15/M output

No provider description is available for this model yet.

together_ai/pearl-ai/gemma-4-31b-it 262.144K context $0.28/M input $0.86/M output

No provider description is available for this model yet.

azure/us/gpt-5.4 1.05M context $2.75/M input $16.5/M output

No provider description is available for this model yet.

azure/eu/gpt-5.4 1.05M context $2.75/M input $16.5/M output

Mistral Medium 3.5 is a dense 128B instruction-following model from Mistral AI. It supports text and image inputs with text output, and is designed for agentic workflows, coding, and complex...

mistralai/mistral-medium-3-5:batch 262.144K context $0.75/M input $3.75/M output

GPT-5 Mini is a compact version of GPT-5, designed to handle lighter-weight reasoning tasks. It provides the same instruction-following and safety-tuning benefits as GPT-5, but with reduced latency and cost....

openai/gpt-5-mini:batch 400K context $0.125/M input $1/M output

No provider description is available for this model yet.

together_ai/zai-org/glm-5.2 1.04858M context $1.4/M input $4.4/M output

Qwen3.8 Max (0803) is the August 3, 2026 checkpoint of Qwen3.8 Max, the flagship model in Alibaba's Qwen3.8 series and the general-availability successor to the Qwen3.8 Max Preview. It is...

qwen/qwen3.8-max 1M context $2/M input $6/M output

No provider description is available for this model yet.

azure/gpt-5.4-2026-03-05 1.05M context $2.5/M input $15/M output

No provider description is available for this model yet.

azure/us/gpt-5.4-2026-03-05 1.05M context $2.75/M input $16.5/M output

No provider description is available for this model yet.

azure/eu/gpt-5.4-2026-03-05 1.05M context $2.75/M input $16.5/M output

No provider description is available for this model yet.

azure/gpt-5.6 1.05M context $5/M input $30/M output

No provider description is available for this model yet.

azure/gpt-5.6-sol 1.05M context $5/M input $30/M output

Claude Haiku 4.5 is Anthropic’s fastest and most efficient model, delivering near-frontier intelligence at a fraction of the cost and latency of larger Claude models. Matching Claude Sonnet 4’s performance...

anthropic/claude-haiku-4.5:batch 200K context $0.5/M input $2.5/M output

GPT-5.2 Pro is OpenAI’s most advanced model, offering major improvements in agentic coding and long context performance over GPT-5 Pro. It is optimized for complex tasks that require step-by-step reasoning,...

openai/gpt-5.2-pro:batch 400K context $10.5/M input $84/M output

Claude Opus 4.1 is an updated version of Anthropic’s flagship model, offering improved performance in coding, reasoning, and agentic tasks. It achieves 74.5% on SWE-bench Verified and shows notable gains...

anthropic/claude-opus-4.1:batch 200K context $7.5/M input $37.5/M output

No provider description is available for this model yet.

azure/gpt-5.6-terra 1.05M context $2.5/M input $15/M output

No provider description is available for this model yet.

azure/gpt-5.6-luna 1.05M context $1/M input $6/M output

Gemini 3.5 Flash Lite is a high-efficiency model from Google with upgraded agentic capabilities. It is suited for subagents that execute focused tasks within complex, multi-agent workflows.

google/gemini-3.5-flash-lite:batch 1.04858M context $0.15/M input $1.25/M output

No provider description is available for this model yet.

azure/us/gpt-5.6 1.05M context $5.5/M input $33/M output

No provider description is available for this model yet.

azure/us/gpt-5.6-sol 1.05M context $5.5/M input $33/M output

No provider description is available for this model yet.

azure/us/gpt-5.6-terra 1.05M context $2.75/M input $16.5/M output

No provider description is available for this model yet.

azure/us/gpt-5.6-luna 1.05M context $1.1/M input $6.6/M output

No provider description is available for this model yet.

azure/eu/gpt-5.6 1.05M context $5.5/M input $33/M output

OpenAI o4-mini-high is the same model as [o4-mini](/openai/o4-mini) with reasoning_effort set to high. OpenAI o4-mini is a compact reasoning model in the o-series, optimized for fast, cost-efficient performance while retaining...

openai/o4-mini-high 200K context $1.1/M input $4.4/M output

o3 is a well-rounded and powerful model across domains. It sets a new standard for math, science, coding, and visual reasoning tasks. It also excels at technical writing and instruction-following....

openai/o3:batch 200K context $1/M input $4/M output