1,424 models

No provider description is available for this model yet.

azure/gpt-4o-2024-08-06 128K context $2.5/M input $10/M output

No provider description is available for this model yet.

azure/gpt-4o-2024-11-20 128K context $2.75/M input $11/M output

No provider description is available for this model yet.

together_ai/qwen/qwen3.5-9b 262.144K context $0.17/M input $0.25/M output

Dots3-Note Preview is an open-weight mixture-of-experts model from Dots Studio, with 16B active parameters out of 280B total. It is the lightest model in the Dots 3 family and is...

dots-studio/dots-3-note-preview:free 512K context Free input Free output

No provider description is available for this model yet.

azure/gpt-4o-mini 128K context $0.165/M input $0.66/M output

No provider description is available for this model yet.

azure/gpt-4o-mini-2024-07-18 128K context $0.165/M input $0.66/M output

No provider description is available for this model yet.

azure/gpt-5.1-2025-11-13 272K context $1.25/M input $10/M output

No provider description is available for this model yet.

azure/gpt-5.1-chat-2025-11-13 128K context $1.25/M input $10/M output

No provider description is available for this model yet.

azure/gpt-5 272K context $1.25/M input $10/M output

No provider description is available for this model yet.

azure/gpt-5-2025-08-07 272K context $1.25/M input $10/M output

No provider description is available for this model yet.

azure/gpt-5-chat 128K context $1.25/M input $10/M output

No provider description is available for this model yet.

azure/gpt-5-chat-latest 128K context $1.25/M input $10/M output

No provider description is available for this model yet.

together_ai/google/gemma-4-31b-it 262.144K context $0.39/M input $0.97/M output

Qwen3.8 Max 0902 is an updated snapshot of Qwen3.8 Max from Alibaba's Qwen team. It is a 2.4-trillion-parameter mixture-of-experts model that accepts text, image, and video input and returns text,...

qwen/qwen3.8-max-0902 1M context $2/M input $6/M output

No provider description is available for this model yet.

azure/gpt-5-mini 272K context $0.25/M input $2/M output

No provider description is available for this model yet.

azure/gpt-5-mini-2025-08-07 272K context $0.25/M input $2/M output

No provider description is available for this model yet.

azure/gpt-5-nano 272K context $0.05/M input $0.4/M output

Qwen3.8 27B is an open-weight dense vision-language model from Qwen. It is suited for coding, professional workflows, research, multimodal interaction, and long-running agent tasks, with flexible thinking that can be...

qwen/qwen3.8-27b 262.144K context $0.214/M input $2.55/M output

No provider description is available for this model yet.

azure/gpt-5-nano-2025-08-07 272K context $0.05/M input $0.4/M output

No provider description is available for this model yet.

azure/gpt-5.1 272K context $1.25/M input $10/M output

No provider description is available for this model yet.

together_ai/moonshotai/kimi-k3 1.04858M context $3/M input $15/M output

No provider description is available for this model yet.

azure/gpt-5.1-chat 128K context $1.25/M input $10/M output

No provider description is available for this model yet.

azure/gpt-5.2 272K context $1.75/M input $14/M output

No provider description is available for this model yet.

azure/gpt-5.2-2025-12-11 272K context $1.75/M input $14/M output

No provider description is available for this model yet.

azure/gpt-5.2-chat 128K context $1.75/M input $14/M output

Inkling Small is an open-weight multimodal mixture-of-experts model from Thinking Machines Lab, with 12B active parameters out of 276B total. It is positioned as the smaller, more efficient member of...

thinkingmachines/inkling-small:batch 524.288K context $0.5/M input $1.2/M output

No provider description is available for this model yet.

azure/gpt-5.2-chat-2025-12-11 128K context $1.75/M input $14/M output

No provider description is available for this model yet.

azure/gpt-5.3-chat 128K context $1.75/M input $14/M output

No provider description is available for this model yet.

azure/gpt-5.4 1.05M context $2.5/M input $15/M output

Mistral Medium 3.5 is a dense 128B instruction-following model from Mistral AI. It supports text and image inputs with text output, and is designed for agentic workflows, coding, and complex...

mistralai/mistral-medium-3-5:batch 262.144K context $0.75/M input $3.75/M output

No provider description is available for this model yet.

azure/us/gpt-5.4 1.05M context $2.75/M input $16.5/M output

No provider description is available for this model yet.

azure/eu/gpt-5.4 1.05M context $2.75/M input $16.5/M output

Gemini 3 Flash Preview is a high speed, high value thinking model designed for agentic workflows, multi turn chat, and coding assistance. It delivers near Pro level reasoning and tool...

google/gemini-3-flash-preview:batch 1.04858M context $0.25/M input $1.5/M output

Qwen3.8 Max (0803) is the August 3, 2026 checkpoint of Qwen3.8 Max, the flagship model in Alibaba's Qwen3.8 series and the general-availability successor to the Qwen3.8 Max Preview. It is...

qwen/qwen3.8-max 1M context $2/M input $6/M output

No provider description is available for this model yet.

azure/gpt-5.4-2026-03-05 1.05M context $2.5/M input $15/M output

No provider description is available for this model yet.

azure/us/gpt-5.4-2026-03-05 1.05M context $2.75/M input $16.5/M output

No provider description is available for this model yet.

azure/eu/gpt-5.4-2026-03-05 1.05M context $2.75/M input $16.5/M output

No provider description is available for this model yet.

azure/gpt-5.6 1.05M context $5/M input $30/M output

No provider description is available for this model yet.

azure/gpt-5.6-sol 1.05M context $5/M input $30/M output

GPT-5 Pro is OpenAI’s most advanced model, offering major improvements in reasoning, code quality, and user experience. It is optimized for complex tasks that require step-by-step reasoning, instruction following, and...

openai/gpt-5-pro:batch 400K context $7.5/M input $60/M output

No provider description is available for this model yet.

azure/gpt-5.6-terra 1.05M context $2.5/M input $15/M output

No provider description is available for this model yet.

azure/gpt-5.6-luna 1.05M context $1/M input $6/M output

Gemini 3.5 Flash Lite is a high-efficiency model from Google with upgraded agentic capabilities. It is suited for subagents that execute focused tasks within complex, multi-agent workflows.

google/gemini-3.5-flash-lite:batch 1.04858M context $0.15/M input $1.25/M output

No provider description is available for this model yet.

azure/us/gpt-5.6 1.05M context $5.5/M input $33/M output

No provider description is available for this model yet.

azure/us/gpt-5.6-sol 1.05M context $5.5/M input $33/M output

No provider description is available for this model yet.

azure/us/gpt-5.6-terra 1.05M context $2.75/M input $16.5/M output

No provider description is available for this model yet.

azure/us/gpt-5.6-luna 1.05M context $1.1/M input $6.6/M output

No provider description is available for this model yet.

azure/eu/gpt-5.6 1.05M context $5.5/M input $33/M output