4,213 models

No provider description is available for this model yet.

azure/eu/gpt-5.4 1.05M context $2.75/M input $16.5/M output
Open weights

No provider description is available for this model yet.

microsoft/Dayhoff-3b-UR90-10 Not documented context Input not listed Output not listed

Mistral Medium 3.5 is a dense 128B instruction-following model from Mistral AI. It supports text and image inputs with text output, and is designed for agentic workflows, coding, and complex...

mistralai/mistral-medium-3-5:batch 262.144K context $0.75/M input $3.75/M output
Open weights

No provider description is available for this model yet.

microsoft/Mage-Flow Not documented context Input not listed Output not listed
Open weights

No provider description is available for this model yet.

meta-llama/Llama-2-13b-chat-hf Not documented context Input not listed Output not listed

No provider description is available for this model yet.

together_ai/zai-org/glm-5.2 1.04858M context $1.4/M input $4.4/M output

Qwen3.8 Max (0803) is the August 3, 2026 checkpoint of Qwen3.8 Max, the flagship model in Alibaba's Qwen3.8 series and the general-availability successor to the Qwen3.8 Max Preview. It is...

qwen/qwen3.8-max 1M context $2/M input $6/M output

No provider description is available for this model yet.

azure/gpt-5.4-2026-03-05 1.05M context $2.5/M input $15/M output

No provider description is available for this model yet.

azure/us/gpt-5.4-2026-03-05 1.05M context $2.75/M input $16.5/M output

No provider description is available for this model yet.

azure/eu/gpt-5.4-2026-03-05 1.05M context $2.75/M input $16.5/M output

No provider description is available for this model yet.

azure/gpt-5.6 1.05M context $5/M input $30/M output

No provider description is available for this model yet.

azure/gpt-5.6-sol 1.05M context $5/M input $30/M output

Mercury 2 is an extremely fast reasoning LLM, and the first reasoning diffusion LLM (dLLM). Instead of generating tokens sequentially, Mercury 2 produces and refines multiple tokens in parallel, achieving...

inception/mercury-2 128K context $0.25/M input $0.75/M output

No provider description is available for this model yet.

azure/gpt-5.6-terra 1.05M context $2.5/M input $15/M output

No provider description is available for this model yet.

azure/gpt-5.6-luna 1.05M context $1/M input $6/M output

Qwen-Plus, based on the Qwen2.5 foundation model, is a 131K context model with a balanced performance, speed, and cost combination.

qwen/qwen-plus 1M context $0.26/M input $0.78/M output

GPT Chat Latest points to OpenAI's stable API alias `chat-latest` that always resolves to the latest Instant chat model used in ChatGPT. As OpenAI rolls out new Instant model updates...

openai/gpt-chat-latest 400K context $5/M input $30/M output

No provider description is available for this model yet.

ollama/mistral-7b-instruct-v0.1 8.192K context Input not listed Output not listed

No provider description is available for this model yet.

ollama/mistral-7b-instruct-v0.2 32.768K context Input not listed Output not listed

No provider description is available for this model yet.

azure/global/gpt-4o-2024-08-06 128K context $2.5/M input $10/M output

No provider description is available for this model yet.

azure/us/gpt-5.6-sol 1.05M context $5.5/M input $33/M output

No provider description is available for this model yet.

azure/us/gpt-5.6-terra 1.05M context $2.75/M input $16.5/M output

No provider description is available for this model yet.

ollama/mistral-large-instruct-2407 65.536K context Input not listed Output not listed

No provider description is available for this model yet.

ollama/mixtral-8x22b-instruct-v0.1 65.536K context Input not listed Output not listed

No provider description is available for this model yet.

nvidia/MiniMax-M2.7-DFlash Not documented context Input not listed Output not listed

No provider description is available for this model yet.

azure/us/gpt-5.6-luna 1.05M context $1.1/M input $6.6/M output

No provider description is available for this model yet.

azure/eu/gpt-5.6 1.05M context $5.5/M input $33/M output

Claude Opus 4 is benchmarked as the world’s best coding model, at time of release, bringing sustained performance on complex, long-running tasks and agent workflows. It sets new benchmarks in...

anthropic/claude-opus-4 200K context $15/M input $75/M output

No provider description is available for this model yet.

perplexity/mixtral-8x7b-instruct 4.096K context $0.07/M input $0.28/M output

No provider description is available for this model yet.

azure/eu/gpt-5.6-sol 1.05M context $5.5/M input $33/M output

No provider description is available for this model yet.

perplexity/pplx-70b-online 4.096K context Input not listed $2.8/M output

No provider description is available for this model yet.

perplexity/pplx-7b-chat 8.192K context $0.07/M input $0.28/M output

No provider description is available for this model yet.

azure/eu/gpt-5.6-terra 1.05M context $2.75/M input $16.5/M output