4,251 models

No provider description is available for this model yet.

cohere_chat/command-r-plus 128K context $2.5/M input $10/M output
Open weights

No provider description is available for this model yet.

meta-llama/Llama-3.2-1B-Instruct Not documented context Input not listed Output not listed

No provider description is available for this model yet.

replicate/meta/llama-2-70b-chat 4.096K context $0.65/M input $2.75/M output

Devstral 2 is a state-of-the-art open-source model by Mistral AI specializing in agentic coding. It is a 123B-parameter dense transformer model supporting a 256K context window. Devstral 2 supports exploring...

mistralai/devstral-2512 262.144K context $0.4/M input $2/M output
Open weights

No provider description is available for this model yet.

meta-llama/Llama-3.2-3B Not documented context Input not listed Output not listed

Qwen Plus 0728, based on the Qwen3 foundation model, is a 1 million context hybrid reasoning model with a balanced performance, speed, and cost combination.

qwen/qwen-plus-2025-07-28 1M context $0.26/M input $0.78/M output

Mistral Medium 3.1 is an updated version of Mistral Medium 3, which is a high-performance enterprise-grade language model designed to deliver frontier-level capabilities at significantly reduced operational cost. It balances...

mistralai/mistral-medium-3.1 131.072K context $0.4/M input $2/M output

This model always redirects to the latest model in the GPT Luna family.

~openai/gpt-luna-latest 1.05M context $0.2/M input $1.2/M output
Open weights

No provider description is available for this model yet.

microsoft/Fara1.5-4B Not documented context Input not listed Output not listed
Open weights

No provider description is available for this model yet.

meta-llama/Llama-3.1-8B Not documented context Input not listed Output not listed
Open weights

No provider description is available for this model yet.

meta-llama/Llama-Guard-3-8B Not documented context Input not listed Output not listed
Open weights

No provider description is available for this model yet.

meta-llama/Meta-Llama-3-70B Not documented context Input not listed Output not listed
Open weights

No provider description is available for this model yet.

microsoft/MagenticBrain Not documented context Input not listed Output not listed
Open weights

No provider description is available for this model yet.

meta-llama/Meta-Llama-3-8B Not documented context Input not listed Output not listed

GPT-5-Codex is a specialized version of GPT-5 optimized for software engineering and coding workflows. It is designed for both interactive development sessions and long, independent execution of complex engineering tasks....

openai/gpt-5-codex:batch 400K context $0.625/M input $5/M output

No provider description is available for this model yet.

bedrock/ai21.j2-mid-v1 8.191K context $12.5/M input $12.5/M output
Open weights

No provider description is available for this model yet.

meta-llama/Llama-3.2-90B-Vision Not documented context Input not listed Output not listed
Open weights

No provider description is available for this model yet.

meta-llama/Llama-3.2-11B-Vision Not documented context Input not listed Output not listed
Open weights

No provider description is available for this model yet.

meta-llama/Llama-Guard-3-1B Not documented context Input not listed Output not listed

No provider description is available for this model yet.

bedrock/ai21.j2-ultra-v1 8.191K context $18.8/M input $18.8/M output

No provider description is available for this model yet.

perplexity/llama-2-70b-chat 4.096K context $0.7/M input $2.8/M output

No provider description is available for this model yet.

bedrock/ai21.jamba-1-5-mini-v1:0 256K context $0.2/M input $0.4/M output
Open weights

No provider description is available for this model yet.

microsoft/Mage-VL Not documented context Input not listed Output not listed

No provider description is available for this model yet.

bedrock/ai21.jamba-instruct-v1:0 70K context $0.5/M input $0.7/M output
Open weights

No provider description is available for this model yet.

meta-llama/Llama-Guard-3-1B-INT4 Not documented context Input not listed Output not listed

Gemini 2.5 Pro is Google’s state-of-the-art AI model designed for advanced reasoning, coding, mathematics, and scientific tasks. It employs “thinking” capabilities, enabling it to reason through responses with enhanced accuracy...

google/gemini-2.5-pro-preview-05-06 1.04858M context $1.25/M input $10/M output
Open weights

No provider description is available for this model yet.

meta-llama/Llama-3.1-405B-Instruct Not documented context Input not listed Output not listed
Open weights

No provider description is available for this model yet.

meta-llama/Llama-3.1-405B Not documented context Input not listed Output not listed
Open weights

No provider description is available for this model yet.

meta-llama/Llama-3.1-70B Not documented context Input not listed Output not listed

No provider description is available for this model yet.

deepinfra/qwen/qwen3.8-27b 262.144K context $0.4/M input $3/M output
Open weights

No provider description is available for this model yet.

microsoft/Mage-Flow-Turbo Not documented context Input not listed Output not listed
Open weights

No provider description is available for this model yet.

meta-llama/Llama-3.1-8B-Instruct Not documented context Input not listed Output not listed

No provider description is available for this model yet.

nvidia/Qwen-Image-Flash Not documented context Input not listed Output not listed

No provider description is available for this model yet.

deepinfra/minimaxai/minimax-m2.7 196.608K context $0.25/M input $1/M output
Open weights

No provider description is available for this model yet.

meta-llama/Llama-Guard-3-8B-INT8 Not documented context Input not listed Output not listed
Open weights

No provider description is available for this model yet.

meta-llama/Meta-Llama-Guard-2-8B Not documented context Input not listed Output not listed

No provider description is available for this model yet.

perplexity/codellama-70b-instruct 16.384K context $0.7/M input $2.8/M output