3,424 models

Mercury 2.5 is the fastest reasoning LLM, and the latest diffusion LLM (dLLM) from Inception. Instead of generating tokens sequentially, Mercury 2.5 produces and refines multiple tokens in parallel, achieving...

inception/mercury-2.5-preview 260K context $0.2/M input $0.75/M output

No provider description is available for this model yet.

anthropic/claude-4-opus-20250514 200K context $15/M input $75/M output

No provider description is available for this model yet.

azure_ai/meta-llama-3-70b-instruct 8.192K context $1.1/M input $0.37/M output

No provider description is available for this model yet.

azure/o3 200K context $2/M input $8/M output

No provider description is available for this model yet.

bedrock/ap-south-1/deepseek.v3.2 163.84K context $0.74/M input $2.22/M output

Mistral-Small-3.2-24B-Instruct-2506 is an updated 24B parameter model from Mistral optimized for instruction following, repetition reduction, and improved function calling. Compared to the 3.1 release, version 3.2 significantly improves accuracy on...

mistralai/mistral-small-3.2-24b-instruct 128K context $0.075/M input $0.2/M output

No provider description is available for this model yet.

cerebras/llama3.1-8b 128K context $0.1/M input $0.1/M output

No provider description is available for this model yet.

azure/eu/gpt-5.5-2026-04-23 1.05M context $5.5/M input $33/M output

No provider description is available for this model yet.

azure/gpt-5.5 1.05M context $5/M input $30/M output

No provider description is available for this model yet.

mistral/mistral-medium-2508 131.072K context $0.4/M input $2/M output

No provider description is available for this model yet.

bedrock/moonshotai.kimi-k2.5 262.144K context $0.6/M input $3.03/M output

No provider description is available for this model yet.

mistral/mistral-medium-2312 32K context $2.7/M input $8.1/M output

No provider description is available for this model yet.

mistral/mistral-medium 32K context $2.7/M input $8.1/M output

No provider description is available for this model yet.

mistral/mistral-large-3 262.144K context $0.5/M input $1.5/M output

No provider description is available for this model yet.

mistral/mistral-large-2407 128K context $3/M input $9/M output