4,213 models

No provider description is available for this model yet.

ollama/mistral-7b-instruct-v0.1 8.192K context Input not listed Output not listed

No provider description is available for this model yet.

ollama/mistral-7b-instruct-v0.2 32.768K context Input not listed Output not listed

Qwen3-14B is a dense 14.8B parameter causal language model from the Qwen3 series, designed for both complex reasoning and efficient dialogue. It supports seamless switching between a "thinking" mode for...

qwen/qwen3-14b 40.96K context $0.12/M input $0.24/M output

No provider description is available for this model yet.

azure/us/gpt-5.6-sol 1.05M context $5.5/M input $33/M output

No provider description is available for this model yet.

azure/us/gpt-5.6-terra 1.05M context $2.75/M input $16.5/M output

No provider description is available for this model yet.

ollama/mistral-large-instruct-2407 65.536K context Input not listed Output not listed

No provider description is available for this model yet.

ollama/mixtral-8x22b-instruct-v0.1 65.536K context Input not listed Output not listed

No provider description is available for this model yet.

azure/us/gpt-5.6-luna 1.05M context $1.1/M input $6.6/M output

No provider description is available for this model yet.

azure/eu/gpt-5.6 1.05M context $5.5/M input $33/M output

Claude Opus 4 is benchmarked as the world’s best coding model, at time of release, bringing sustained performance on complex, long-running tasks and agent workflows. It sets new benchmarks in...

anthropic/claude-opus-4 200K context $15/M input $75/M output

A cost-efficient version of GPT Audio. The new snapshot features an upgraded decoder for more natural sounding voices and maintains better voice consistency. Input is priced at $0.60 per million...

openai/gpt-audio-mini 128K context $0.6/M input $2.4/M output

No provider description is available for this model yet.

azure/eu/gpt-5.6-sol 1.05M context $5.5/M input $33/M output

No provider description is available for this model yet.

perplexity/pplx-70b-online 4.096K context Input not listed $2.8/M output

No provider description is available for this model yet.

perplexity/pplx-7b-chat 8.192K context $0.07/M input $0.28/M output

No provider description is available for this model yet.

azure/eu/gpt-5.6-terra 1.05M context $2.75/M input $16.5/M output

No provider description is available for this model yet.

azure/eu/gpt-5.6-luna 1.05M context $1.1/M input $6.6/M output

No provider description is available for this model yet.

perplexity/pplx-7b-online 4.096K context Input not listed $0.28/M output

No provider description is available for this model yet.

perplexity/sonar-medium-chat 16.384K context $0.6/M input $1.8/M output

No provider description is available for this model yet.

perplexity/sonar-medium-online 12K context Input not listed $1.8/M output

No provider description is available for this model yet.

azure/gpt-5.5 1.05M context $5/M input $30/M output

No provider description is available for this model yet.

azure/us/gpt-5.5 1.05M context $5.5/M input $33/M output

No provider description is available for this model yet.

azure/eu/gpt-5.5 1.05M context $5.5/M input $33/M output

No provider description is available for this model yet.

azure/gpt-5.5-2026-04-23 1.05M context $5/M input $30/M output

No provider description is available for this model yet.

cerebras/llama-3.3-70b 128K context $0.85/M input $1.2/M output

No provider description is available for this model yet.

azure/us/gpt-5.5-2026-04-23 1.05M context $5.5/M input $33/M output

No provider description is available for this model yet.

azure/eu/gpt-5.5-2026-04-23 1.05M context $5.5/M input $33/M output

Mercury 2.5 is the fastest reasoning LLM, and the latest diffusion LLM (dLLM) from Inception. Instead of generating tokens sequentially, Mercury 2.5 produces and refines multiple tokens in parallel, achieving...

inception/mercury-2.5-preview 260K context $0.2/M input $0.75/M output

No provider description is available for this model yet.

azure/gpt-5.4-mini 272K context $0.75/M input $4.5/M output

No provider description is available for this model yet.

cerebras/llama3.1-70b 128K context $0.6/M input $0.6/M output
Open weights

No provider description is available for this model yet.

ibm-granite/granite-4.2-30b Not documented context Input not listed Output not listed