4,182 models

Claude Haiku 4.5 is Anthropic’s fastest and most efficient model, delivering near-frontier intelligence at a fraction of the cost and latency of larger Claude models. Matching Claude Sonnet 4’s performance...

anthropic/claude-haiku-4.5 200K context $1/M input $5/M output

No provider description is available for this model yet.

azure_ai/claude-fable-5 1M context $10/M input $50/M output

No provider description is available for this model yet.

azure_ai/claude-opus-5 1M context $5/M input $25/M output

Hunyuan-A13B is a 13B active parameter Mixture-of-Experts (MoE) language model developed by Tencent, with a total parameter count of 80B and support for reasoning via Chain-of-Thought. It offers competitive benchmark...

tencent/hunyuan-a13b-instruct 131.072K context $0.14/M input $0.57/M output

GPT-5.4 Pro is OpenAI's most advanced model, building on GPT-5.4's unified architecture with enhanced reasoning capabilities for complex, high-stakes tasks. It features a 1M+ token context window (922K input, 128K...

openai/gpt-5.4-pro:batch 1.05M context $15/M input $90/M output

Fast-mode variant of [Opus 4.7](/anthropic/claude-opus-4.7) - identical capabilities with higher output speed at premium 6x pricing. Learn more in Anthropic's docs: https://platform.claude.com/docs/en/build-with-claude/fast-mode

anthropic/claude-opus-4.7-fast 1M context $30/M input $150/M output
Open weights

No provider description is available for this model yet.

Qwen/WebWorld-14B Not documented context Input not listed Output not listed

Llama 3.2 1B is a 1-billion-parameter language model focused on efficiently performing natural language tasks, such as summarization, dialogue, and multilingual text analysis. Its smaller size allows it to operate...

meta-llama/llama-3.2-1b-instruct 60K context $0.027/M input $0.201/M output
Open weights

No provider description is available for this model yet.

Qwen/WebWorld-8B Not documented context Input not listed Output not listed
Open weights

No provider description is available for this model yet.

Qwen/Qwen3.5-397B-A17B Not documented context Input not listed Output not listed

No provider description is available for this model yet.

azure_ai/claude-opus-4-8 1M context $5/M input $25/M output

Inflection 3 Productivity is optimized for following instructions. It is better for tasks requiring JSON output or precise adherence to provided guidelines. It has access to recent news. For emotional...

inflection/inflection-3-productivity 8K context $2.5/M input $10/M output

Llama 3.2 11B Vision is a multimodal model with 11 billion parameters, designed to handle tasks combining visual and textual data. It excels in tasks such as image captioning and...

meta-llama/llama-3.2-11b-vision-instruct 131.072K context $0.345/M input $0.345/M output

No provider description is available for this model yet.

azure_ai/claude-opus-4-1 200K context $15/M input $75/M output
Open weights

No provider description is available for this model yet.

Qwen/Qwen3.5-27B Not documented context Input not listed Output not listed
Open weights

No provider description is available for this model yet.

Qwen/Qwen3.6-35B-A3B Not documented context Input not listed Output not listed

Schematron V2 Turbo is a 3B-parameter HTML-to-JSON extraction model from Inference.net. It prioritizes throughput for high-volume extraction workloads. Extraction instructions must be supplied through a JSON schema in response_format rather...

inference-net/schematron-v2-turbo 128K context $0.03/M input $0.15/M output

Sakana Namazu is a Japanese-specialized reasoning model from Sakana AI, based on Kimi K2.6 with additional training for Japanese language and business contexts. It is suited for Japanese instruction following,...

sakana/namazu 262.144K context $0.95/M input $4/M output

No provider description is available for this model yet.

azure_ai/claude-sonnet-4-5 200K context $3/M input $15/M output

No provider description is available for this model yet.

cerebras/zai-glm-4.7 128K context $2.25/M input $2.75/M output

No provider description is available for this model yet.

ollama/llama3 8.192K context Input not listed Output not listed
Open weights

No provider description is available for this model yet.

Qwen/Qwen3.5-35B-A3B Not documented context Input not listed Output not listed
Open weights

No provider description is available for this model yet.

Qwen/Qwen3.5-9B-Base Not documented context Input not listed Output not listed
Open weights

No provider description is available for this model yet.

Qwen/Qwen3.5-35B-A3B-Base Not documented context Input not listed Output not listed
Open weights

No provider description is available for this model yet.

Qwen/Qwen3.5-0.8B-Base Not documented context Input not listed Output not listed
Open weights

No provider description is available for this model yet.

Qwen/Qwen3.5-2B-Base Not documented context Input not listed Output not listed
Open weights

No provider description is available for this model yet.

Qwen/Qwen3.5-4B-Base Not documented context Input not listed Output not listed

No provider description is available for this model yet.

azure_ai/claude-sonnet-5 1M context $2/M input $10/M output

No provider description is available for this model yet.

azure/computer-use-preview 8.192K context $3/M input $12/M output
Open weights

No provider description is available for this model yet.

Qwen/Qwen3.5-0.8B Not documented context Input not listed Output not listed

No provider description is available for this model yet.

azure/container Not documented context Input not listed Output not listed

No provider description is available for this model yet.

vertex_ai/gemini-3.8-flash 1.04858M context $0.75/M input $3.75/M output

No provider description is available for this model yet.

azure_ai/gpt-oss-120b 131.072K context $0.15/M input $0.6/M output

No provider description is available for this model yet.

ollama/llama2:7b 4.096K context Input not listed Output not listed
Open weights

No provider description is available for this model yet.

Qwen/Qwen3.5-2B Not documented context Input not listed Output not listed
Open weights

No provider description is available for this model yet.

Qwen/Qwen3.5-4B Not documented context Input not listed Output not listed
Open weights

No provider description is available for this model yet.

Qwen/Qwen3.5-9B Not documented context Input not listed Output not listed
Open weights

No provider description is available for this model yet.

Qwen/Qwen3-Coder-Next Not documented context Input not listed Output not listed
Open weights

No provider description is available for this model yet.

Qwen/Qwen-Image-2512 Not documented context Input not listed Output not listed
Open weights

No provider description is available for this model yet.

meta-llama/CodeLlama-34b-hf Not documented context Input not listed Output not listed