3,229 models

No provider description is available for this model yet.

azure/gpt-4o-2024-05-13 128K context $5/M input $15/M output

No provider description is available for this model yet.

together_ai/minimaxai/minimax-m3 524.288K context $0.3/M input $1.2/M output

No provider description is available for this model yet.

azure/gpt-4o-2024-08-06 128K context $2.5/M input $10/M output

No provider description is available for this model yet.

azure/gpt-4o-2024-11-20 128K context $2.75/M input $11/M output

GPT-4o Search Previewis a specialized model for web search in Chat Completions. It is trained to understand and execute web search queries.

openai/gpt-4o-search-preview 128K context $2.5/M input $10/M output

No provider description is available for this model yet.

azure/gpt-audio-2025-08-28 128K context $2.5/M input $10/M output

No provider description is available for this model yet.

together_ai/qwen/qwen3.5-9b 262.144K context $0.17/M input $0.25/M output

No provider description is available for this model yet.

together_ai/qwen/qwen3.6-plus 1M context $0.5/M input $3/M output

No provider description is available for this model yet.

azure/gpt-audio-1.5-2026-02-23 128K context $2.5/M input $10/M output

Dots3-Note Preview is an open-weight mixture-of-experts model from Dots Studio, with 16B active parameters out of 280B total. It is the lightest model in the Dots 3 family and is...

dots-studio/dots-3-note-preview:free 512K context Free input Free output

No provider description is available for this model yet.

together_ai/qwen/qwen3.7-max 1M context $1.25/M input $3.75/M output

No provider description is available for this model yet.

azure/gpt-audio-mini-2025-10-06 128K context $0.6/M input $2.4/M output

Llama 4 Scout 17B Instruct (16E) is a mixture-of-experts (MoE) language model developed by Meta, activating 17 billion parameters out of a total of 109B. It supports native multimodal input...

meta-llama/llama-4-scout 327.68K context $0.1/M input $0.3/M output

No provider description is available for this model yet.

azure/gpt-4o-mini 128K context $0.165/M input $0.66/M output

No provider description is available for this model yet.

bedrock/us-west-2/deepseek.v3.2 163.84K context $0.62/M input $1.85/M output

No provider description is available for this model yet.

azure/gpt-5.1-2025-11-13 272K context $1.25/M input $10/M output

Qwen3-32B is a dense 32.8B parameter causal language model from the Qwen3 series, optimized for both complex reasoning and efficient dialogue. It supports seamless switching between a "thinking" mode for...

qwen/qwen3-32b 40.96K context $0.08/M input $0.28/M output

Morph's high-accuracy apply model for complex code edits. ~4,500 tokens/sec with 98% accuracy for precise code transformations. The model requires the prompt to be in the following format: <instruction>{instruction}</instruction> <code>{initial_code}</code>...

morph/morph-v3-large 262.144K context $0.9/M input $1.9/M output

No provider description is available for this model yet.

azure/gpt-5-chat-latest 128K context $1.25/M input $10/M output

Qwen3-8B is a dense 8.2B parameter causal language model from the Qwen3 series, designed for both reasoning-heavy tasks and efficient dialogue. It supports seamless switching between "thinking" mode for math,...

qwen/qwen3-8b 131.072K context $0.117/M input $0.455/M output

OpenAI o4-mini-high is the same model as [o4-mini](/openai/o4-mini) with reasoning_effort set to high. OpenAI o4-mini is a compact reasoning model in the o-series, optimized for fast, cost-efficient performance while retaining...

openai/o4-mini-high 200K context $1.1/M input $4.4/M output

No provider description is available for this model yet.

azure/gpt-5-mini-2025-08-07 272K context $0.25/M input $2/M output

No provider description is available for this model yet.

azure/gpt-5-nano 272K context $0.05/M input $0.4/M output

No provider description is available for this model yet.

azure/gpt-5.1 272K context $1.25/M input $10/M output

No provider description is available for this model yet.

azure/gpt-5.1-chat 128K context $1.25/M input $10/M output

No provider description is available for this model yet.

azure/gpt-5.2 272K context $1.75/M input $14/M output

No provider description is available for this model yet.

bedrock_converse/us.xai.grok-4.6 500K context $2.2/M input $6.6/M output

Qwen3, the latest generation in the Qwen large language model series, features both dense and mixture-of-experts (MoE) architectures to excel in reasoning, multilingual support, and advanced agent tasks. Its unique...

qwen/qwen3-30b-a3b 40.96K context $0.12/M input $0.5/M output