No provider description is available for this model yet.
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
GPT-4o Search Previewis a specialized model for web search in Chat Completions. It is trained to understand and execute web search queries.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Dots3-Note Preview is an open-weight mixture-of-experts model from Dots Studio, with 16B active parameters out of 280B total. It is the lightest model in the Dots 3 family and is...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Llama 4 Scout 17B Instruct (16E) is a mixture-of-experts (MoE) language model developed by Meta, activating 17 billion parameters out of a total of 109B. It supports native multimodal input...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Qwen3-32B is a dense 32.8B parameter causal language model from the Qwen3 series, optimized for both complex reasoning and efficient dialogue. It supports seamless switching between a "thinking" mode for...
Morph's high-accuracy apply model for complex code edits. ~4,500 tokens/sec with 98% accuracy for precise code transformations. The model requires the prompt to be in the following format: <instruction>{instruction}</instruction> <code>{initial_code}</code>...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Qwen3-8B is a dense 8.2B parameter causal language model from the Qwen3 series, designed for both reasoning-heavy tasks and efficient dialogue. It supports seamless switching between "thinking" mode for math,...
No provider description is available for this model yet.
OpenAI o4-mini-high is the same model as [o4-mini](/openai/o4-mini) with reasoning_effort set to high. OpenAI o4-mini is a compact reasoning model in the o-series, optimized for fast, cost-efficient performance while retaining...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Qwen3, the latest generation in the Qwen large language model series, features both dense and mixture-of-experts (MoE) architectures to excel in reasoning, multilingual support, and advanced agent tasks. Its unique...
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| gpt-4o-2024-05-13azure/gpt-4o-2024-05-13 | 128K | $5 | $15 | — | |||
| MiniMaxAI/MiniMax-M3together_ai/minimaxai/minimax-m3 | 524.288K | $0.3 | $1.2 | — | |||
| gpt-4o-2024-08-06azure/gpt-4o-2024-08-06 | 128K | $2.5 | $10 | — | |||
| gpt-4o-2024-11-20azure/gpt-4o-2024-11-20 | 128K | $2.75 | $11 | — | |||
| OpenAI: GPT-4o Search Previewopenai/gpt-4o-search-preview | 128K | $2.5 | $10 | — | |||
| gpt-audio-2025-08-28azure/gpt-audio-2025-08-28 | 128K | $2.5 | $10 | — | |||
| Qwen/Qwen3.5-9Btogether_ai/qwen/qwen3.5-9b | 262.144K | $0.17 | $0.25 | — | |||
| Qwen/Qwen3.6-Plustogether_ai/qwen/qwen3.6-plus | 1M | $0.5 | $3 | — | |||
| gpt-audio-1.5-2026-02-23azure/gpt-audio-1.5-2026-02-23 | 128K | $2.5 | $10 | — | |||
| Dots Studio: Dots3-Note Preview (free)dots-studio/dots-3-note-preview:free | 512K | Free | Free | — | |||
| Qwen/Qwen3.7-Maxtogether_ai/qwen/qwen3.7-max | 1M | $1.25 | $3.75 | — | |||
| gpt-audio-mini-2025-10-06azure/gpt-audio-mini-2025-10-06 | 128K | $0.6 | $2.4 | — | |||
| us-gov.anthropic.claude-3-haiku-20240307-v1:0bedrock_converse/us-gov.anthropic.claude-3-haiku-20240307-v1:0 | 200K | $0.3 | $1.5 | — | |||
| gpt-4o-audio-preview-2024-12-17azure/gpt-4o-audio-preview-2024-12-17 | 128K | $2.5 | $10 | — | |||
| Meta: Llama 4 Scoutmeta-llama/llama-4-scout | 327.68K | $0.1 | $0.3 | — | |||
| arize-ai/qwen-2-1.5b-instructtogether_ai/arize-ai/qwen-2-1.5b-instruct | 32.768K | $0.1 | $0.1 | — | |||
| deepseek-ai/DeepSeek-V4-Flash-0731together_ai/deepseek-ai/deepseek-v4-flash-0731 | 1.04858M | $0.14 | $0.28 | — | |||
| gpt-4o-miniazure/gpt-4o-mini | 128K | $0.165 | $0.66 | — | |||
| us-east-1/mistral.mistral-large-2402-v1:0bedrock/us-east-1/mistral.mistral-large-2402-v1:0 | 32K | $8 | $24 | — | |||
| deepseek-ai/DeepSeek-V4-Protogether_ai/deepseek-ai/deepseek-v4-pro | 512K | $1.74 | $3.48 | — | |||
| us.anthropic.claude-sonnet-4-6bedrock_converse/us.anthropic.claude-sonnet-4-6 | 1M | $3.3 | $16.5 | — | |||
| gpt-4o-mini-audio-preview-2024-12-17azure/gpt-4o-mini-audio-preview-2024-12-17 | 128K | $2.5 | $10 | — | |||
| us-west-2/deepseek.v3.2bedrock/us-west-2/deepseek.v3.2 | 163.84K | $0.62 | $1.85 | — | |||
| gpt-5.1-2025-11-13azure/gpt-5.1-2025-11-13 | 272K | $1.25 | $10 | — | |||
| us-east-1/mistral.mistral-7b-instruct-v0:2bedrock/us-east-1/mistral.mistral-7b-instruct-v0:2 | 32K | $0.15 | $0.2 | — | |||
| Qwen: Qwen3 32Bqwen/qwen3-32b | 40.96K | $0.08 | $0.28 | — | |||
| Morph: Morph V3 Largemorph/morph-v3-large | 262.144K | $0.9 | $1.9 | — | |||
| global.anthropic.claude-sonnet-4-6bedrock_converse/global.anthropic.claude-sonnet-4-6 | 1M | $3 | $15 | — | |||
| anthropic.claude-sonnet-4-6bedrock_converse/anthropic.claude-sonnet-4-6 | 1M | $3 | $15 | — | |||
| us-east-1/anthropic.claude-v2:1bedrock/us-east-1/anthropic.claude-v2:1 | 100K | $8 | $24 | — | |||
| jp.anthropic.claude-sonnet-5bedrock_converse/jp.anthropic.claude-sonnet-5 | 1M | $2.2 | $11 | — | |||
| gpt-5-chat-latestazure/gpt-5-chat-latest | 128K | $1.25 | $10 | — | |||
| au.anthropic.claude-sonnet-5bedrock_converse/au.anthropic.claude-sonnet-5 | 1M | $2.2 | $11 | — | |||
| us-east-1/anthropic.claude-v1bedrock/us-east-1/anthropic.claude-v1 | 100K | $8 | $24 | — | |||
| Qwen: Qwen3 8Bqwen/qwen3-8b | 131.072K | $0.117 | $0.455 | — | |||
| @cf/zai-org/glm-4.7-flashcloudflare/@cf/zai-org/glm-4.7-flash | 131.072K | $0.061 | $0.4 | — | |||
| OpenAI: o4 Mini Highopenai/o4-mini-high | 200K | $1.1 | $4.4 | — | |||
| gpt-5-mini-2025-08-07azure/gpt-5-mini-2025-08-07 | 272K | $0.25 | $2 | — | |||
| gpt-5-nanoazure/gpt-5-nano | 272K | $0.05 | $0.4 | — | |||
| us-west-2/minimax.minimax-m2.1bedrock/us-west-2/minimax.minimax-m2.1 | 196K | $0.3 | $1.2 | — | |||
| global.xai.grok-4.6bedrock_converse/global.xai.grok-4.6 | 500K | $2 | $6 | — | |||
| moonshotai/Kimi-K2.7-Codetogether_ai/moonshotai/kimi-k2.7-code | 262.144K | $0.95 | $4 | — | |||
| eu.anthropic.claude-sonnet-5bedrock_converse/eu.anthropic.claude-sonnet-5 | 1M | $2.2 | $11 | — | |||
| gpt-5.1azure/gpt-5.1 | 272K | $1.25 | $10 | — | |||
| us.anthropic.claude-sonnet-5bedrock_converse/us.anthropic.claude-sonnet-5 | 1M | $2.2 | $11 | — | |||
| gpt-5.1-chatazure/gpt-5.1-chat | 128K | $1.25 | $10 | — | |||
| us-west-2/minimax.minimax-m2.5bedrock/us-west-2/minimax.minimax-m2.5 | 1M | $0.3 | $1.2 | — | |||
| gpt-5.2azure/gpt-5.2 | 272K | $1.75 | $14 | — | |||
| us.xai.grok-4.6bedrock_converse/us.xai.grok-4.6 | 500K | $2.2 | $6.6 | — | |||
| Qwen: Qwen3 30B A3Bqwen/qwen3-30b-a3b | 40.96K | $0.12 | $0.5 | — |