No provider description is available for this model yet.
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
DeepSeek V3, a 685B-parameter, mixture-of-experts model, is the latest iteration of the flagship chat model family from the DeepSeek team. It succeeds the [DeepSeek V3](/deepseek/deepseek-chat-v3) model and performs really well...
No provider description is available for this model yet.
Gemma 3n E4B-it is optimized for efficient execution on mobile and low-resource devices, such as phones, laptops, and tablets. It supports multimodal inputs—including text, visual data, and audio—enabling diverse tasks...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Dots3-Note Preview is an open-weight mixture-of-experts model from Dots Studio, with 16B active parameters out of 280B total. It is the lightest model in the Dots 3 family and is...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Qwen3.6 27B is a dense 27-billion-parameter language model from the Qwen Team at Alibaba, released in April 2026. It features hybrid multimodal capabilities — accepting text, image, and video inputs...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| us-gov.nvidia.nemotron-nano-12b-v2bedrock_converse/us-gov.nvidia.nemotron-nano-12b-v2 | 128K | $0.24 | $0.72 | — | |||
| gpt-4.1-miniazure/gpt-4.1-mini | 1.04758M | $0.4 | $1.6 | — | |||
| @cf/moonshotai/kimi-k2.6cloudflare/@cf/moonshotai/kimi-k2.6 | 262.144K | $0.95 | $4 | — | |||
| databricks-claude-sonnet-5databricks/databricks-claude-sonnet-5 | 1M | $3 | $15 | — | |||
| gpt-4.1-mini-2025-04-14azure/gpt-4.1-mini-2025-04-14 | 1.04758M | $0.4 | $1.6 | — | |||
| gpt-4.1-nanoazure/gpt-4.1-nano | 1.04758M | $0.1 | $0.4 | — | |||
| gpt-4.1-nano-2025-04-14azure/gpt-4.1-nano-2025-04-14 | 1.04758M | $0.1 | $0.4 | — | |||
| gpt-4.5-previewazure/gpt-4.5-preview | 128K | $75 | $150 | — | |||
| gpt-4oazure/gpt-4o | 128K | $2.5 | $10 | — | |||
| gpt-4o-2024-05-13azure/gpt-4o-2024-05-13 | 128K | $5 | $15 | — | |||
| MiniMaxAI/MiniMax-M3together_ai/minimaxai/minimax-m3 | 524.288K | $0.3 | $1.2 | — | |||
| gpt-4o-2024-08-06azure/gpt-4o-2024-08-06 | 128K | $2.5 | $10 | — | |||
| DeepSeek: DeepSeek V3 0324deepseek/deepseek-chat-v3-0324 | 163.84K | $0.25 | $1 | — | |||
| Prism-ML/Ternary-Bonsai-27Btogether_ai/prism-ml/ternary-bonsai-27b | 262.144K | — | — | — | |||
| Google: Gemma 3n 4Bgoogle/gemma-3n-e4b-it | 32.768K | $0.06 | $0.12 | — | |||
| Qwen/Qwen3.5-9Btogether_ai/qwen/qwen3.5-9b | 262.144K | $0.17 | $0.25 | — | |||
| Qwen/Qwen3.6-Plustogether_ai/qwen/qwen3.6-plus | 1M | $0.5 | $3 | — | |||
| gpt-audio-1.5-2026-02-23azure/gpt-audio-1.5-2026-02-23 | 128K | $2.5 | $10 | — | |||
| Qwen/Qwen3.7-Maxtogether_ai/qwen/qwen3.7-max | 1M | $1.25 | $3.75 | — | |||
| gpt-audio-mini-2025-10-06azure/gpt-audio-mini-2025-10-06 | 128K | $0.6 | $2.4 | — | |||
| Dots Studio: Dots3-Note Preview (free)dots-studio/dots-3-note-preview:free | 512K | Free | Free | — | |||
| Qwen/Qwen3.7-Plustogether_ai/qwen/qwen3.7-plus | 1M | $0.32 | $1.28 | — | |||
| gpt-4o-audio-preview-2024-12-17azure/gpt-4o-audio-preview-2024-12-17 | 128K | $2.5 | $10 | — | |||
| Qwen/Qwen3.8-2.4T-A95Btogether_ai/qwen/qwen3.8-2.4t-a95b | 1.01M | $2.5 | $6.25 | — | |||
| arize-ai/qwen-2-1.5b-instructtogether_ai/arize-ai/qwen-2-1.5b-instruct | 32.768K | $0.1 | $0.1 | — | |||
| deepseek-ai/DeepSeek-V4-Flash-0731together_ai/deepseek-ai/deepseek-v4-flash-0731 | 1.04858M | $0.14 | $0.28 | — | |||
| gpt-4o-miniazure/gpt-4o-mini | 128K | $0.165 | $0.66 | — | |||
| deepseek-ai/DeepSeek-V4-Protogether_ai/deepseek-ai/deepseek-v4-pro | 512K | $1.74 | $3.48 | — | |||
| gpt-4o-mini-2024-07-18azure/gpt-4o-mini-2024-07-18 | 128K | $0.165 | $0.66 | — | |||
| gpt-4o-mini-audio-preview-2024-12-17azure/gpt-4o-mini-audio-preview-2024-12-17 | 128K | $2.5 | $10 | — | |||
| Qwen: Qwen3.6 27Bqwen/qwen3.6-27b | 262.144K | $0.3 | $2 | — | |||
| gpt-5.1-2025-11-13azure/gpt-5.1-2025-11-13 | 272K | $1.25 | $10 | — | |||
| gpt-5.1-chat-2025-11-13azure/gpt-5.1-chat-2025-11-13 | 128K | $1.25 | $10 | — | |||
| gpt-5azure/gpt-5 | 272K | $1.25 | $10 | — | |||
| deepseek-ai/DeepSeek-V4-Pro-0813together_ai/deepseek-ai/deepseek-v4-pro-0813 | 1.04858M | $1.32 | $3.96 | — | |||
| Qwen/Qwen3-VL-32B-Instructtogether_ai/qwen/qwen3-vl-32b-instruct | 262.144K | $0.5 | $1.5 | — | |||
| google/gemma-3n-E4B-ittogether_ai/google/gemma-3n-e4b-it | 32.768K | $0.06 | $0.12 | — | |||
| gpt-5-2025-08-07azure/gpt-5-2025-08-07 | 272K | $1.25 | $10 | — | |||
| @cf/meta/llama-3.2-3b-instructcloudflare/@cf/meta/llama-3.2-3b-instruct | 80K | $0.051 | $0.335 | — | |||
| gpt-5-chat-latestazure/gpt-5-chat-latest | 128K | $1.25 | $10 | — | |||
| us-gov.nvidia.nemotron-nano-3-30bbedrock_converse/us-gov.nvidia.nemotron-nano-3-30b | 262.144K | $0.072 | $0.288 | — | |||
| google/gemma-4-31B-ittogether_ai/google/gemma-4-31b-it | 262.144K | $0.39 | $0.97 | — | |||
| meta-llama/Llama-Guard-4-12Btogether_ai/meta-llama/llama-guard-4-12b | 1.04858M | $0.2 | $0.2 | — | |||
| @cf/zai-org/glm-4.7-flashcloudflare/@cf/zai-org/glm-4.7-flash | 131.072K | $0.061 | $0.4 | — | |||
| apac.amazon.nova-micro-v1:0bedrock_converse/apac.amazon.nova-micro-v1:0 | 128K | $0.037 | $0.148 | — | |||
| apac.amazon.nova-lite-v1:0bedrock_converse/apac.amazon.nova-lite-v1:0 | 300K | $0.063 | $0.252 | — | |||
| gpt-5-mini-2025-08-07azure/gpt-5-mini-2025-08-07 | 272K | $0.25 | $2 | — | |||
| gpt-5-nanoazure/gpt-5-nano | 272K | $0.05 | $0.4 | — | |||
| us-west-2/minimax.minimax-m2.1bedrock/us-west-2/minimax.minimax-m2.1 | 196K | $0.3 | $1.2 | — | |||
| FW-GLM-5.2azure_ai/fw-glm-5.2 | 1.04858M | $1.54 | $4.84 | — |