No provider description is available for this model yet.
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Voxtral Small is an enhancement of Mistral Small 3, incorporating state-of-the-art audio input capabilities while retaining best-in-class text performance. It excels at speech transcription, translation and audio understanding. Input audio...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Claude Haiku 4.5 is Anthropic’s fastest and most efficient model, delivering near-frontier intelligence at a fraction of the cost and latency of larger Claude models. Matching Claude Sonnet 4’s performance...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Jamba Large 1.7 is the latest model in the Jamba open family, offering improvements in grounding, instruction-following, and overall efficiency. Built on a hybrid SSM-Transformer architecture with a 256K context...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Mistral Medium 3.5 is a dense 128B instruction-following model from Mistral AI. It supports text and image inputs with text output, and is designed for agentic workflows, coding, and complex...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Mistral's cutting-edge language model for coding released end of July 2025. Codestral specializes in low-latency, high-frequency tasks such as fill-in-the-middle (FIM), code correction and test generation. [Blog Post](https://mistral.ai/news/codestral-25-08)
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
MiniMax-M2 is a compact, high-efficiency large language model optimized for end-to-end coding and agentic workflows. With 10 billion activated parameters (230 billion total), it delivers near-frontier intelligence across general reasoning,...
Qwen3.8 Max (0803) is the August 3, 2026 checkpoint of Qwen3.8 Max, the flagship model in Alibaba's Qwen3.8 series and the general-availability successor to the Qwen3.8 Max Preview. It is...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Claude Opus 4.5 is Anthropic’s frontier reasoning model optimized for complex software engineering, agentic workflows, and long-horizon computer use. It offers strong multimodal capabilities, competitive performance across real-world coding and...
Qwen3-VL-30B-A3B-Instruct is a multimodal model that unifies strong text generation with visual understanding for images and videos. Its Instruct variant optimizes instruction-following for general multimodal tasks. It excels in perception...
No provider description is available for this model yet.
No provider description is available for this model yet.
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| gpt-5-2025-08-07azure/gpt-5-2025-08-07 | 272K | $1.25 | $10 | — | |||
| gpt-5-chatazure/gpt-5-chat | 128K | $1.25 | $10 | — | |||
| gpt-5-chat-latestazure/gpt-5-chat-latest | 128K | $1.25 | $10 | — | |||
| google/gemma-4-31B-ittogether_ai/google/gemma-4-31b-it | 262.144K | $0.39 | $0.97 | — | |||
| us-gov-west-1/amazon.nova-lite-v1:0bedrock/us-gov-west-1/amazon.nova-lite-v1:0 | 300K | $0.072 | $0.288 | — | |||
| meta-llama/Llama-Guard-4-12Btogether_ai/meta-llama/llama-guard-4-12b | 1.04858M | $0.2 | $0.2 | — | |||
| @cf/zai-org/glm-4.7-flashcloudflare/@cf/zai-org/glm-4.7-flash | 131.072K | $0.061 | $0.4 | — | |||
| Mistral: Voxtral Small 24B 2507mistralai/voxtral-small-24b-2507 | 32.768K | $0.1 | $0.3 | — | |||
| gpt-5-miniazure/gpt-5-mini | 272K | $0.25 | $2 | — | |||
| gpt-5-mini-2025-08-07azure/gpt-5-mini-2025-08-07 | 272K | $0.25 | $2 | — | |||
| gpt-5-nanoazure/gpt-5-nano | 272K | $0.05 | $0.4 | — | |||
| us-west-2/minimax.minimax-m2.1bedrock/us-west-2/minimax.minimax-m2.1 | 196K | $0.3 | $1.2 | — | |||
| meta-models/Muse-Glimmer-30Btogether_ai/meta-models/muse-glimmer-30b | 131.072K | $0.35 | $1.5 | — | |||
| moonshotai/Kimi-K2.7-Codetogether_ai/moonshotai/kimi-k2.7-code | 262.144K | $0.95 | $4 | — | |||
| Anthropic: Claude Haiku 4.5 (batch)anthropic/claude-haiku-4.5:batch | 200K | $0.5 | $2.5 | — | |||
| gpt-5-nano-2025-08-07azure/gpt-5-nano-2025-08-07 | 272K | $0.05 | $0.4 | — | |||
| gpt-5.1azure/gpt-5.1 | 272K | $1.25 | $10 | — | |||
| moonshotai/Kimi-K3together_ai/moonshotai/kimi-k3 | 1.04858M | $3 | $15 | — | |||
| gpt-5.1-chatazure/gpt-5.1-chat | 128K | $1.25 | $10 | — | |||
| us-west-2/minimax.minimax-m2.5bedrock/us-west-2/minimax.minimax-m2.5 | 1M | $0.3 | $1.2 | — | |||
| gpt-5.2azure/gpt-5.2 | 272K | $1.75 | $14 | — | |||
| gpt-5.2-2025-12-11azure/gpt-5.2-2025-12-11 | 272K | $1.75 | $14 | — | |||
| gpt-5.2-chatazure/gpt-5.2-chat | 128K | $1.75 | $14 | — | |||
| nvidia/nemotron-3-ultra-550b-a55btogether_ai/nvidia/nemotron-3-ultra-550b-a55b | 512.288K | $0.6 | $3.6 | — | |||
| us-west-2/moonshotai.kimi-k2-thinkingbedrock/us-west-2/moonshotai.kimi-k2-thinking | 262.144K | $0.6 | $2.5 | — | |||
| AI21: Jamba Large 1.7ai21/jamba-large-1.7 | 256K | $2 | $8 | — | |||
| gpt-5.2-chat-2025-12-11azure/gpt-5.2-chat-2025-12-11 | 128K | $1.75 | $14 | — | |||
| gpt-5.3-chatazure/gpt-5.3-chat | 128K | $1.75 | $14 | — | |||
| au.anthropic.claude-opus-4-8bedrock_converse/au.anthropic.claude-opus-4-8 | 1M | $5.5 | $27.5 | — | |||
| us-east-1/6-month-commitment/anthropic.claude-v2:1bedrock/us-east-1/6-month-commitment/anthropic.claude-v2:1 | 100K | — | — | — | |||
| Mistral: Mistral Medium 3.5 (batch)mistralai/mistral-medium-3-5:batch | 262.144K | $0.75 | $3.75 | — | |||
| pearl-ai/gemma-4-31b-ittogether_ai/pearl-ai/gemma-4-31b-it | 262.144K | $0.28 | $0.86 | — | |||
| us/gpt-5.4azure/us/gpt-5.4 | 1.05M | $2.75 | $16.5 | — | |||
| eu/gpt-5.4azure/eu/gpt-5.4 | 1.05M | $2.75 | $16.5 | — | |||
| Mistral: Codestral 2508 (batch)mistralai/codestral-2508:batch | 256K | $0.15 | $0.45 | — | |||
| thinkingmachines/Inklingtogether_ai/thinkingmachines/inkling | 524.288K | $1 | $4.05 | — | |||
| thinkingmachines/Inkling-Smalltogether_ai/thinkingmachines/inkling-small | 524.288K | $0.5 | $1.2 | — | |||
| us-west-2/moonshotai.kimi-k2.5bedrock/us-west-2/moonshotai.kimi-k2.5 | 262.144K | $0.6 | $3 | — | |||
| MiniMax: MiniMax M2minimax/minimax-m2 | 204.8K | $0.255 | $1.02 | — | |||
| Qwen: Qwen3.8 Max (0803)qwen/qwen3.8-max | 1M | $2 | $6 | — | |||
| gpt-5.4-2026-03-05azure/gpt-5.4-2026-03-05 | 1.05M | $2.5 | $15 | — | |||
| us/gpt-5.4-2026-03-05azure/us/gpt-5.4-2026-03-05 | 1.05M | $2.75 | $16.5 | — | |||
| eu/gpt-5.4-2026-03-05azure/eu/gpt-5.4-2026-03-05 | 1.05M | $2.75 | $16.5 | — | |||
| xai.grok-4.6bedrock_mantle/xai.grok-4.6 | 500K | $2.2 | $6.6 | — | |||
| gpt-5.6azure/gpt-5.6 | 1.05M | $5 | $30 | — | |||
| gpt-5.6-solazure/gpt-5.6-sol | 1.05M | $5 | $30 | — | |||
| Anthropic: Claude Opus 4.5anthropic/claude-opus-4.5 | 200K | $5 | $25 | — | |||
| Qwen: Qwen3 VL 30B A3B Instructqwen/qwen3-vl-30b-a3b-instruct | 262.144K | $0.15 | $0.6 | — | |||
| eu.anthropic.claude-opus-4-8bedrock_converse/eu.anthropic.claude-opus-4-8 | 1M | $5.5 | $27.5 | — | |||
| us.anthropic.claude-opus-4-8bedrock_converse/us.anthropic.claude-opus-4-8 | 1M | $5.5 | $27.5 | — |