No provider description is available for this model yet.

azure/gpt-5-2025-08-07 272K context $1.25/M input $10/M output

No provider description is available for this model yet.

azure/gpt-5-chat 128K context $1.25/M input $10/M output

No provider description is available for this model yet.

azure/gpt-5-chat-latest 128K context $1.25/M input $10/M output

No provider description is available for this model yet.

together_ai/google/gemma-4-31b-it 262.144K context $0.39/M input $0.97/M output

Voxtral Small is an enhancement of Mistral Small 3, incorporating state-of-the-art audio input capabilities while retaining best-in-class text performance. It excels at speech transcription, translation and audio understanding. Input audio...

mistralai/voxtral-small-24b-2507 32.768K context $0.1/M input $0.3/M output

No provider description is available for this model yet.

azure/gpt-5-mini 272K context $0.25/M input $2/M output

No provider description is available for this model yet.

azure/gpt-5-mini-2025-08-07 272K context $0.25/M input $2/M output

No provider description is available for this model yet.

azure/gpt-5-nano 272K context $0.05/M input $0.4/M output

Claude Haiku 4.5 is Anthropic’s fastest and most efficient model, delivering near-frontier intelligence at a fraction of the cost and latency of larger Claude models. Matching Claude Sonnet 4’s performance...

anthropic/claude-haiku-4.5:batch 200K context $0.5/M input $2.5/M output

No provider description is available for this model yet.

azure/gpt-5-nano-2025-08-07 272K context $0.05/M input $0.4/M output

No provider description is available for this model yet.

azure/gpt-5.1 272K context $1.25/M input $10/M output

No provider description is available for this model yet.

together_ai/moonshotai/kimi-k3 1.04858M context $3/M input $15/M output

No provider description is available for this model yet.

azure/gpt-5.1-chat 128K context $1.25/M input $10/M output

No provider description is available for this model yet.

azure/gpt-5.2 272K context $1.75/M input $14/M output

No provider description is available for this model yet.

azure/gpt-5.2-2025-12-11 272K context $1.75/M input $14/M output

No provider description is available for this model yet.

azure/gpt-5.2-chat 128K context $1.75/M input $14/M output

Jamba Large 1.7 is the latest model in the Jamba open family, offering improvements in grounding, instruction-following, and overall efficiency. Built on a hybrid SSM-Transformer architecture with a 256K context...

ai21/jamba-large-1.7 256K context $2/M input $8/M output

No provider description is available for this model yet.

azure/gpt-5.2-chat-2025-12-11 128K context $1.75/M input $14/M output

No provider description is available for this model yet.

azure/gpt-5.3-chat 128K context $1.75/M input $14/M output

Mistral Medium 3.5 is a dense 128B instruction-following model from Mistral AI. It supports text and image inputs with text output, and is designed for agentic workflows, coding, and complex...

mistralai/mistral-medium-3-5:batch 262.144K context $0.75/M input $3.75/M output

No provider description is available for this model yet.

together_ai/pearl-ai/gemma-4-31b-it 262.144K context $0.28/M input $0.86/M output

No provider description is available for this model yet.

azure/us/gpt-5.4 1.05M context $2.75/M input $16.5/M output

No provider description is available for this model yet.

azure/eu/gpt-5.4 1.05M context $2.75/M input $16.5/M output

Mistral's cutting-edge language model for coding released end of July 2025. Codestral specializes in low-latency, high-frequency tasks such as fill-in-the-middle (FIM), code correction and test generation. [Blog Post](https://mistral.ai/news/codestral-25-08)

mistralai/codestral-2508:batch 256K context $0.15/M input $0.45/M output

MiniMax-M2 is a compact, high-efficiency large language model optimized for end-to-end coding and agentic workflows. With 10 billion activated parameters (230 billion total), it delivers near-frontier intelligence across general reasoning,...

minimax/minimax-m2 204.8K context $0.255/M input $1.02/M output

Qwen3.8 Max (0803) is the August 3, 2026 checkpoint of Qwen3.8 Max, the flagship model in Alibaba's Qwen3.8 series and the general-availability successor to the Qwen3.8 Max Preview. It is...

qwen/qwen3.8-max 1M context $2/M input $6/M output

No provider description is available for this model yet.

azure/gpt-5.4-2026-03-05 1.05M context $2.5/M input $15/M output

No provider description is available for this model yet.

azure/us/gpt-5.4-2026-03-05 1.05M context $2.75/M input $16.5/M output

No provider description is available for this model yet.

azure/eu/gpt-5.4-2026-03-05 1.05M context $2.75/M input $16.5/M output

No provider description is available for this model yet.

bedrock_mantle/xai.grok-4.6 500K context $2.2/M input $6.6/M output

No provider description is available for this model yet.

azure/gpt-5.6 1.05M context $5/M input $30/M output

No provider description is available for this model yet.

azure/gpt-5.6-sol 1.05M context $5/M input $30/M output

Claude Opus 4.5 is Anthropic’s frontier reasoning model optimized for complex software engineering, agentic workflows, and long-horizon computer use. It offers strong multimodal capabilities, competitive performance across real-world coding and...

anthropic/claude-opus-4.5 200K context $5/M input $25/M output

Qwen3-VL-30B-A3B-Instruct is a multimodal model that unifies strong text generation with visual understanding for images and videos. Its Instruct variant optimizes instruction-following for general multimodal tasks. It excels in perception...

qwen/qwen3-vl-30b-a3b-instruct 262.144K context $0.15/M input $0.6/M output