MiniMax-M3 is a multimodal foundation model from MiniMax. It supports text, image, and video inputs with text output, a 1M-token context window, and is suited for long-horizon agentic work, coding,...
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
No provider description is available for this model yet.
No provider description is available for this model yet.
This model always redirects to the latest model in the GPT Terra family.
No provider description is available for this model yet.
Claude Sonnet 4.5 is Anthropic’s most advanced Sonnet model to date, optimized for real-world agents and coding workflows. It delivers state-of-the-art performance on coding benchmarks such as SWE-bench Verified, with...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Sonnet 5 is Anthropic's most capable Sonnet-class model, with frontier performance across coding, agents, and professional work. It supports adaptive thinking with selectable reasoning effort levels (low, medium, high, max,...
No provider description is available for this model yet.
Gemini 2.5 Pro is Google’s state-of-the-art AI model designed for advanced reasoning, coding, mathematics, and scientific tasks. It employs “thinking” capabilities, enabling it to reason through responses with enhanced accuracy...
GPT-4.1 is a flagship large language model optimized for advanced instruction following, real-world software engineering, and long-context reasoning. It supports a 1 million token context window and outperforms GPT-4o and...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
GLM-5.3-Flash is a native multimodal model from Z.ai. It is suited for efficient coding and long-horizon agent tasks. Its hybrid sparse and linear attention architecture maintains accurate long-context behavior while...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Claude Opus 4.8 is Anthropic's most capable generally available model in the Opus family. It supports text, image, and file inputs with text output, with reasoning support and a 1M-token...
No provider description is available for this model yet.
Claude Sonnet 4 significantly enhances the capabilities of its predecessor, Sonnet 3.7, excelling in both coding and reasoning tasks with improved precision and controllability. Achieving state-of-the-art performance on SWE-bench (72.7%),...
Mistral-Small-3.2-24B-Instruct-2506 is an updated 24B parameter model from Mistral optimized for instruction following, repetition reduction, and improved function calling. Compared to the 3.1 release, version 3.2 significantly improves accuracy on...
GLM-4.5V is a vision-language foundation model for multimodal agent applications. Built on a Mixture-of-Experts (MoE) architecture with 106B parameters and 12B activated parameters, it achieves state-of-the-art results in video understanding,...
No provider description is available for this model yet.
Opus 4.7 is the next generation of Anthropic's Opus family, built for long-running, asynchronous agents. Building on the coding and agentic strengths of Opus 4.6, it delivers stronger performance on...
Gemini 2.5 Flash is Google's state-of-the-art workhorse model, specifically designed for advanced reasoning, coding, mathematics, and scientific tasks. It includes built-in "thinking" capabilities, enabling it to provide responses with greater...
No provider description is available for this model yet.
Claude Haiku 4.5 is Anthropic’s fastest and most efficient model, delivering near-frontier intelligence at a fraction of the cost and latency of larger Claude models. Matching Claude Sonnet 4’s performance...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Qwen3-VL-235B-A22B Thinking is a multimodal model that unifies strong text generation with visual understanding across images and video. The Thinking model is optimized for multimodal reasoning in STEM and math....
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Mistral Medium 3.1 is an updated version of Mistral Medium 3, which is a high-performance enterprise-grade language model designed to deliver frontier-level capabilities at significantly reduced operational cost. It balances...
No provider description is available for this model yet.
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| MiniMax: MiniMax M3 (free)minimax/minimax-m3:free | 1.04858M | Free | Free | — | |||
| openai/gpt-4.1vercel_ai_gateway/openai/gpt-4.1 | 1.04758M | $2 | $8 | — | |||
| kimi-k2.7-codeqwencloud/kimi-k2.7-code | 229.376K | $0.95 | $4 | — | |||
| OpenAI: GPT Terra Latest~openai/gpt-terra-latest | 1.05M | $2 | $12 | — | |||
| openai/gpt-4-turbovercel_ai_gateway/openai/gpt-4-turbo | 128K | $10 | $30 | — | |||
| Anthropic: Claude Sonnet 4.5 (batch)anthropic/claude-sonnet-4.5:batch | 1M | $1.5 | $7.5 | — | |||
| meta.llama3-2-11b-instruct-v1:0bedrock/meta.llama3-2-11b-instruct-v1:0 | 128K | $0.35 | $0.35 | — | |||
| gemini-2.0-flash-litegemini/gemini-2.0-flash-lite | 1.04858M | $0.075 | $0.3 | — | |||
| mistral/pixtral-largevercel_ai_gateway/mistral/pixtral-large | 128K | $2 | $6 | — | |||
| Anthropic: Claude Sonnet 5 (batch)anthropic/claude-sonnet-5:batch | 1M | $1 | $5 | — | |||
| kimi-k3moonshot/kimi-k3 | 1.04858M | $3 | $15 | — | |||
| Google: Gemini 2.5 Pro Preview 05-06google/gemini-2.5-pro-preview-05-06 | 1.04858M | $1.25 | $10 | — | |||
| OpenAI: GPT-4.1 (batch)openai/gpt-4.1:batch | 1.04758M | $1 | $4 | — | |||
| mistral/pixtral-12bvercel_ai_gateway/mistral/pixtral-12b | 128K | $0.15 | $0.15 | — | |||
| llama3.2-11b-vision-instructlambda_ai/llama3.2-11b-vision-instruct | 131.072K | $0.015 | $0.025 | — | |||
| claude-fable-5-1azure_ai/claude-fable-5-1 | 1M | $10 | $50 | — | |||
| Z.ai: GLM 5.3 Flashz-ai/glm-5.3-flash | 1.04858M | $0.15 | $0.5 | — | |||
| mistral/ministral-8bvercel_ai_gateway/mistral/ministral-8b | 128K | $0.1 | $0.1 | — | |||
| eu.anthropic.claude-fable-5-1bedrock_converse/eu.anthropic.claude-fable-5-1 | 1M | $11 | $55 | — | |||
| meta/llama-4-scoutvercel_ai_gateway/meta/llama-4-scout | 131.072K | $0.1 | $0.3 | — | |||
| google/gemma-3-12b-itcrusoe/google/gemma-3-12b-it | 131.072K | $0.1 | $0.1 | — | |||
| gemini-2.0-flash-001gemini/gemini-2.0-flash-001 | 1.04858M | $0.1 | $0.4 | — | |||
| gpt-5.1github_copilot/gpt-5.1 | 128K | — | — | — | |||
| us.anthropic.claude-fable-5-1bedrock_converse/us.anthropic.claude-fable-5-1 | 1M | $11 | $55 | — | |||
| meta/llama-3.2-90bvercel_ai_gateway/meta/llama-3.2-90b | 128K | $0.72 | $0.72 | — | |||
| global.anthropic.claude-fable-5-1bedrock_converse/global.anthropic.claude-fable-5-1 | 1M | $10 | $50 | — | |||
| anthropic.claude-fable-5-1bedrock_converse/anthropic.claude-fable-5-1 | 1M | $10 | $50 | — | |||
| meta/llama-3.2-11bvercel_ai_gateway/meta/llama-3.2-11b | 128K | $0.16 | $0.16 | — | |||
| jp.anthropic.claude-haiku-4-5-20251001-v1:0bedrock_converse/jp.anthropic.claude-haiku-4-5-20251001-v1:0 | 200K | $1.1 | $5.5 | — | |||
| Anthropic: Claude Opus 4.8 (batch)anthropic/claude-opus-4.8:batch | 1M | $2.5 | $12.5 | — | |||
| google/gemini-2.5-provercel_ai_gateway/google/gemini-2.5-pro | 1.04858M | $2.5 | $10 | — | |||
| Anthropic: Claude Sonnet 4anthropic/claude-sonnet-4 | 200K | $3 | $15 | — | |||
| Mistral: Mistral Small 3.2 24Bmistralai/mistral-small-3.2-24b-instruct | 128K | $0.075 | $0.2 | — | |||
| Z.ai: GLM 4.5Vz-ai/glm-4.5v | 65.536K | $0.6 | $1.8 | — | |||
| google/gemini-2.5-flashvercel_ai_gateway/google/gemini-2.5-flash | 1M | $0.3 | $2.5 | — | |||
| Anthropic: Claude Opus 4.7 (batch)anthropic/claude-opus-4.7:batch | 1M | $2.5 | $12.5 | — | |||
| Google: Gemini 2.5 Flash (batch)google/gemini-2.5-flash:batch | 1.04858M | $0.15 | $1.25 | — | |||
| jp.anthropic.claude-sonnet-4-5-20250929-v1:0bedrock_converse/jp.anthropic.claude-sonnet-4-5-20250929-v1:0 | 200K | $3.3 | $16.5 | — | |||
| Anthropic: Claude Haiku 4.5 (batch)anthropic/claude-haiku-4.5:batch | 200K | $0.5 | $2.5 | — | |||
| gemini-2.0-flashgemini/gemini-2.0-flash | 1.04858M | $0.1 | $0.4 | — | |||
| google/gemma-4-31B-itdeepinfra/google/gemma-4-31b-it | 262.144K | $0.13 | $0.38 | — | |||
| gpt-chat-latestazure_ai/gpt-chat-latest | 272K | $5 | $30 | — | |||
| google/gemini-2.0-flash-litevercel_ai_gateway/google/gemini-2.0-flash-lite | 1.04858M | $0.075 | $0.3 | — | |||
| Qwen: Qwen3 VL 235B A22B Thinkingqwen/qwen3-vl-235b-a22b-thinking | 131.072K | $0.4 | $4 | — | |||
| xai/grok-4.3vertex_ai/xai/grok-4.3 | 200K | $1.25 | $2.5 | — | |||
| grok-4-20-reasoningazure_ai/grok-4-20-reasoning | 262K | $1.25 | $2.5 | — | |||
| xai/grok-4.6vertex_ai/xai/grok-4.6 | 524.288K | $2 | $6 | — | |||
| grok-4-20-non-reasoningazure_ai/grok-4-20-non-reasoning | 262K | $1.25 | $2.5 | — | |||
| Mistral: Mistral Medium 3.1mistralai/mistral-medium-3.1 | 131.072K | $0.4 | $2 | — | |||
| Qwen/Qwen3.5-9Bdeepinfra/qwen/qwen3.5-9b | 262.144K | $0.1 | $0.15 | — |