4,324 models

MiniMax-M3 is a multimodal foundation model from MiniMax. It supports text, image, and video inputs with text output, a 1M-token context window, and is suited for long-horizon agentic work, coding,...

minimax/minimax-m3:batch 524.288K context $0.3/M input $1.2/M output

Qwen3.8 Max 0902 is an updated snapshot of Qwen3.8 Max from Alibaba's Qwen team. It is a 2.4-trillion-parameter mixture-of-experts model that accepts text, image, and video input and returns text,...

qwen/qwen3.8-max-0902 1M context $2/M input $6/M output

Gemini 3.5 Flash is Google's high-efficiency multimodal model, bringing near-Pro level coding and reasoning at Flash-tier cost and speed. It is highly optimized for coding proficiency and parallel agentic execution...

google/gemini-3.5-flash:batch 1.04858M context $0.75/M input $4.5/M output
Open weights

No provider description is available for this model yet.

Qwen/Qwen-Drive-1.0-4B Not documented context Input not listed Output not listed

Opus 4.6 is Anthropic’s strongest model for coding and long-running professional tasks. It is built for agents that operate across entire workflows rather than single prompts, making it especially effective...

anthropic/claude-opus-4.6:batch 1M context $2.5/M input $12.5/M output

A balanced model in the Ministral 3 family, Ministral 3 8B is a powerful, efficient tiny language model with vision capabilities.

mistralai/ministral-8b-2512 262.144K context $0.15/M input $0.15/M output

GPT-5.4 is OpenAI’s latest frontier model, unifying the Codex and GPT lines into a single system. It features a 1M+ token context window (922K input, 128K output) with support for...

openai/gpt-5.4:batch 1.05M context $1.25/M input $7.5/M output

Mistral Large 3 2512 is Mistral’s most capable model to date, featuring a sparse mixture-of-experts architecture with 41B active parameters (675B total), and released under the Apache 2.0 license.

mistralai/mistral-large-2512 262.144K context $0.5/M input $1.5/M output
Open weights

No provider description is available for this model yet.

deepseek-ai/DeepSeek-V4.1-Flash Not documented context Input not listed Output not listed

DeepSeek-V3.1 Terminus is an update to [DeepSeek V3.1](/deepseek/deepseek-chat-v3.1) that maintains the model's original capabilities while addressing issues reported by users, including language consistency and agent capabilities, further optimizing the model's...

deepseek/deepseek-v3.1-terminus 131.072K context $0.27/M input $1/M output

GPT-5 Pro is OpenAI’s most advanced model, offering major improvements in reasoning, code quality, and user experience. It is optimized for complex tasks that require step-by-step reasoning, instruction following, and...

openai/gpt-5-pro:batch 400K context $7.5/M input $60/M output

Compared with GLM-4.5, this generation brings several key improvements: Longer context window: The context window has been expanded from 128K to 200K tokens, enabling the model to handle more complex...

z-ai/glm-4.6 198K context $0.43/M input $1.75/M output

GPT-5 is OpenAI’s most advanced model, offering major improvements in reasoning, code quality, and user experience. It is optimized for complex tasks that require step-by-step reasoning, instruction following, and accuracy...

openai/gpt-5:batch 400K context $0.625/M input $5/M output

GPT-5.1 is the latest frontier-grade model in the GPT-5 series, offering stronger general-purpose reasoning, improved instruction adherence, and a more natural conversational style compared to GPT-5. It uses adaptive reasoning...

openai/gpt-5.1:batch 400K context $0.625/M input $5/M output

Qwen3-235B-A22B-Thinking-2507 is a high-performance, open-weight Mixture-of-Experts (MoE) language model optimized for complex reasoning tasks. It activates 22B of its 235B parameters per forward pass and natively supports up to 262,144...

qwen/qwen3-235b-a22b-thinking-2507 131.072K context $0.23/M input $2.3/M output

No provider description is available for this model yet.

azure/gpt-chat-latest 272K context $5/M input $30/M output

No provider description is available for this model yet.

azure/chat-latest 272K context $5/M input $30/M output

No provider description is available for this model yet.

azure/us/gpt-chat-latest 272K context $5.5/M input $33/M output

OpenAI o3-mini is a cost-efficient language model optimized for STEM reasoning tasks, particularly excelling in science, mathematics, and coding. This model supports the `reasoning_effort` parameter, which can be set to...

openai/o3-mini:batch 200K context $0.55/M input $2.2/M output

DeepSeek-V3.2-Exp is an experimental large language model released by DeepSeek as an intermediate step between V3.1 and future architectures. It introduces DeepSeek Sparse Attention (DSA), a fine-grained sparse attention mechanism...

deepseek/deepseek-v3.2-exp 163.84K context $0.27/M input $0.41/M output

Claude Sonnet 4 significantly enhances the capabilities of its predecessor, Sonnet 3.7, excelling in both coding and reasoning tasks with improved precision and controllability. Achieving state-of-the-art performance on SWE-bench (72.7%),...

anthropic/claude-sonnet-4 200K context $3/M input $15/M output

No provider description is available for this model yet.

azure_ai/gpt-chat-latest 272K context $5/M input $30/M output

No provider description is available for this model yet.

azure_ai/model-router 200K context $0.14/M input Output not listed

No provider description is available for this model yet.

azure_ai/cohere-command-a 131.072K context $2.5/M input $10/M output

No provider description is available for this model yet.

vertex_ai/xai/grok-4.3 200K context $1.25/M input $2.5/M output

No provider description is available for this model yet.

azure_ai/grok-4-20-reasoning 262K context $1.25/M input $2.5/M output

No provider description is available for this model yet.

vertex_ai/xai/grok-4.6 524.288K context $2/M input $6/M output

No provider description is available for this model yet.

azure_ai/grok-4-20-non-reasoning 262K context $1.25/M input $2.5/M output

No provider description is available for this model yet.

gemini/lyria-3.5 1.04858M context Input not listed Output not listed

No provider description is available for this model yet.

azure/gpt-6-astra 922K context $10/M input $50/M output

No provider description is available for this model yet.

azure/us/gpt-6-astra 922K context $11/M input $55/M output

No provider description is available for this model yet.

azure_ai/codestral-2501 256K context $0.3/M input $0.9/M output

No provider description is available for this model yet.

azure_ai/mai-thinking-1 256K context $2/M input $8/M output

No provider description is available for this model yet.

azure_ai/grok-4.6 200K context $2/M input $6/M output

MiniMax-M2 is a compact, high-efficiency large language model optimized for end-to-end coding and agentic workflows. With 10 billion activated parameters (230 billion total), it delivers near-frontier intelligence across general reasoning,...

minimax/minimax-m2 204.8K context $0.255/M input $1.02/M output

No provider description is available for this model yet.

databricks/databricks-glm-5-3 1.04858M context $1.4/M input $4.4/M output