MiniMax-M2.7 is a next-generation large language model designed for autonomous, real-world productivity and continuous improvement. Built to actively participate in its own evolution, M2.7 integrates advanced agentic capabilities through multi-agent...
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
No provider description is available for this model yet.
No provider description is available for this model yet.
GPT-6 Astra is OpenAI's flagship model for demanding end-to-end work. It is suited for advanced analysis, software engineering, deep research, scientific work, and document creation, with particular strengths in long-horizon...
No provider description is available for this model yet.
No provider description is available for this model yet.
GLM 5.2 is a large-scale reasoning model from Z.ai. It supports text input and output with a 1M-token context window, and is suited for long-horizon agent workflows, project-level software engineering,...
Claude Sonnet 4.5 is Anthropic’s most advanced Sonnet model to date, optimized for real-world agents and coding workflows. It delivers state-of-the-art performance on coding benchmarks such as SWE-bench Verified, with...
No provider description is available for this model yet.
No provider description is available for this model yet.
Mistral-Small-3.2-24B-Instruct-2506 is an updated 24B parameter model from Mistral optimized for instruction following, repetition reduction, and improved function calling. Compared to the 3.1 release, version 3.2 significantly improves accuracy on...
No provider description is available for this model yet.
Mistral's cutting-edge language model for coding released end of July 2025. Codestral specializes in low-latency, high-frequency tasks such as fill-in-the-middle (FIM), code correction and test generation. [Blog Post](https://mistral.ai/news/codestral-25-08)
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Qwen Plus 0728, based on the Qwen3 foundation model, is a 1 million context hybrid reasoning model with a balanced performance, speed, and cost combination.
No provider description is available for this model yet.
No provider description is available for this model yet.
Claude Sonnet 4 significantly enhances the capabilities of its predecessor, Sonnet 3.7, excelling in both coding and reasoning tasks with improved precision and controllability. Achieving state-of-the-art performance on SWE-bench (72.7%),...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
GPT-5.6 Terra is a balanced model in OpenAI's GPT-5.6 series, positioned between the flagship Sol tier and the cost-efficient Luna tier. It is suited for everyday coding, reasoning, and agentic...
MiniMax-M3 is a multimodal foundation model from MiniMax. It supports text, image, and video inputs with text output, a 1M-token context window, and is suited for long-horizon agentic work, coding,...
Nex-N2.5 is an agentic model built to turn goals into working, verified outcomes. Its core strength is agentic coding within a visual feedback loop: it can explore codebases, implement multi-file...
MiniMax-M3 is a multimodal foundation model from MiniMax. It supports text, image, and video inputs with text output, a 1M-token context window, and is suited for long-horizon agentic work, coding,...
Muse Glimmer 30B is a dense, open-weight multimodal model from Meta Superintelligence Labs, distilled from Muse Spark and optimized for autonomous agents on consumer hardware. It is suited for long-horizon...
OpenAI o3-mini is a cost-efficient language model optimized for STEM reasoning tasks, particularly excelling in science, mathematics, and coding. This model supports the `reasoning_effort` parameter, which can be set to...
GPT-5.4 is OpenAI’s latest frontier model, unifying the Codex and GPT lines into a single system. It features a 1M+ token context window (922K input, 128K output) with support for...
DeepSeek V4 Flash 0731 is a sparse mixture-of-experts model from DeepSeek, with 13B active parameters out of 284B total. This re-post-trained revision is suited for coding, reasoning, and agent workflows....
Sonnet 5 is Anthropic's most capable Sonnet-class model, with frontier performance across coding, agents, and professional work. It supports adaptive thinking with selectable reasoning effort levels (low, medium, high, max,...
Claude Opus 4.8 is Anthropic's most capable generally available model in the Opus family. It supports text, image, and file inputs with text output, with reasoning support and a 1M-token...
Devstral 2 is a state-of-the-art open-source model by Mistral AI specializing in agentic coding. It is a 123B-parameter dense transformer model supporting a 256K context window. Devstral 2 supports exploring...
Mistral Large 3 2512 is Mistral’s most capable model to date, featuring a sparse mixture-of-experts architecture with 41B active parameters (675B total), and released under the Apache 2.0 license.
MiniMax-M2 is a compact, high-efficiency large language model optimized for end-to-end coding and agentic workflows. With 10 billion activated parameters (230 billion total), it delivers near-frontier intelligence across general reasoning,...
Compared with GLM-4.5, this generation brings several key improvements: Longer context window: The context window has been expanded from 128K to 200K tokens, enabling the model to handle more complex...
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| MiniMax: MiniMax M2.7minimax/minimax-m2.7 | 204.8K | $0.3 | $1.2 | — | |||
| qwen3-coder-flash-2025-07-28qwen_ai_platform/qwen3-coder-flash-2025-07-28 | 997.952K | — | — | — | |||
| qwen3-coder-plusqwen_ai_platform/qwen3-coder-plus | 997.952K | — | — | — | |||
| OpenAI: GPT-6 Astra (batch)openai/gpt-6-astra:batch | 1.05M | $5 | $25 | — | |||
| qwen3-coder-plus-2025-07-22qwen_ai_platform/qwen3-coder-plus-2025-07-22 | 997.952K | — | — | — | |||
| qwen3-max-previewqwen_ai_platform/qwen3-max-preview | 258.048K | — | — | — | |||
| Z.ai: GLM 5.2 (batch)z-ai/glm-5.2:batch | 1.04858M | $0.7 | $2.2 | — | |||
| Anthropic: Claude Sonnet 4.5anthropic/claude-sonnet-4.5 | 1M | $3 | $15 | — | |||
| qwen3-maxqwen_ai_platform/qwen3-max | 258.048K | — | — | — | |||
| qwen3-max-2026-01-23qwen_ai_platform/qwen3-max-2026-01-23 | 258.048K | — | — | — | |||
| Mistral: Mistral Small 3.2 24Bmistralai/mistral-small-3.2-24b-instruct | 128K | $0.075 | $0.2 | — | |||
| qwen3-next-80b-a3b-instructqwen_ai_platform/qwen3-next-80b-a3b-instruct | 262.144K | $0.15 | $1.2 | — | |||
| Mistral: Codestral 2508 (batch)mistralai/codestral-2508:batch | 256K | $0.15 | $0.45 | — | |||
| qwen3-next-80b-a3b-thinkingqwen_ai_platform/qwen3-next-80b-a3b-thinking | 262.144K | $0.15 | $1.2 | — | |||
| qwen3-vl-235b-a22b-instructqwen_ai_platform/qwen3-vl-235b-a22b-instruct | 131.072K | $0.4 | $1.6 | — | |||
| qwen3-vl-235b-a22b-thinkingqwen_ai_platform/qwen3-vl-235b-a22b-thinking | 131.072K | $0.4 | $4 | — | |||
| qwen3-vl-32b-instructqwen_ai_platform/qwen3-vl-32b-instruct | 131.072K | $0.16 | $0.64 | — | |||
| qwen3-vl-32b-thinkingqwen_ai_platform/qwen3-vl-32b-thinking | 131.072K | $0.16 | $2.87 | — | |||
| qwen3-vl-plusqwen_ai_platform/qwen3-vl-plus | 260.096K | — | — | — | |||
| qwen3.5-plusqwen_ai_platform/qwen3.5-plus | 991.808K | — | — | — | |||
| Qwen: Qwen Plus 0728qwen/qwen-plus-2025-07-28 | 1M | $0.26 | $0.78 | — | |||
| qwen3.7-maxqwen_ai_platform/qwen3.7-max | 991.808K | $2.5 | $7.5 | — | |||
| qwen3.7-plusqwen_ai_platform/qwen3.7-plus | 991.808K | — | — | — | |||
| Anthropic: Claude Sonnet 4anthropic/claude-sonnet-4 | 200K | $3 | $15 | — | |||
| qwen3.8-maxqwen_ai_platform/qwen3.8-max | 991.808K | $2 | $6 | — | |||
| qwq-plusqwen_ai_platform/qwq-plus | 98.304K | $0.8 | $2.4 | — | |||
| databricks-deepseek-v4-flash-0731databricks/databricks-deepseek-v4-flash-0731 | 1M | $0.14 | $0.28 | — | |||
| databricks-deepseek-v4-pro-0813databricks/databricks-deepseek-v4-pro-0813 | 1M | $1.32 | $3.96 | — | |||
| zai-org/GLM-5.3-Flashfriendliai/zai-org/glm-5.3-flash | 1.04858M | $0.15 | $0.5 | — | |||
| zai-org/GLM-5.3friendliai/zai-org/glm-5.3 | 1.04858M | $1.26 | $3.96 | — | |||
| GigaChat-2gigachat/gigachat-2 | 128K | — | — | — | |||
| claude-fable-5-1vertex_ai-anthropic_models/claude-fable-5-1 | 1M | $10 | $50 | — | |||
| claude-fable-5-1@defaultvertex_ai-anthropic_models/claude-fable-5-1@default | 1M | $10 | $50 | — | |||
| glm-5.2zai/glm-5.2 | 1M | $1.4 | $4.4 | — | |||
| Qwen/Qwen3.8-Flashtogether_ai/qwen/qwen3.8-flash | 1M | $0.15 | $0.47 | — | |||
| gemma-4-31bcerebras/gemma-4-31b | 131.072K | $0.99 | $1.49 | — | |||
| OpenAI: GPT-5.6 Terra (batch)openai/gpt-5.6-terra:batch | 1.05M | $1 | $6 | — | |||
| MiniMax: MiniMax M3 (free)minimax/minimax-m3:free | 1.04858M | Free | Free | — | |||
| Nex AGI: Nex-N2.5-Pro (free)nex-agi/nex-n2.5-pro:free | 262.144K | Free | Free | — | |||
| MiniMax: MiniMax M3 (batch)minimax/minimax-m3:batch | 524.288K | $0.3 | $1.2 | — | |||
| Meta: Muse Glimmer 30B (batch)meta/muse-glimmer-30b:batch | 131.072K | $0.175 | $0.75 | — | |||
| OpenAI: o3 Mini (batch)openai/o3-mini:batch | 200K | $0.55 | $2.2 | — | |||
| OpenAI: GPT-5.4 (batch)openai/gpt-5.4:batch | 1.05M | $1.25 | $7.5 | — | |||
| DeepSeek: DeepSeek V4 Flash 0731 (batch)deepseek/deepseek-v4-flash-0731:batch | 1.04858M | $0.11 | $0.33 | — | |||
| Anthropic: Claude Sonnet 5 (batch)anthropic/claude-sonnet-5:batch | 1M | $1 | $5 | — | |||
| Anthropic: Claude Opus 4.8 (batch)anthropic/claude-opus-4.8:batch | 1M | $2.5 | $12.5 | — | |||
| Mistral: Devstral 2 2512mistralai/devstral-2512 | 262.144K | $0.4 | $2 | — | |||
| Mistral: Mistral Large 3 2512 (batch)mistralai/mistral-large-2512:batch | 262.144K | $0.25 | $0.75 | — | |||
| MiniMax: MiniMax M2minimax/minimax-m2 | 204.8K | $0.255 | $1.02 | — | |||
| Z.ai: GLM 4.6z-ai/glm-4.6 | 198K | $0.43 | $1.75 | — |