MiniMax-M2.7 is a next-generation large language model designed for autonomous, real-world productivity and continuous improvement. Built to actively participate in its own evolution, M2.7 integrates advanced agentic capabilities through multi-agent...
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
GPT-4.1 Mini is a mid-sized model delivering performance competitive with GPT-4o at substantially lower latency and cost. It retains a 1 million token context window and scores 45.1% on hard...
No provider description is available for this model yet.
No provider description is available for this model yet.
Mistral Medium 3 is a high-performance enterprise-grade language model designed to deliver frontier-level capabilities at significantly reduced operational cost. It balances state-of-the-art reasoning and multimodal performance with 8× lower cost...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
GPT-5.4 nano is the most lightweight and cost-efficient variant of the GPT-5.4 family, optimized for speed-critical and high-volume tasks. It supports text and image inputs and is designed for low-latency...
No provider description is available for this model yet.
GPT-5-Nano is the smallest and fastest variant in the GPT-5 system, optimized for developer tools, rapid interactions, and ultra-low latency environments. While limited in reasoning depth compared to its larger...
Muse Glimmer 30B is a dense, open-weight multimodal model from Meta Superintelligence Labs, distilled from Muse Spark and optimized for autonomous agents on consumer hardware. It is suited for long-horizon...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
GPT-5.5 is OpenAI’s frontier model designed for complex professional workloads, building on GPT-5.4 with stronger reasoning, higher reliability, and improved token efficiency on hard tasks. It features a 1M+ token...
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| MiniMax: MiniMax M2.7 (free)minimax/minimax-m2.7:free | 196.608K | Free | Free | — | |||
| zai-org/glm-5-turbonovita/zai-org/glm-5-turbo | 202.8K | $1.2 | $4 | — | |||
| qwen3-vl-8bllamagate/qwen3-vl-8b | 32.768K | $0.15 | $0.55 | — | |||
| mistralai/mistral-small-3.2-24b-instruct-2506scaleway/mistralai/mistral-small-3.2-24b-instruct-2506 | 128K | $0.15 | $0.35 | — | |||
| accounts/fireworks/models/qwen3-1p7bfireworks_ai/accounts/fireworks/models/qwen3-1p7b | 131.072K | $0.1 | $0.1 | — | |||
| google/gemma-4-31b-itnovita/google/gemma-4-31b-it | 262.144K | $0.14 | $0.4 | — | |||
| google/gemma-4-26b-a4b-itnovita/google/gemma-4-26b-a4b-it | 262.144K | $0.13 | $0.4 | — | |||
| qwen2.5-coder-7bllamagate/qwen2.5-coder-7b | 32.768K | $0.06 | $0.12 | — | |||
| zai-org/glm-5v-turbonovita/zai-org/glm-5v-turbo | 204.8K | $1.2 | $4 | — | |||
| OpenAI: GPT-4.1 Mini (batch)openai/gpt-4.1-mini:batch | 1.04758M | $0.2 | $0.8 | — | |||
| openthinker-7bllamagate/openthinker-7b | 32.768K | $0.08 | $0.15 | — | |||
| mistralai/voxtral-small-24b-2507scaleway/mistralai/voxtral-small-24b-2507 | 32K | $0.15 | $0.35 | — | |||
| Mistral: Mistral Medium 3mistralai/mistral-medium-3 | 131.072K | $0.4 | $2 | — | |||
| deepseek-r1-7b-qwenllamagate/deepseek-r1-7b-qwen | 131.072K | $0.08 | $0.15 | — | |||
| minimax/minimax-m2.7-highspeednovita/minimax/minimax-m2.7-highspeed | 204.8K | $0.6 | $2.4 | — | |||
| zai-org/glm-5.1novita/zai-org/glm-5.1 | 204.8K | $1.38 | $4.4 | — | |||
| deepseek-r1-8bllamagate/deepseek-r1-8b | 65.536K | $0.1 | $0.2 | — | |||
| mistralai/devstral-2-123b-instruct-2512scaleway/mistralai/devstral-2-123b-instruct-2512 | 200K | $0.4 | $2 | — | |||
| accounts/fireworks/models/qwen3-14bfireworks_ai/accounts/fireworks/models/qwen3-14b | 40.96K | $0.2 | $0.2 | — | |||
| accounts/fireworks/models/qwen2p5-coder-0p5bfireworks_ai/accounts/fireworks/models/qwen2p5-coder-0p5b | 32.768K | $0.1 | $0.1 | — | |||
| moonshotai/kimi-k2.6novita/moonshotai/kimi-k2.6 | 262.144K | $0.8 | $3.4 | — | |||
| qwen/qwen3.6-27bnovita/qwen/qwen3.6-27b | 262.144K | $0.6 | $3.6 | — | |||
| dolphin3-8bllamagate/dolphin3-8b | 128K | $0.08 | $0.15 | — | |||
| xiaomimimo/mimo-v2.5-pronovita/xiaomimimo/mimo-v2.5-pro | 1.04858M | $0.522 | $1.044 | — | |||
| qwen/qwen3.7-maxnovita/qwen/qwen3.7-max | 1M | $1.25 | $3.75 | — | |||
| qwen3-8bllamagate/qwen3-8b | 32.768K | $0.04 | $0.14 | — | |||
| mistralai/mistral-medium-3.5-128bscaleway/mistralai/mistral-medium-3.5-128b | 256K | $1.5 | $7.5 | — | |||
| xiaomimimo/mimo-v2.5novita/xiaomimimo/mimo-v2.5 | 1.04858M | $0.168 | $0.336 | — | |||
| qwq-plusqwen_ai_platform/qwq-plus | 98.304K | $0.8 | $2.4 | — | |||
| mistral-7b-v0.3llamagate/mistral-7b-v0.3 | 32.768K | $0.1 | $0.15 | — | |||
| qwen3.7-plusqwen_ai_platform/qwen3.7-plus | 991.808K | — | — | — | |||
| baidu/cobuddynovita/baidu/cobuddy | 131.072K | $0.28 | $1.13 | — | |||
| llama-3.2-3bllamagate/llama-3.2-3b | 131.072K | $0.04 | $0.08 | — | |||
| google/gemma-3-27b-itscaleway/google/gemma-3-27b-it | 40K | $0.25 | $0.5 | — | |||
| accounts/fireworks/models/qwen3-0p6bfireworks_ai/accounts/fireworks/models/qwen3-0p6b | 40.96K | $0.1 | $0.1 | — | |||
| qwen3.7-maxqwen_ai_platform/qwen3.7-max | 991.808K | $2.5 | $7.5 | — | |||
| qwen3.5-plusqwen_ai_platform/qwen3.5-plus | 991.808K | — | — | — | |||
| nvidia/nemotron-3-nano-30b-a3bnovita/nvidia/nemotron-3-nano-30b-a3b | 262.144K | $0.05 | $0.2 | — | |||
| OpenAI: GPT-5.4 Nano (batch)openai/gpt-5.4-nano:batch | 400K | $0.1 | $0.625 | — | |||
| llama-3.1-8bllamagate/llama-3.1-8b | 131.072K | $0.03 | $0.05 | — | |||
| OpenAI: GPT-5 Nano (batch)openai/gpt-5-nano:batch | 400K | $0.025 | $0.2 | — | |||
| Meta: Muse Glimmer 30B (batch)meta/muse-glimmer-30b:batch | 131.072K | $0.175 | $0.75 | — | |||
| qwen3-vl-plusqwen_ai_platform/qwen3-vl-plus | 260.096K | — | — | — | |||
| qwen3.8-maxqwen_ai_platform/qwen3.8-max | 991.808K | $2 | $6 | — | |||
| qwen3-vl-32b-thinkingqwen_ai_platform/qwen3-vl-32b-thinking | 131.072K | $0.16 | $2.87 | — | |||
| stepfun/step-3.7-flashnovita/stepfun/step-3.7-flash | 262.144K | $0.2 | $1.15 | — | |||
| databricks-deepseek-v4-flash-0731databricks/databricks-deepseek-v4-flash-0731 | 1M | $0.14 | $0.28 | — | |||
| databricks-deepseek-v4-pro-0813databricks/databricks-deepseek-v4-pro-0813 | 1M | $1.32 | $3.96 | — | |||
| zai-org/GLM-5.3-Flashfriendliai/zai-org/glm-5.3-flash | 1.04858M | $0.15 | $0.5 | — | |||
| OpenAI: GPT-5.5 (batch)openai/gpt-5.5:batch | 1.05M | $2.5 | $15 | — |