MiniMax-M2.7 is a next-generation large language model designed for autonomous, real-world productivity and continuous improvement. Built to actively participate in its own evolution, M2.7 integrates advanced agentic capabilities through multi-agent...

minimax/minimax-m2.7:free 196.608K context Free input Free output

No provider description is available for this model yet.

novita/zai-org/glm-5-turbo 202.8K context $1.2/M input $4/M output

No provider description is available for this model yet.

llamagate/qwen3-vl-8b 32.768K context $0.15/M input $0.55/M output

No provider description is available for this model yet.

novita/google/gemma-4-31b-it 262.144K context $0.14/M input $0.4/M output

No provider description is available for this model yet.

novita/google/gemma-4-26b-a4b-it 262.144K context $0.13/M input $0.4/M output

No provider description is available for this model yet.

llamagate/qwen2.5-coder-7b 32.768K context $0.06/M input $0.12/M output

No provider description is available for this model yet.

novita/zai-org/glm-5v-turbo 204.8K context $1.2/M input $4/M output

GPT-4.1 Mini is a mid-sized model delivering performance competitive with GPT-4o at substantially lower latency and cost. It retains a 1 million token context window and scores 45.1% on hard...

openai/gpt-4.1-mini:batch 1.04758M context $0.2/M input $0.8/M output

No provider description is available for this model yet.

llamagate/openthinker-7b 32.768K context $0.08/M input $0.15/M output

Mistral Medium 3 is a high-performance enterprise-grade language model designed to deliver frontier-level capabilities at significantly reduced operational cost. It balances state-of-the-art reasoning and multimodal performance with 8× lower cost...

mistralai/mistral-medium-3 131.072K context $0.4/M input $2/M output

No provider description is available for this model yet.

llamagate/deepseek-r1-7b-qwen 131.072K context $0.08/M input $0.15/M output

No provider description is available for this model yet.

novita/zai-org/glm-5.1 204.8K context $1.38/M input $4.4/M output

No provider description is available for this model yet.

llamagate/deepseek-r1-8b 65.536K context $0.1/M input $0.2/M output

No provider description is available for this model yet.

novita/moonshotai/kimi-k2.6 262.144K context $0.8/M input $3.4/M output

No provider description is available for this model yet.

novita/qwen/qwen3.6-27b 262.144K context $0.6/M input $3.6/M output

No provider description is available for this model yet.

llamagate/dolphin3-8b 128K context $0.08/M input $0.15/M output

No provider description is available for this model yet.

novita/xiaomimimo/mimo-v2.5-pro 1.04858M context $0.522/M input $1.044/M output

No provider description is available for this model yet.

novita/qwen/qwen3.7-max 1M context $1.25/M input $3.75/M output

No provider description is available for this model yet.

llamagate/qwen3-8b 32.768K context $0.04/M input $0.14/M output

No provider description is available for this model yet.

novita/xiaomimimo/mimo-v2.5 1.04858M context $0.168/M input $0.336/M output

No provider description is available for this model yet.

qwen_ai_platform/qwq-plus 98.304K context $0.8/M input $2.4/M output

No provider description is available for this model yet.

llamagate/mistral-7b-v0.3 32.768K context $0.1/M input $0.15/M output

No provider description is available for this model yet.

qwen_ai_platform/qwen3.7-plus 991.808K context Input not listed Output not listed

No provider description is available for this model yet.

novita/baidu/cobuddy 131.072K context $0.28/M input $1.13/M output

No provider description is available for this model yet.

llamagate/llama-3.2-3b 131.072K context $0.04/M input $0.08/M output

No provider description is available for this model yet.

scaleway/google/gemma-3-27b-it 40K context $0.25/M input $0.5/M output

No provider description is available for this model yet.

qwen_ai_platform/qwen3.7-max 991.808K context $2.5/M input $7.5/M output

No provider description is available for this model yet.

qwen_ai_platform/qwen3.5-plus 991.808K context Input not listed Output not listed

GPT-5.4 nano is the most lightweight and cost-efficient variant of the GPT-5.4 family, optimized for speed-critical and high-volume tasks. It supports text and image inputs and is designed for low-latency...

openai/gpt-5.4-nano:batch 400K context $0.1/M input $0.625/M output

No provider description is available for this model yet.

llamagate/llama-3.1-8b 131.072K context $0.03/M input $0.05/M output

GPT-5-Nano is the smallest and fastest variant in the GPT-5 system, optimized for developer tools, rapid interactions, and ultra-low latency environments. While limited in reasoning depth compared to its larger...

openai/gpt-5-nano:batch 400K context $0.025/M input $0.2/M output

Muse Glimmer 30B is a dense, open-weight multimodal model from Meta Superintelligence Labs, distilled from Muse Spark and optimized for autonomous agents on consumer hardware. It is suited for long-horizon...

meta/muse-glimmer-30b:batch 131.072K context $0.175/M input $0.75/M output

No provider description is available for this model yet.

qwen_ai_platform/qwen3-vl-plus 260.096K context Input not listed Output not listed

No provider description is available for this model yet.

qwen_ai_platform/qwen3.8-max 991.808K context $2/M input $6/M output

No provider description is available for this model yet.

novita/stepfun/step-3.7-flash 262.144K context $0.2/M input $1.15/M output

No provider description is available for this model yet.

friendliai/zai-org/glm-5.3-flash 1.04858M context $0.15/M input $0.5/M output

GPT-5.5 is OpenAI’s frontier model designed for complex professional workloads, building on GPT-5.4 with stronger reasoning, higher reliability, and improved token efficiency on hard tasks. It features a 1M+ token...

openai/gpt-5.5:batch 1.05M context $2.5/M input $15/M output