No provider description is available for this model yet.

libertai/gemma-4-31b-it-thinking 262.144K context $0.15/M input $0.4/M output

No provider description is available for this model yet.

llamagate/deepseek-r1-7b-qwen 131.072K context $0.08/M input $0.15/M output

No provider description is available for this model yet.

llamagate/qwen3-8b 32.768K context $0.04/M input $0.14/M output

No provider description is available for this model yet.

databricks/databricks-gpt-5-4 272K context $2.5/M input $15/M output

No provider description is available for this model yet.

databricks/databricks-gpt-5-2 272K context $1.75/M input $14/M output

No provider description is available for this model yet.

openrouter/openai/gpt-5.6-sol 1.05M context $2/M input $10/M output

Fugu Max is the cost-performance model in Sakana AI's Fugu family. Rather than a single monolithic model, Fugu is a learned multi-agent orchestration system: a language model trained to route...

sakana/fugu-max 1M context $2/M input $6/M output

No provider description is available for this model yet.

databricks/databricks-gemini-3-pro 1.04858M context $2.5/M input $15/M output

No provider description is available for this model yet.

novita/deepseek/deepseek-r1 64K context $4/M input $4/M output

Gemini 2.5 Flash-Lite is a lightweight reasoning model in the Gemini 2.5 family, optimized for ultra-low latency and cost efficiency. It offers improved throughput, faster token generation, and better performance...

google/gemini-2.5-flash-lite:batch 1.04858M context $0.05/M input $0.2/M output

No provider description is available for this model yet.

gemini/gemini-2.0-flash-lite-001 1.04858M context $0.075/M input $0.3/M output

No provider description is available for this model yet.

novita/deepseek/deepseek_v3 64K context $0.89/M input $0.89/M output

Hermes 4 is a large-scale reasoning model built on Meta-Llama-3.1-405B and released by Nous Research. It introduces a hybrid reasoning mode, where the model can choose to deliberate internally with...

nousresearch/hermes-4-405b 131.072K context $1/M input $3/M output

No provider description is available for this model yet.

novita/qwen/qwen3.6-35b-a3b 262.144K context $0.248/M input $1.485/M output

No provider description is available for this model yet.

novita/zai-org/glm-4.7-flash 200K context $0.07/M input $0.4/M output

GPT-5 Pro is OpenAI’s most advanced model, offering major improvements in reasoning, code quality, and user experience. It is optimized for complex tasks that require step-by-step reasoning, instruction following, and...

openai/gpt-5-pro:batch 400K context $7.5/M input $60/M output

No provider description is available for this model yet.

bedrock/us-west-2/zai.glm-5 200K context $1/M input $3.2/M output

No provider description is available for this model yet.

novita/zai-org/glm-4.7-h 204.8K context $0.6/M input $2.2/M output

No provider description is available for this model yet.

novita/moonshotai/kimi-k2.5 262.144K context $0.6/M input $3/M output

Mistral Medium 3.1 is an updated version of Mistral Medium 3, which is a high-performance enterprise-grade language model designed to deliver frontier-level capabilities at significantly reduced operational cost. It balances...

mistralai/mistral-medium-3.1:batch 131.072K context $0.2/M input $1/M output

Ministral 8B is an 8B parameter model featuring a unique interleaved sliding-window attention pattern for faster, memory-efficient inference. Designed for edge use cases, it supports up to 128k context length...

mistralai/ministral-8b 128K context $0.11/M input $0.11/M output

No provider description is available for this model yet.

bedrock/us-east-1/zai.glm-5 200K context $1/M input $3.2/M output

No provider description is available for this model yet.

libertai/gemma-4-31b-it 262.144K context $0.15/M input $0.4/M output

No provider description is available for this model yet.

novita/qwen/qwen3-coder-next 262.144K context $0.2/M input $1.5/M output

No provider description is available for this model yet.

qwencloud/qwen-flash-2025-07-28 997.952K context Input not listed Output not listed

No provider description is available for this model yet.

novita/zai-org/glm-5 202.8K context $1/M input $3.2/M output

No provider description is available for this model yet.

novita/minimax/minimax-m2.5 204.8K context $0.3/M input $1.2/M output

No provider description is available for this model yet.

qwencloud/qwen-flash 997.952K context Input not listed Output not listed

No provider description is available for this model yet.

qwencloud/qwen-coder 1M context $0.3/M input $1.5/M output

No provider description is available for this model yet.

novita/qwen/qwen3.5-397b-a17b 262.144K context $0.6/M input $3.6/M output

GPT-5.5 is OpenAI’s frontier model designed for complex professional workloads, building on GPT-5.4 with stronger reasoning, higher reliability, and improved token efficiency on hard tasks. It features a 1M+ token...

openai/gpt-5.5:batch 1.05M context $2.5/M input $15/M output

No provider description is available for this model yet.

qwencloud/kimi-k2.7-code 229.376K context $0.95/M input $4/M output

Muse Spark 1.3 Contributor is the cost-efficient contributor tier of Meta’s multimodal reasoning model for experimentation, learning, and early-stage agentic, multi-agent, and coding workflows. It is designed to track information...

meta/muse-spark-1.3-contributor 1.04858M context $0.1/M input $0.2/M output

GLM-5 is Z.ai’s flagship open-source foundation model engineered for complex systems design and long-horizon agent workflows. Built for expert developers, it delivers production-grade performance on large-scale programming tasks, rivaling leading...

z-ai/glm-5 198K context $0.6/M input $1.92/M output

No provider description is available for this model yet.

qwencloud/glm-5.2 1.04858M context $1.4/M input $4.4/M output

No provider description is available for this model yet.

qwencloud/qwen-plus 129.024K context $0.4/M input $1.2/M output

No provider description is available for this model yet.

qwencloud/qwen-plus-2025-01-25 129.024K context $0.4/M input $1.2/M output

No provider description is available for this model yet.

qwencloud/glm-5.1 202.745K context $1.4/M input $4.4/M output