3,239 models

Ministral 8B is an 8B parameter model featuring a unique interleaved sliding-window attention pattern for faster, memory-efficient inference. Designed for edge use cases, it supports up to 128k context length...

mistralai/ministral-8b 128K context $0.11/M input $0.11/M output

No provider description is available for this model yet.

dashscope/qwen3-max 258.048K context Input not listed Output not listed

No provider description is available for this model yet.

azure/gpt-4.1-mini-2025-04-14 1.04758M context $0.4/M input $1.6/M output

Fugu Max is the cost-performance model in Sakana AI's Fugu family. Rather than a single monolithic model, Fugu is a learned multi-agent orchestration system: a language model trained to route...

sakana/fugu-max 1M context $2/M input $6/M output

Gemini 2.5 Flash-Lite is a lightweight reasoning model in the Gemini 2.5 family, optimized for ultra-low latency and cost efficiency. It offers improved throughput, faster token generation, and better performance...

google/gemini-2.5-flash-lite:batch 1.04858M context $0.05/M input $0.2/M output

No provider description is available for this model yet.

openrouter/openai/gpt-5.6-sol 1.05M context $2/M input $10/M output

No provider description is available for this model yet.

dashscope/qwen3-max-preview 258.048K context Input not listed Output not listed

OpenAI o4-mini-high is the same model as [o4-mini](/openai/o4-mini) with reasoning_effort set to high. OpenAI o4-mini is a compact reasoning model in the o-series, optimized for fast, cost-efficient performance while retaining...

openai/o4-mini-high 200K context $1.1/M input $4.4/M output

No provider description is available for this model yet.

dashscope/qwen3-coder-plus-2025-07-22 997.952K context Input not listed Output not listed
Open weights

Granite 4.2 8B is a dense reasoning model from IBM. It is suited for mathematics, code generation, multilingual dialogue, and agentic workflows that need multi-step reasoning. It supports full, low-effort,...

ibm-granite/granite-4.2-8b 131.072K context $0.06/M input $0.25/M output

No provider description is available for this model yet.

openrouter/google/gemini-2.5-pro 1.04858M context $1.25/M input $10/M output

No provider description is available for this model yet.

dashscope/qwen3-coder-plus 997.952K context Input not listed Output not listed

No provider description is available for this model yet.

openrouter/google/gemini-2.5-flash 1.04858M context $0.3/M input $2.5/M output

No provider description is available for this model yet.

openrouter/deepseek/deepseek-r1 65.336K context $0.55/M input $2.19/M output

No provider description is available for this model yet.

databricks/databricks-glm-5-2 1M context $1.4/M input $4.4/M output

No provider description is available for this model yet.

databricks/databricks-kimi-k3 1M context $3/M input $15/M output

No provider description is available for this model yet.

mistral/ministral-14b-2512 262.144K context $0.2/M input $0.2/M output

No provider description is available for this model yet.

mistral/ministral-14b-latest 262.144K context $0.2/M input $0.2/M output

No provider description is available for this model yet.

mistral/ministral-3b-2512 131.072K context $0.1/M input $0.1/M output

No provider description is available for this model yet.

mistral/ministral-3b-latest 131.072K context $0.1/M input $0.1/M output

No provider description is available for this model yet.

mistral/mistral-medium-3 262.144K context $1.5/M input $7.5/M output

No provider description is available for this model yet.

mistral/voxtral-small-2507 32.768K context $0.1/M input $0.4/M output

No provider description is available for this model yet.

together_ai/zai-org/glm-5.3-flash 1.04858M context $0.15/M input $0.5/M output

No provider description is available for this model yet.

zai/glm-5.3 1M context $1.4/M input $4.4/M output

No provider description is available for this model yet.

tencent/minimax-m3 1M context $0.3/M input $1.2/M output

No provider description is available for this model yet.

novita/zai-org/glm-5.3 1.04858M context $1.4/M input $4.4/M output

No provider description is available for this model yet.

novita/moonshotai/kimi-k3 1.04858M context $3/M input $15/M output

No provider description is available for this model yet.

novita/tencent/hy3 262.144K context $0.14/M input $0.58/M output

No provider description is available for this model yet.

novita/zai-org/glm-5.2 1.04858M context $1.4/M input $4.4/M output

No provider description is available for this model yet.

dashscope/qwen3-coder-flash 997.952K context Input not listed Output not listed

No provider description is available for this model yet.

novita/moonshotai/kimi-k2.7-code 262.144K context $0.95/M input $4/M output

No provider description is available for this model yet.

openrouter/deepseek/deepseek-v3.2 163.84K context $0.28/M input $0.4/M output

Grok 4.20 is a reasoning model from SpaceXAI with industry-leading speed and agentic tool calling capabilities. It combines the lowest hallucination rate on the market with strict prompt adherance, delivering...

x-ai/grok-4.20 2M context $1.25/M input $2.5/M output

No provider description is available for this model yet.

dashscope/qwen3-30b-a3b 129.024K context Input not listed Output not listed