3,311 models

No provider description is available for this model yet.

mistral/ministral-14b-latest 262.144K context $0.2/M input $0.2/M output

No provider description is available for this model yet.

aihubmix/muse-spark-1.2 1.04858M context $1.375/M input $4.675/M output

No provider description is available for this model yet.

mistral/ministral-3b-2512 131.072K context $0.1/M input $0.1/M output

No provider description is available for this model yet.

mistral/ministral-3b-latest 131.072K context $0.1/M input $0.1/M output

No provider description is available for this model yet.

mistral/mistral-medium-3 262.144K context $1.5/M input $7.5/M output

No provider description is available for this model yet.

mistral/voxtral-small-2507 32.768K context $0.1/M input $0.4/M output

No provider description is available for this model yet.

together_ai/zai-org/glm-5.3-flash 1.04858M context $0.15/M input $0.5/M output

No provider description is available for this model yet.

zai/glm-5.3 1M context $1.4/M input $4.4/M output

No provider description is available for this model yet.

tencent/minimax-m3 1M context $0.3/M input $1.2/M output

No provider description is available for this model yet.

novita/zai-org/glm-5.3 1.04858M context $1.4/M input $4.4/M output

No provider description is available for this model yet.

novita/moonshotai/kimi-k3 1.04858M context $3/M input $15/M output

No provider description is available for this model yet.

novita/tencent/hy3 262.144K context $0.14/M input $0.58/M output

No provider description is available for this model yet.

novita/zai-org/glm-5.2 1.04858M context $1.4/M input $4.4/M output

MiniMax-M1 is a large-scale, open-weight reasoning model designed for extended context and high-efficiency inference. It leverages a hybrid Mixture-of-Experts (MoE) architecture paired with a custom "lightning attention" mechanism, allowing it...

minimax/minimax-m1 1M context $0.4/M input $2.2/M output

No provider description is available for this model yet.

aihubmix/qwen3-coder-next 262.144K context $0.137/M input $0.548/M output

No provider description is available for this model yet.

novita/moonshotai/kimi-k2.7-code 262.144K context $0.95/M input $4/M output

Mistral Large 3 2512 is Mistral’s most capable model to date, featuring a sparse mixture-of-experts architecture with 41B active parameters (675B total), and released under the Apache 2.0 license.

mistralai/mistral-large-2512:batch 262.144K context $0.25/M input $0.75/M output

Fugu Max is the cost-performance model in Sakana AI's Fugu family. Rather than a single monolithic model, Fugu is a learned multi-agent orchestration system: a language model trained to route...

sakana/fugu-max 1M context $2/M input $6/M output

No provider description is available for this model yet.

aihubmix/qwen3.5-122b-a10b 262.144K context $0.113/M input $0.901/M output

No provider description is available for this model yet.

aihubmix/qwen3.5-397b-a17b 262.144K context $0.164/M input $0.986/M output

No provider description is available for this model yet.

aihubmix/qwen3.6-27b 262.144K context $0.422/M input $2.532/M output

No provider description is available for this model yet.

novita/mindai/macaron-v1-venti 1.04858M context $1.5/M input $4.5/M output

No provider description is available for this model yet.

aihubmix/qwen3.6-35b-a3b 262.144K context $0.254/M input $1.524/M output

No provider description is available for this model yet.

aihubmix/qwen3.6-max-preview 262.144K context $1.268/M input $7.608/M output

No provider description is available for this model yet.

novita/minimax/minimax-m3 1M context $0.3/M input $1.2/M output

No provider description is available for this model yet.

novita/deepseek/deepseek-v4-flash 1.04858M context $0.14/M input $0.28/M output

No provider description is available for this model yet.

novita/deepseek/deepseek-v4-pro 1.04858M context $1.6/M input $3.2/M output

No provider description is available for this model yet.

aihubmix/qwen3.7-plus 1M context $0.282/M input $1.128/M output

No provider description is available for this model yet.

aihubmix/qwen3.8-2.4t-a95b 1M context $2/M input $6/M output

No provider description is available for this model yet.

novita/qwen/qwen3.8-max 1M context $2/M input $6/M output

No provider description is available for this model yet.

aihubmix/qwen3.8-flash 1M context $0.113/M input $0.38/M output

No provider description is available for this model yet.

aihubmix/qwen3.8-max 1M context $1.69/M input $5.07/M output

No provider description is available for this model yet.

novita/inclusionai/ling-3.0-flash 262.144K context $0.06/M input $0.18/M output

No provider description is available for this model yet.

aihubmix/step-3.7-flash 256K context $0.22/M input $1.32/M output

No provider description is available for this model yet.

novita/mindai/macaron-v1-tall 262.144K context $0.45/M input $2.6/M output

No provider description is available for this model yet.

novita/stepfun/step-3.7-flash 262.144K context $0.2/M input $1.15/M output

Schematron V2 Turbo is a 3B-parameter HTML-to-JSON extraction model from Inference.net. It prioritizes throughput for high-volume extraction workloads. Extraction instructions must be supplied through a JSON schema in response_format rather...

inference-net/schematron-v2-turbo 128K context $0.03/M input $0.15/M output

No provider description is available for this model yet.

novita/baidu/cobuddy 131.072K context $0.28/M input $1.13/M output

GPT-6 Astra Pro is the same underlying model as [GPT-6 Astra](https://openrouter.ai/openai/gpt-6-astra), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. Learn more in OpenAI's docs: https://developers.openai.com/api/docs/guides/reasoning#reasoning-mode

openai/gpt-6-astra-pro 1.05M context $10/M input $50/M output

No provider description is available for this model yet.

novita/xiaomimimo/mimo-v2.5 1.04858M context $0.168/M input $0.336/M output

No provider description is available for this model yet.

novita/qwen/qwen3.7-max 1M context $1.25/M input $3.75/M output

No provider description is available for this model yet.

novita/xiaomimimo/mimo-v2.5-pro 1.04858M context $0.522/M input $1.044/M output

No provider description is available for this model yet.

novita/qwen/qwen3.6-27b 262.144K context $0.6/M input $3.6/M output

*Ling-3.0-flash* is a *124B-parameter Mixture-of-Experts (MoE) model*, with approximately *5.1B parameters activated per token*. The model is designed with *token efficiency and production-scale agentic inference* as key priorities, enabling developers...

inclusionai/ling-3.0-flash 262.144K context $0.021/M input $0.063/M output

No provider description is available for this model yet.

novita/moonshotai/kimi-k2.6 262.144K context $0.8/M input $3.4/M output