1,808 models

Qwen3.7 Flash is a vision-language reasoning model from Alibaba. It is suited for multimodal agents, visual coding, search, and computer interaction, with strengths in object recognition, spatial understanding, and real-world...

qwen/qwen3.7-flash 1M context $0.03/M input $0.13/M output

*Ling-3.0-flash* is a *124B-parameter Mixture-of-Experts (MoE) model*, with approximately *5.1B parameters activated per token*. The model is designed with *token efficiency and production-scale agentic inference* as key priorities, enabling developers...

inclusionai/ling-3.0-flash:free 262.144K context Free input Free output

MiniMax-M3 is a multimodal foundation model from MiniMax. It supports text, image, and video inputs with text output, a 1M-token context window, and is suited for long-horizon agentic work, coding,...

minimax/minimax-m3:batch 524.288K context $0.3/M input $1.2/M output

No provider description is available for this model yet.

gemini/gemma-4-26b-a4b-it 262.144K context Input not listed Output not listed

No provider description is available for this model yet.

novita/moonshotai/kimi-k2.6 262.144K context $0.8/M input $3.4/M output

GPT-5.6 Luna Pro is the same underlying model as [GPT-5.6 Luna](https://openrouter.ai/openai/gpt-5.6-luna), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. Learn more in OpenAI's docs: https://developers.openai.com/api/docs/guides/reasoning#reasoning-mode

openai/gpt-5.6-luna-pro 1.05M context $0.2/M input $1.2/M output

Seed 1.6 is a general-purpose model released by the ByteDance Seed team. It incorporates multimodal capabilities and adaptive deep thinking with a 256K context window.

bytedance-seed/seed-1.6 262.144K context $0.25/M input $2/M output

Seed 1.6 Flash is an ultra-fast multimodal deep thinking model by ByteDance Seed, supporting both text and visual understanding. It features a 256k context window and can generate outputs of...

bytedance-seed/seed-1.6-flash 262.144K context $0.075/M input $0.3/M output

Kimi K2 0905 is the September update of [Kimi K2 0711](moonshotai/kimi-k2). It is a large-scale Mixture-of-Experts (MoE) language model developed by Moonshot AI, featuring 1 trillion total parameters with 32...

moonshotai/kimi-k2-0905 262.144K context $0.6/M input $2.5/M output

No provider description is available for this model yet.

databricks/databricks-glm-5-3-flash 1.04858M context Input not listed Output not listed

Qwen3.6 Flash is a fast, efficient language model from Alibaba's Qwen 3.6 series. It supports text, image, and video input with a 1M token context window. Tiered pricing kicks in...

qwen/qwen3.6-flash 1M context $0.188/M input $1.125/M output

No provider description is available for this model yet.

qwen_ai_platform/qwen-plus-latest 997.952K context Input not listed Output not listed

Inkling is an open-weight multimodal mixture-of-experts model from Thinking Machines Lab, with 41B active parameters out of 975B total. It is designed for general-purpose reasoning, coding, agentic and tool-use systems,...

thinkingmachines/inkling:free 1.04858M context Free input Free output

No provider description is available for this model yet.

novita/qwen/qwen3.6-27b 262.144K context $0.6/M input $3.6/M output

No provider description is available for this model yet.

mistral/ministral-14b-latest 262.144K context $0.2/M input $0.2/M output

No provider description is available for this model yet.

mistral/ministral-14b-2512 262.144K context $0.2/M input $0.2/M output

No provider description is available for this model yet.

qwen_ai_platform/qwen-plus-2025-09-11 997.952K context Input not listed Output not listed

No provider description is available for this model yet.

qwen_ai_platform/qwen-plus-2025-07-28 997.952K context Input not listed Output not listed

No provider description is available for this model yet.

deepinfra/qwen/qwen3-max 256K context $1.2/M input $6/M output

Ox Alpha is a reasoning model designed for coding, sustained agentic work, and production workloads. It is suited for long-horizon software engineering, complex reasoning, and workflows that combine text with...

stealth/ox-alpha 1.04858M context Input not listed Output not listed

Qwen Plus 0728, based on the Qwen3 foundation model, is a 1 million context hybrid reasoning model with a balanced performance, speed, and cost combination.

qwen/qwen-plus-2025-07-28:thinking 1M context $0.26/M input $0.78/M output

No provider description is available for this model yet.

deepinfra/xiaomimimo/mimo-v2.5 262.144K context $0.4/M input $2/M output

A balanced model in the Ministral 3 family, Ministral 3 8B is a powerful, efficient tiny language model with vision capabilities.

mistralai/ministral-8b-2512:batch 262.144K context $0.075/M input $0.075/M output

No provider description is available for this model yet.

qwen_ai_platform/qwen-flash 997.952K context Input not listed Output not listed

No provider description is available for this model yet.

qwen_ai_platform/qwen-coder 1M context $0.3/M input $1.5/M output

No provider description is available for this model yet.

qwen_ai_platform/kimi-k2.7-code 229.376K context $0.95/M input $4/M output

No provider description is available for this model yet.

qwen_ai_platform/glm-5.2 1.04858M context $1.4/M input $4.4/M output

No provider description is available for this model yet.

openai/chat-latest 400K context $5/M input $30/M output

No provider description is available for this model yet.

novita/xiaomimimo/mimo-v2.5-pro 1.04858M context $0.522/M input $1.044/M output

No provider description is available for this model yet.

mistral/glm-5-2 1.04858M context $1.4/M input $4.4/M output