3,229 models

Qwen3-VL-8B-Thinking is the reasoning-optimized variant of the Qwen3-VL-8B multimodal model, designed for advanced visual and textual reasoning across complex scenes, documents, and temporal sequences. It integrates enhanced multimodal alignment and...

qwen/qwen3-vl-8b-thinking 131.072K context $0.18/M input $2.1/M output

Gemma 3n E4B-it is optimized for efficient execution on mobile and low-resource devices, such as phones, laptops, and tablets. It supports multimodal inputs—including text, visual data, and audio—enabling diverse tasks...

google/gemma-3n-e4b-it 32.768K context $0.06/M input $0.12/M output

No provider description is available for this model yet.

dashscope/kimi-k2.7-code 229.376K context $0.95/M input $4/M output

No provider description is available for this model yet.

dashscope/glm-5.2 1.04858M context $1.4/M input $4.4/M output

No provider description is available for this model yet.

openai/gpt-5.6-cyber 400K context $12.5/M input $75/M output

Claude Haiku 4.5 is Anthropic’s fastest and most efficient model, delivering near-frontier intelligence at a fraction of the cost and latency of larger Claude models. Matching Claude Sonnet 4’s performance...

anthropic/claude-haiku-4.5 200K context $1/M input $5/M output

Qwen3-Next-80B-A3B-Instruct is an instruction-tuned chat model in the Qwen3-Next series optimized for fast, stable responses without “thinking” traces. It targets complex tasks across reasoning, code generation, knowledge QA, and multilingual...

qwen/qwen3-next-80b-a3b-instruct 262.144K context $0.09/M input $1.1/M output

Morph's high-accuracy apply model for complex code edits. ~4,500 tokens/sec with 98% accuracy for precise code transformations. The model requires the prompt to be in the following format: <instruction>{instruction}</instruction> <code>{initial_code}</code>...

morph/morph-v3-large 262.144K context $0.9/M input $1.9/M output

Mistral Small 4 is the next major release in the Mistral Small family, unifying the capabilities of several flagship Mistral models into a single system. It combines strong reasoning from...

mistralai/mistral-small-2603 262.144K context $0.15/M input $0.6/M output

This model always redirects to the latest Grok model from xAI.

~x-ai/grok-latest 500K context $2/M input $6/M output

GPT-5.6 Terra Pro is the same underlying model as [GPT-5.6 Terra](https://openrouter.ai/openai/gpt-5.6-terra), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. Learn more in OpenAI's docs: https://developers.openai.com/api/docs/guides/reasoning#reasoning-mode

openai/gpt-5.6-terra-pro 1.05M context $2/M input $12/M output

No provider description is available for this model yet.

dashscope/qwen3-vl-32b-instruct 131.072K context $0.16/M input $0.64/M output

No provider description is available for this model yet.

dashscope/glm-5.1 202.745K context $1.4/M input $4.4/M output

No provider description is available for this model yet.

azure_ai/gpt-6-astra 922K context $10/M input $50/M output

No provider description is available for this model yet.

dashscope/deepseek-v4-pro 1M context $2.4/M input $4.8/M output

No provider description is available for this model yet.

dashscope/qwen3-max-2026-01-23 258.048K context Input not listed Output not listed

No provider description is available for this model yet.

dashscope/qwen3-max 258.048K context Input not listed Output not listed

No provider description is available for this model yet.

dashscope/qwen3-max-preview 258.048K context Input not listed Output not listed

No provider description is available for this model yet.

dashscope/qwen3-coder-plus-2025-07-22 997.952K context Input not listed Output not listed

No provider description is available for this model yet.

friendliai/minimaxai/minimax-m2.5 196.608K context $0.3/M input $1.2/M output

NVIDIA Nemotron Nano 2 VL is a 12-billion-parameter open multimodal reasoning model designed for video understanding and document intelligence. It introduces a hybrid Transformer-Mamba architecture, combining transformer-level accuracy with Mamba’s...

nvidia/nemotron-nano-12b-v2-vl:free 128K context Free input Free output

Hunyuan-A13B is a 13B active parameter Mixture-of-Experts (MoE) language model developed by Tencent, with a total parameter count of 80B and support for reasoning via Chain-of-Thought. It offers competitive benchmark...

tencent/hunyuan-a13b-instruct 131.072K context $0.14/M input $0.57/M output

No provider description is available for this model yet.

dashscope/qwen3-coder-plus 997.952K context Input not listed Output not listed

No provider description is available for this model yet.

dashscope/qwen3-vl-32b-thinking 131.072K context $0.16/M input $2.87/M output

No provider description is available for this model yet.

dashscope/qwen3-vl-plus 260.096K context Input not listed Output not listed

No provider description is available for this model yet.

dashscope/qwen3.5-plus 991.808K context Input not listed Output not listed

No provider description is available for this model yet.

dashscope/qwen3.7-max 991.808K context $2.5/M input $7.5/M output

No provider description is available for this model yet.

dashscope/qwen3.7-plus 991.808K context Input not listed Output not listed

No provider description is available for this model yet.

dashscope/qwq-plus 98.304K context $0.8/M input $2.4/M output