4,182 models

Qwen3.8 Max 0902 is an updated snapshot of Qwen3.8 Max from Alibaba's Qwen team. It is a 2.4-trillion-parameter mixture-of-experts model that accepts text, image, and video input and returns text,...

qwen/qwen3.8-max-0902 1M context $2/M input $6/M output

No provider description is available for this model yet.

dashscope/deepseek-v4-flash 1M context $0.2/M input $0.4/M output

No provider description is available for this model yet.

dashscope/deepseek-v4-flash-0731 1M context $0.2/M input $0.4/M output
Open weights

No provider description is available for this model yet.

google/gemma-4-31B-it-assistant Not documented context Input not listed Output not listed

No provider description is available for this model yet.

dashscope/deepseek-v4-pro 1M context $2.4/M input $4.8/M output

GPT-5.6 Terra Pro is the same underlying model as [GPT-5.6 Terra](https://openrouter.ai/openai/gpt-5.6-terra), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. Learn more in OpenAI's docs: https://developers.openai.com/api/docs/guides/reasoning#reasoning-mode

openai/gpt-5.6-terra-pro:batch 1.05M context $1/M input $6/M output
Open weights

No provider description is available for this model yet.

meta-llama/CodeLlama-7b-Python-hf Not documented context Input not listed Output not listed

No provider description is available for this model yet.

dashscope/glm-5.2 1.04858M context $1.4/M input $4.4/M output

No provider description is available for this model yet.

dashscope/kimi-k2.7-code 229.376K context $0.95/M input $4/M output

Hy-MT2-1.8B is a compact 1.8B-parameter translation model from Tencent. It supports 33 language pairs and five Chinese dialect and minority-language pairs, with workflows for structured, delimiter-based, contextual, glossary-based, and style-guided...

tencent/hy-mt2-1.8b 8.192K context $0.044/M input $0.177/M output

No provider description is available for this model yet.

dashscope/qwen3.8-max 991.808K context $2/M input $6/M output

No provider description is available for this model yet.

vertex_ai/gemini-3.7-flash 1.04858M context $0.75/M input $3.75/M output
Open weights

No provider description is available for this model yet.

meta-llama/CodeLlama-7b-Instruct-hf Not documented context Input not listed Output not listed

Qwen Plus 0728, based on the Qwen3 foundation model, is a 1 million context hybrid reasoning model with a balanced performance, speed, and cost combination.

qwen/qwen-plus-2025-07-28 1M context $0.26/M input $0.78/M output

No provider description is available for this model yet.

gemini/gemini-3.7-flash 1.04858M context $0.75/M input $3.75/M output
Open weights

No provider description is available for this model yet.

microsoft/UniRG-CXR Not documented context Input not listed Output not listed

Qwen3-Next-80B-A3B-Instruct is an instruction-tuned chat model in the Qwen3-Next series optimized for fast, stable responses without “thinking” traces. It targets complex tasks across reasoning, code generation, knowledge QA, and multilingual...

qwen/qwen3-next-80b-a3b-instruct 262.144K context $0.09/M input $1.1/M output

Morph's high-accuracy apply model for complex code edits. ~4,500 tokens/sec with 98% accuracy for precise code transformations. The model requires the prompt to be in the following format: <instruction>{instruction}</instruction> <code>{initial_code}</code>...

morph/morph-v3-large 262.144K context $0.9/M input $1.9/M output

Coder‑Large is a 32 B‑parameter offspring of Qwen 2.5‑Instruct that has been further trained on permissively‑licensed GitHub, CodeSearchNet and synthetic bug‑fix corpora. It supports a 32k context window, enabling multi‑file...

arcee-ai/coder-large 32.768K context $0.5/M input $0.8/M output

Schematron V2 Small is a 3B-parameter HTML-to-JSON extraction model from Inference.net. It prioritizes extraction quality for complex schemas and long pages. Extraction instructions must be supplied through a JSON schema...

inference-net/schematron-v2-small 128K context $0.05/M input $0.23/M output

No provider description is available for this model yet.

openai/daybreak-blue-latest 1.05M context $5/M input $30/M output

No provider description is available for this model yet.

groq/qwen/qwen3.6-27b 131.072K context $0.6/M input $3/M output