4,182 models

Claude Opus 4.8 is Anthropic's most capable generally available model in the Opus family. It supports text, image, and file inputs with text output, with reasoning support and a 1M-token...

anthropic/claude-opus-4.8:batch 1M context $2.5/M input $12.5/M output

No provider description is available for this model yet.

baseten/deepseek-ai/deepseek-v3.1 Not documented context $0.5/M input $1.5/M output

No provider description is available for this model yet.

baseten/deepseek-ai/deepseek-v3-0324 Not documented context $0.77/M input $0.77/M output

A cost-efficient version of GPT Audio. The new snapshot features an upgraded decoder for more natural sounding voices and maintains better voice consistency. Input is priced at $0.60 per million...

openai/gpt-audio-mini 128K context $0.6/M input $2.4/M output

No provider description is available for this model yet.

gmi/zai-org/glm-4.7-fp8 202.752K context $0.4/M input $2/M output

Ling 3.0 Flash Fin is a finance-focused mixture-of-experts model from InclusionAI, built on Ling 3.0 Flash with 5.1B active parameters out of 124B total. It is designed for real-world investment...

inclusionai/ling-3.0-flash-fin:free 262.144K context Free input Free output

No provider description is available for this model yet.

azure/gpt-audio-mini 128K context $0.6/M input $2.4/M output
Open weights

No provider description is available for this model yet.

google/gemma-4-12B Not documented context Input not listed Output not listed
Open weights

No provider description is available for this model yet.

microsoft/Mage-Flow-Base Not documented context Input not listed Output not listed

No provider description is available for this model yet.

scx-ai/glm-5.2 1.04858M context $0.61/M input $1.98/M output

No provider description is available for this model yet.

scx-ai/qwen3.8-max 1M context $1.65/M input $4.99/M output
Open weights

No provider description is available for this model yet.

google/gemma-4-12B-it Not documented context Input not listed Output not listed

GLM-4.5-Air is the lightweight variant of our latest flagship model family, also purpose-built for agent-centric applications. Like GLM-4.5, it adopts the Mixture-of-Experts (MoE) architecture but with a more compact parameter...

z-ai/glm-4.5-air 131.072K context $0.13/M input $0.85/M output

No provider description is available for this model yet.

together_ai/together-ai-8.1b-21b 1K context $0.3/M input $0.3/M output

No provider description is available for this model yet.

together_ai/together-ai-81.1b-110b Not documented context $1.8/M input $1.8/M output

No provider description is available for this model yet.

together_ai/together-ai-up-to-4b Not documented context $0.1/M input $0.1/M output

Qwen3-Coder-480B-A35B-Instruct is a Mixture-of-Experts (MoE) code generation model developed by the Qwen team. It is optimized for agentic coding tasks such as function calling, tool use, and long-context reasoning over...

qwen/qwen3-coder:free 262K context Free input Free output
Open weights

No provider description is available for this model yet.

google/gemma-4-12B-it-assistant Not documented context Input not listed Output not listed

GPT-6 Astra Pro is the same underlying model as [GPT-6 Astra](https://openrouter.ai/openai/gpt-6-astra), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. Learn more in OpenAI's docs: https://developers.openai.com/api/docs/guides/reasoning#reasoning-mode

openai/gpt-6-astra-pro:batch 1.05M context $5/M input $25/M output

No provider description is available for this model yet.

dashscope/deepseek-v4-flash 1M context $0.2/M input $0.4/M output

No provider description is available for this model yet.

dashscope/deepseek-v4-flash-0731 1M context $0.2/M input $0.4/M output

Qwen3.8 Max 0902 is an updated snapshot of Qwen3.8 Max from Alibaba's Qwen team. It is a 2.4-trillion-parameter mixture-of-experts model that accepts text, image, and video input and returns text,...

qwen/qwen3.8-max-0902 1M context $2/M input $6/M output
Open weights

No provider description is available for this model yet.

google/gemma-4-31B-it-assistant Not documented context Input not listed Output not listed