3,646 models

No provider description is available for this model yet.

azure_ai/claude-sonnet-4-5 200K context $3/M input $15/M output

Qwen3.8 2.4T A95B is an open-weight sparse mixture-of-experts model from Qwen and the open-weight variant of [Qwen3.8 Max](/qwen/qwen3.8-max), with 95 billion active parameters out of 2.4 trillion total. It is...

qwen/qwen3.8-2.4t-a95b:batch 1.01M context $2/M input $6/M output

No provider description is available for this model yet.

azure_ai/claude-sonnet-5 1M context $2/M input $10/M output

No provider description is available for this model yet.

azure_ai/claude-sonnet-4-6 1M context $3/M input $15/M output

No provider description is available for this model yet.

azure/computer-use-preview 8.192K context $3/M input $12/M output

No provider description is available for this model yet.

azure/container Not documented context Input not listed Output not listed

No provider description is available for this model yet.

vertex_ai/gemini-3.8-flash 1.04858M context $0.75/M input $3.75/M output

No provider description is available for this model yet.

azure_ai/gpt-oss-120b 131.072K context $0.15/M input $0.6/M output

GPT-3.5 Turbo is OpenAI's fastest model. It can understand and generate natural language or code, and is optimized for chat and traditional completion tasks. Training data up to Sep 2021.

openai/gpt-3.5-turbo:batch 16.385K context $0.25/M input $0.75/M output

MoonshotAI: Kimi K2.7 Code is a coding-focused model in Moonshot AI's Kimi K2 family, built to complete end-to-end programming tasks reliably over long contexts. It uses a native multimodal mixture-of-experts...

moonshotai/kimi-k2.7-code:batch 262.144K context $0.95/M input $4/M output

No provider description is available for this model yet.

azure_ai/gpt-5.5 1.05M context $5/M input $30/M output

No provider description is available for this model yet.

gemini/gemini-3.8-flash 1.04858M context $0.75/M input $3.75/M output

DeepSeek V4 Flash Vision Exp is an experimental vision-enabled version of [DeepSeek V4 Flash 0731](https://openrouter.ai/deepseek/deepseek-v4-flash-0731) from DeepSeek, adding image understanding while matching the base model on text capabilities including agents,...

deepseek/deepseek-v4-flash-vision-exp:batch 1.04858M context $0.11/M input $0.33/M output

No provider description is available for this model yet.

azure_ai/gpt-5.5-2026-04-23 1.05M context $5/M input $30/M output

GPT-5.1 is the latest frontier-grade model in the GPT-5 series, offering stronger general-purpose reasoning, improved instruction adherence, and a more natural conversational style compared to GPT-5. It uses adaptive reasoning...

openai/gpt-5.1:batch 400K context $0.625/M input $5/M output

Gemini 3.8 Flash is Google's most intelligent Flash model with significant gains from 3.7 Flash across software engineering, agentic tasks, and multi-step reasoning.

google/gemini-3.8-flash:batch 1.04858M context $0.375/M input $1.875/M output

No provider description is available for this model yet.

xai/grok-build-latest 500K context $2/M input $6/M output

No provider description is available for this model yet.

scaleway/glm-5.2 256K context $1.8/M input $5.5/M output

The largest model in the Ministral 3 family, Ministral 3 14B offers frontier capabilities and performance comparable to its larger Mistral Small 3.2 24B counterpart. A powerful and efficient language...

mistralai/ministral-14b-2512 262.144K context $0.2/M input $0.2/M output

No provider description is available for this model yet.

scaleway/deepseek-v4-flash-0731 256K context $0.4/M input $0.8/M output

No provider description is available for this model yet.

azure_ai/kimi-k2.7-code 262.144K context $0.95/M input $4/M output

Qwen3-VL-30B-A3B-Thinking is a multimodal model that unifies strong text generation with visual understanding for images and videos. Its Thinking variant enhances reasoning in STEM, math, and complex tasks. It excels...

qwen/qwen3-vl-30b-a3b-thinking 131.072K context $0.2/M input $2.4/M output