2,852 models

No provider description is available for this model yet.

azure_ai/claude-haiku-4-5 200K context $1/M input $5/M output

No provider description is available for this model yet.

azure_ai/claude-opus-4-5 200K context $5/M input $25/M output

No provider description is available for this model yet.

azure_ai/claude-opus-4-6 1M context $5/M input $25/M output

No provider description is available for this model yet.

azure_ai/claude-opus-4-7 1M context $5/M input $25/M output

No provider description is available for this model yet.

openrouter/openai/gpt-6-astra-pro 1.05M context $10/M input $50/M output

No provider description is available for this model yet.

azure_ai/claude-fable-5 1M context $10/M input $50/M output

No provider description is available for this model yet.

azure_ai/claude-opus-5 1M context $5/M input $25/M output

No provider description is available for this model yet.

azure_ai/claude-opus-4-8 1M context $5/M input $25/M output

No provider description is available for this model yet.

azure_ai/claude-opus-4-1 200K context $15/M input $75/M output

Sakana Namazu is a Japanese-specialized reasoning model from Sakana AI, based on Kimi K2.6 with additional training for Japanese language and business contexts. It is suited for Japanese instruction following,...

sakana/namazu 262.144K context $0.95/M input $4/M output

No provider description is available for this model yet.

azure_ai/claude-sonnet-4-5 200K context $3/M input $15/M output

No provider description is available for this model yet.

azure_ai/claude-sonnet-5 1M context $2/M input $10/M output

No provider description is available for this model yet.

azure_ai/claude-sonnet-4-6 1M context $3/M input $15/M output

Qwen3.8 2.4T A95B is an open-weight sparse mixture-of-experts model from Qwen and the open-weight variant of [Qwen3.8 Max](/qwen/qwen3.8-max), with 95 billion active parameters out of 2.4 trillion total. It is...

qwen/qwen3.8-2.4t-a95b:batch 1.01M context $2/M input $6/M output

No provider description is available for this model yet.

vertex_ai/gemini-3.8-flash 1.04858M context $0.75/M input $3.75/M output

No provider description is available for this model yet.

azure_ai/gpt-oss-120b 131.072K context $0.15/M input $0.6/M output

MoonshotAI: Kimi K2.7 Code is a coding-focused model in Moonshot AI's Kimi K2 family, built to complete end-to-end programming tasks reliably over long contexts. It uses a native multimodal mixture-of-experts...

moonshotai/kimi-k2.7-code:batch 262.144K context $0.95/M input $4/M output

No provider description is available for this model yet.

azure_ai/gpt-5.5 1.05M context $5/M input $30/M output

No provider description is available for this model yet.

gemini/gemini-3.8-flash 1.04858M context $0.75/M input $3.75/M output

DeepSeek V4 Flash Vision Exp is an experimental vision-enabled version of [DeepSeek V4 Flash 0731](https://openrouter.ai/deepseek/deepseek-v4-flash-0731) from DeepSeek, adding image understanding while matching the base model on text capabilities including agents,...

deepseek/deepseek-v4-flash-vision-exp:batch 1.04858M context $0.11/M input $0.33/M output

No provider description is available for this model yet.

azure_ai/gpt-5.5-2026-04-23 1.05M context $5/M input $30/M output

No provider description is available for this model yet.

azure_ai/gpt-5.4 1.05M context $2.5/M input $15/M output

No provider description is available for this model yet.

xai/grok-build-latest 500K context $2/M input $6/M output

No provider description is available for this model yet.

scaleway/glm-5.2 256K context $1.8/M input $5.5/M output

No provider description is available for this model yet.

scaleway/deepseek-v4-flash-0731 256K context $0.4/M input $0.8/M output

No provider description is available for this model yet.

azure_ai/kimi-k2.7-code 262.144K context $0.95/M input $4/M output

GPT-5.2 Pro is OpenAI’s most advanced model, offering major improvements in agentic coding and long context performance over GPT-5 Pro. It is optimized for complex tasks that require step-by-step reasoning,...

openai/gpt-5.2-pro:batch 400K context $10.5/M input $84/M output

Gemini 2.5 Flash-Lite is a lightweight reasoning model in the Gemini 2.5 family, optimized for ultra-low latency and cost efficiency. It offers improved throughput, faster token generation, and better performance...

google/gemini-2.5-flash-lite:batch 1.04858M context $0.05/M input $0.2/M output

Llama 3.2 3B is a 3-billion-parameter multilingual large language model, optimized for advanced natural language processing tasks like dialogue generation, reasoning, and summarization. Designed with the latest transformer architecture, it...

meta-llama/llama-3.2-3b-instruct:free 131.072K context Free input Free output

No provider description is available for this model yet.

azure_ai/gpt-5.4-2026-03-05 1.05M context $2.5/M input $15/M output

Seed 2.1 Turbo is a multimodal model from ByteDance Seed for coding and long-horizon agent workflows. It is suited for end-to-end software delivery, multi-step task execution, and understanding visual and...

bytedance-seed/seed-2-1-turbo 262.144K context $0.5/M input $2.5/M output