3,424 models

No provider description is available for this model yet.

bedrock_mantle/xai.grok-4.6 500K context $2.2/M input $6.6/M output

Seed 2.1 Turbo is a multimodal model from ByteDance Seed for coding and long-horizon agent workflows. It is suited for end-to-end software delivery, multi-step task execution, and understanding visual and...

bytedance-seed/seed-2-1-turbo 262.144K context $0.5/M input $2.5/M output

No provider description is available for this model yet.

azure_ai/gpt-5.4-mini 272K context $0.75/M input $4.5/M output

No provider description is available for this model yet.

azure_ai/gpt-5.4-mini-2026-03-17 272K context $0.75/M input $4.5/M output

Qwen3-Coder-30B-A3B-Instruct is a 30.5B parameter Mixture-of-Experts (MoE) model with 128 experts (8 active per forward pass), designed for advanced code generation, repository-scale understanding, and agentic tool use. Built on the...

qwen/qwen3-coder-30b-a3b-instruct 262.144K context $0.07/M input $0.28/M output

No provider description is available for this model yet.

azure_ai/gpt-5.4-nano-2026-03-17 272K context $0.2/M input $1.25/M output

No provider description is available for this model yet.

azure/eu/gpt-4o-2024-08-06 128K context $2.75/M input $11/M output

No provider description is available for this model yet.

azure/eu/gpt-4o-2024-11-20 128K context $2.75/M input $11/M output

No provider description is available for this model yet.

azure/eu/gpt-4o-mini-2024-07-18 128K context $0.165/M input $0.66/M output

No provider description is available for this model yet.

azure/eu/gpt-5-2025-08-07 272K context $1.375/M input $11/M output

No provider description is available for this model yet.

azure/eu/gpt-5-mini-2025-08-07 272K context $0.275/M input $2.2/M output

No provider description is available for this model yet.

azure/eu/gpt-5.1 272K context $1.38/M input $11/M output

No provider description is available for this model yet.

azure/eu/gpt-5.1-chat 128K context $1.38/M input $11/M output

No provider description is available for this model yet.

azure/eu/gpt-5-nano-2025-08-07 272K context $0.055/M input $0.44/M output

No provider description is available for this model yet.

azure/eu/o1-2024-12-17 200K context $16.5/M input $66/M output

No provider description is available for this model yet.

azure/eu/o1-preview-2024-09-12 128K context $16.5/M input $66/M output

Qwen3.8 27B is an open-weight dense vision-language model from Qwen. It is suited for coding, professional workflows, research, multimodal interaction, and long-running agent tasks, with flexible thinking that can be...

qwen/qwen3.8-27b 262.144K context $0.214/M input $2.55/M output

No provider description is available for this model yet.

azure/us-gov/gpt-5.1 272K context $1.719/M input $13.75/M output

No provider description is available for this model yet.

azure/us-gov/o3-mini 200K context $1.513/M input $6.05/M output

GLM 5.2 is a large-scale reasoning model from Z.ai. It supports text input and output with a 1M-token context window, and is suited for long-horizon agent workflows, project-level software engineering,...

z-ai/glm-5.2:free 256K context Free input Free output

No provider description is available for this model yet.

openrouter/deepseek/deepseek-v4-pro 1.04858M context $1.32/M input $3.96/M output

Qwen3-VL-8B-Thinking is the reasoning-optimized variant of the Qwen3-VL-8B multimodal model, designed for advanced visual and textual reasoning across complex scenes, documents, and temporal sequences. It integrates enhanced multimodal alignment and...

qwen/qwen3-vl-8b-thinking 131.072K context $0.18/M input $2.1/M output

No provider description is available for this model yet.

azure/global-standard/gpt-4o-mini 128K context $0.15/M input $0.6/M output

No provider description is available for this model yet.

azure/global/gpt-4o-2024-08-06 128K context $2.5/M input $10/M output

No provider description is available for this model yet.

azure/global/gpt-4o-2024-11-20 128K context $2.5/M input $10/M output

No provider description is available for this model yet.

azure/global/gpt-5.1 272K context $1.25/M input $10/M output

Hunyuan-A13B is a 13B active parameter Mixture-of-Experts (MoE) language model developed by Tencent, with a total parameter count of 80B and support for reasoning via Chain-of-Thought. It offers competitive benchmark...

tencent/hunyuan-a13b-instruct 131.072K context $0.14/M input $0.57/M output

No provider description is available for this model yet.

azure/global/gpt-5.1-chat 128K context $1.25/M input $10/M output

The Meta Llama 3.3 multilingual large language model (LLM) is a pretrained and instruction tuned generative model in 70B (text in/text out). The Llama 3.3 instruction tuned text only model...

meta-llama/llama-3.3-70b-instruct:free 65.536K context Free input Free output

No provider description is available for this model yet.

azure/gpt-3.5-turbo-0125 16.384K context $0.5/M input $1.5/M output

No provider description is available for this model yet.

azure/gpt-35-turbo-0125 16.384K context $0.5/M input $1.5/M output

Gemini 2.5 Flash-Lite is a lightweight reasoning model in the Gemini 2.5 family, optimized for ultra-low latency and cost efficiency. It offers improved throughput, faster token generation, and better performance...

google/gemini-2.5-flash-lite:batch 1.04858M context $0.05/M input $0.2/M output