3,455 models

No provider description is available for this model yet.

qwencloud/qwen3-max-2026-01-23 258.048K context Input not listed Output not listed

Claude Opus 4.5 is Anthropic’s frontier reasoning model optimized for complex software engineering, agentic workflows, and long-horizon computer use. It offers strong multimodal capabilities, competitive performance across real-world coding and...

anthropic/claude-opus-4.5:batch 200K context $2.5/M input $12.5/M output

No provider description is available for this model yet.

qwencloud/qwen3-vl-32b-instruct 131.072K context $0.16/M input $0.64/M output

No provider description is available for this model yet.

qwencloud/qwen3-vl-32b-thinking 131.072K context $0.16/M input $2.87/M output

No provider description is available for this model yet.

azure_ai/grok-4-1-fast-reasoning 131.072K context $0.2/M input $0.5/M output

No provider description is available for this model yet.

qwencloud/qwen3.5-plus 991.808K context Input not listed Output not listed

No provider description is available for this model yet.

qwencloud/qwen3.7-max 991.808K context $2.5/M input $7.5/M output

No provider description is available for this model yet.

qwencloud/qwen3.7-plus 991.808K context Input not listed Output not listed

Qwen3.8 Flash is a multimodal reasoning model from Alibaba. It is suited for coding assistance, agentic workflows, visual understanding, document and codebase analysis, desktop interaction, chart analysis, and long-video analysis.

qwen/qwen3.8-flash 1M context $0.15/M input $0.47/M output

No provider description is available for this model yet.

qwencloud/qwen3.8-max 991.808K context $2/M input $6/M output

No provider description is available for this model yet.

qwencloud/qwq-plus 98.304K context $0.8/M input $2.4/M output

Qwen3-Coder-480B-A35B-Instruct is a Mixture-of-Experts (MoE) code generation model developed by the Qwen team. It is optimized for agentic coding tasks such as function calling, tool use, and long-context reasoning over...

qwen/qwen3-coder 262.144K context $0.3/M input $1/M output

NVIDIA Nemotron 3 Nano 30B A3B is a small language MoE model with highest compute efficiency and accuracy for developers to build specialized agentic AI systems. The model is fully...

nvidia/nemotron-3-nano-30b-a3b:free 256K context Free input Free output

No provider description is available for this model yet.

qwen_ai_platform/deepseek-v4-pro 1M context $2.4/M input $4.8/M output

No provider description is available for this model yet.

qwen_ai_platform/glm-5.1 202.745K context $1.4/M input $4.4/M output

No provider description is available for this model yet.

qwen_ai_platform/kimi-k2.7-code 229.376K context $0.95/M input $4/M output

No provider description is available for this model yet.

qwen_ai_platform/qwen-coder 1M context $0.3/M input $1.5/M output

No provider description is available for this model yet.

qwen_ai_platform/qwen-flash 997.952K context Input not listed Output not listed

No provider description is available for this model yet.

qwen_ai_platform/qwen-max 30.72K context $1.6/M input $6.4/M output

No provider description is available for this model yet.

qwen_ai_platform/qwen-plus 129.024K context $0.4/M input $1.2/M output

Fugu Ultra v2 is the higher-performance model in Sakana AI's Fugu family. Rather than a single monolithic model, Fugu is a learned multi-agent orchestration system: a language model trained to...

sakana/fugu-ultra-v2 1M context $5/M input $30/M output

GLM-4.5V is a vision-language foundation model for multimodal agent applications. Built on a Mixture-of-Experts (MoE) architecture with 106B parameters and 12B activated parameters, it achieves state-of-the-art results in video understanding,...

z-ai/glm-4.5v 65.536K context $0.6/M input $1.8/M output

No provider description is available for this model yet.

qwen_ai_platform/qwen-plus-2025-07-28 997.952K context Input not listed Output not listed

No provider description is available for this model yet.

qwen_ai_platform/qwen-plus-2025-09-11 997.952K context Input not listed Output not listed

No provider description is available for this model yet.

qwen_ai_platform/qwen-plus-latest 997.952K context Input not listed Output not listed

No provider description is available for this model yet.

qwen_ai_platform/qwen-turbo 129.024K context $0.05/M input $0.2/M output

Kimi K2 Instruct is a large-scale Mixture-of-Experts (MoE) language model developed by Moonshot AI, featuring 1 trillion total parameters with 32 billion active per forward pass. It is optimized for...

moonshotai/kimi-k2 131.072K context $0.57/M input $2.3/M output

No provider description is available for this model yet.

qwen_ai_platform/qwen3-30b-a3b 129.024K context Input not listed Output not listed

No provider description is available for this model yet.

qwen_ai_platform/qwen3-coder-flash 997.952K context Input not listed Output not listed

No provider description is available for this model yet.

vertex_ai-ai21_models/jamba-1.5 256K context $0.2/M input $0.4/M output