4,213 models

No provider description is available for this model yet.

qwencloud/qwen3.5-plus 991.808K context Input not listed Output not listed

No provider description is available for this model yet.

qwencloud/qwen3.7-max 991.808K context $2.5/M input $7.5/M output

No provider description is available for this model yet.

qwencloud/qwen3.7-plus 991.808K context Input not listed Output not listed

Qwen3.8 Flash is a multimodal reasoning model from Alibaba. It is suited for coding assistance, agentic workflows, visual understanding, document and codebase analysis, desktop interaction, chart analysis, and long-video analysis.

qwen/qwen3.8-flash 1M context $0.15/M input $0.47/M output

Qwen3-Coder-480B-A35B-Instruct is a Mixture-of-Experts (MoE) code generation model developed by the Qwen team. It is optimized for agentic coding tasks such as function calling, tool use, and long-context reasoning over...

qwen/qwen3-coder 262.144K context $0.3/M input $1/M output

No provider description is available for this model yet.

qwencloud/qwen3.8-max 991.808K context $2/M input $6/M output

No provider description is available for this model yet.

qwencloud/qwq-plus 98.304K context $0.8/M input $2.4/M output

No provider description is available for this model yet.

qwen_ai_platform/deepseek-v4-pro 1M context $2.4/M input $4.8/M output

No provider description is available for this model yet.

qwen_ai_platform/glm-5.1 202.745K context $1.4/M input $4.4/M output

No provider description is available for this model yet.

qwen_ai_platform/glm-5.2 1.04858M context $1.4/M input $4.4/M output

No provider description is available for this model yet.

qwen_ai_platform/kimi-k2.7-code 229.376K context $0.95/M input $4/M output

No provider description is available for this model yet.

qwen_ai_platform/qwen-coder 1M context $0.3/M input $1.5/M output

No provider description is available for this model yet.

qwen_ai_platform/qwen-flash 997.952K context Input not listed Output not listed

No provider description is available for this model yet.

qwen_ai_platform/qwen-max 30.72K context $1.6/M input $6.4/M output

Fugu Ultra v2 is the higher-performance model in Sakana AI's Fugu family. Rather than a single monolithic model, Fugu is a learned multi-agent orchestration system: a language model trained to...

sakana/fugu-ultra-v2 1M context $5/M input $30/M output

No provider description is available for this model yet.

qwen_ai_platform/qwen-plus 129.024K context $0.4/M input $1.2/M output

No provider description is available for this model yet.

qwen_ai_platform/qwen-plus-2025-07-28 997.952K context Input not listed Output not listed

No provider description is available for this model yet.

qwen_ai_platform/qwen-plus-2025-09-11 997.952K context Input not listed Output not listed

No provider description is available for this model yet.

qwen_ai_platform/qwen-plus-latest 997.952K context Input not listed Output not listed

No provider description is available for this model yet.

qwen_ai_platform/qwen-turbo 129.024K context $0.05/M input $0.2/M output

No provider description is available for this model yet.

qwen_ai_platform/qwen3-30b-a3b 129.024K context Input not listed Output not listed

OpenAI o3-mini is a cost-efficient language model optimized for STEM reasoning tasks, particularly excelling in science, mathematics, and coding. This model supports the `reasoning_effort` parameter, which can be set to...

openai/o3-mini:batch 200K context $0.55/M input $2.2/M output

No provider description is available for this model yet.

qwen_ai_platform/qwen3-coder-flash 997.952K context Input not listed Output not listed

GPT-5.5 is OpenAI’s frontier model designed for complex professional workloads, building on GPT-5.4 with stronger reasoning, higher reliability, and improved token efficiency on hard tasks. It features a 1M+ token...

openai/gpt-5.5:batch 1.05M context $2.5/M input $15/M output

No provider description is available for this model yet.

qwen_ai_platform/qwen3-coder-plus 997.952K context Input not listed Output not listed

MiniMax-M2.7 is a next-generation large language model designed for autonomous, real-world productivity and continuous improvement. Built to actively participate in its own evolution, M2.7 integrates advanced agentic capabilities through multi-agent...

minimax/minimax-m2.7 204.8K context $0.3/M input $1.2/M output

No provider description is available for this model yet.

qwen_ai_platform/qwen3-max-preview 258.048K context Input not listed Output not listed

No provider description is available for this model yet.

qwen_ai_platform/qwen3-max 258.048K context Input not listed Output not listed

GLM-5.3 is a large-scale reasoning model from Z.ai, built for complex software engineering and long-horizon agent tasks. It supports text input and output with a 1M-token context window, and improves...

z-ai/glm-5.3 1.04858M context $1.4/M input $4.4/M output

No provider description is available for this model yet.

qwen_ai_platform/qwen3-max-2026-01-23 258.048K context Input not listed Output not listed

Gemma 4 31B Instruct is Google DeepMind's 30.7B dense multimodal model supporting text and image input with text output. Features a 256K token context window, configurable thinking/reasoning mode, native function...

google/gemma-4-31b-it:batch 262.144K context $0.39/M input $0.97/M output

A balanced model in the Ministral 3 family, Ministral 3 8B is a powerful, efficient tiny language model with vision capabilities.

mistralai/ministral-8b-2512:batch 262.144K context $0.075/M input $0.075/M output

GLM 5.2 is a large-scale reasoning model from Z.ai. It supports text input and output with a 1M-token context window, and is suited for long-horizon agent workflows, project-level software engineering,...

z-ai/glm-5.2:batch 1.04858M context $0.7/M input $2.2/M output