No provider description is available for this model yet.

qwencloud/qwen-turbo-2025-04-28 1M context $0.05/M input $0.2/M output

Kimi K3 is a 2.8T parameter open-weight multimodal reasoning model from Moonshot AI. It is suited for complex coding, knowledge work, and long-horizon agentic workflows, and is particularly strong at...

moonshotai/kimi-k3:batch 1.04858M context $3/M input $15/M output

No provider description is available for this model yet.

qwencloud/qwen-turbo-latest 1M context $0.05/M input $0.2/M output

No provider description is available for this model yet.

qwencloud/qwen3-coder-flash 997.952K context Input not listed Output not listed

This model always redirects to the latest GLM model from Z.ai.

~z-ai/glm-latest 262.144K context $0.936/M input $3.168/M output

No provider description is available for this model yet.

qwencloud/qwen3-coder-plus 997.952K context Input not listed Output not listed

The o-series of models are trained with reinforcement learning to think before they answer and perform complex reasoning. The o3-pro model uses more compute to think harder and provide consistently...

openai/o3-pro:batch 200K context $10/M input $40/M output

No provider description is available for this model yet.

qwencloud/qwen3-coder-plus-2025-07-22 997.952K context Input not listed Output not listed

No provider description is available for this model yet.

qwencloud/qwen3-max-preview 258.048K context Input not listed Output not listed

No provider description is available for this model yet.

qwencloud/qwen3-max 258.048K context Input not listed Output not listed

No provider description is available for this model yet.

qwencloud/qwen3-max-2026-01-23 258.048K context Input not listed Output not listed

GLM-5.3-Flash is a native multimodal model from Z.ai. It is suited for efficient coding and long-horizon agent tasks. Its hybrid sparse and linear attention architecture maintains accurate long-context behavior while...

z-ai/glm-5.3-flash:batch 1.04858M context $0.075/M input $0.25/M output

No provider description is available for this model yet.

qwencloud/qwen3-vl-plus 260.096K context Input not listed Output not listed

No provider description is available for this model yet.

qwencloud/qwen3.5-plus 991.808K context Input not listed Output not listed

No provider description is available for this model yet.

qwencloud/qwen3.7-max 991.808K context $2.5/M input $7.5/M output

No provider description is available for this model yet.

qwencloud/qwen3.7-plus 991.808K context Input not listed Output not listed

No provider description is available for this model yet.

qwencloud/qwen3.8-max 991.808K context $2/M input $6/M output

Qwen3.8 Flash is a multimodal reasoning model from Alibaba. It is suited for coding assistance, agentic workflows, visual understanding, document and codebase analysis, desktop interaction, chart analysis, and long-video analysis.

qwen/qwen3.8-flash 1M context $0.15/M input $0.47/M output

Fugu Ultra v2 is the higher-performance model in Sakana AI's Fugu family. Rather than a single monolithic model, Fugu is a learned multi-agent orchestration system: a language model trained to...

sakana/fugu-ultra-v2 1M context $5/M input $30/M output

No provider description is available for this model yet.

qwen_ai_platform/deepseek-v4-pro 1M context $2.4/M input $4.8/M output

No provider description is available for this model yet.

qwen_ai_platform/glm-5.1 202.745K context $1.4/M input $4.4/M output

No provider description is available for this model yet.

qwen_ai_platform/glm-5.2 1.04858M context $1.4/M input $4.4/M output

No provider description is available for this model yet.

qwen_ai_platform/kimi-k2.7-code 229.376K context $0.95/M input $4/M output

No provider description is available for this model yet.

qwen_ai_platform/qwen-coder 1M context $0.3/M input $1.5/M output

No provider description is available for this model yet.

qwen_ai_platform/qwen-flash 997.952K context Input not listed Output not listed

Grok 4.20 Multi-Agent is a variant of SpaceXAI’s Grok 4.20 designed for collaborative, agent-based workflows. Multiple agents operate in parallel to conduct deep research, coordinate tool use, and synthesize information...

x-ai/grok-4.20-multi-agent 2M context $1.25/M input $2.5/M output

No provider description is available for this model yet.

qwen_ai_platform/qwen-plus-2025-07-28 997.952K context Input not listed Output not listed

No provider description is available for this model yet.

qwen_ai_platform/qwen-plus-2025-09-11 997.952K context Input not listed Output not listed

No provider description is available for this model yet.

qwen_ai_platform/qwen-plus-latest 997.952K context Input not listed Output not listed

No provider description is available for this model yet.

qwen_ai_platform/qwen3-coder-flash 997.952K context Input not listed Output not listed

No provider description is available for this model yet.

qwen_ai_platform/qwen3-coder-plus 997.952K context Input not listed Output not listed

Qwen 3.6 Plus builds on a hybrid architecture that combines efficient linear attention with sparse mixture-of-experts routing, enabling strong scalability and high-performance inference. Compared to the 3.5 series, it delivers...

qwen/qwen3.6-plus 1M context $0.325/M input $1.95/M output

No provider description is available for this model yet.

qwen_ai_platform/qwen3-max-preview 258.048K context Input not listed Output not listed

GPT-6 Astra is OpenAI's flagship model for demanding end-to-end work. It is suited for advanced analysis, software engineering, deep research, scientific work, and document creation, with particular strengths in long-horizon...

openai/gpt-6-astra:batch 1.05M context $5/M input $25/M output

Claude Sonnet 4.5 is Anthropic’s most advanced Sonnet model to date, optimized for real-world agents and coding workflows. It delivers state-of-the-art performance on coding benchmarks such as SWE-bench Verified, with...

anthropic/claude-sonnet-4.5 1M context $3/M input $15/M output

GLM 5.2 is a large-scale reasoning model from Z.ai. It supports text input and output with a 1M-token context window, and is suited for long-horizon agent workflows, project-level software engineering,...

z-ai/glm-5.2:batch 1.04858M context $0.7/M input $2.2/M output

No provider description is available for this model yet.

qwen_ai_platform/qwen3-max 258.048K context Input not listed Output not listed

No provider description is available for this model yet.

qwen_ai_platform/qwen3-max-2026-01-23 258.048K context Input not listed Output not listed

MiniMax-M2.7 is a next-generation large language model designed for autonomous, real-world productivity and continuous improvement. Built to actively participate in its own evolution, M2.7 integrates advanced agentic capabilities through multi-agent...

minimax/minimax-m2.7 204.8K context $0.3/M input $1.2/M output