No provider description is available for this model yet.
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
No provider description is available for this model yet.
No provider description is available for this model yet.
Qwen3.8 Flash is a multimodal reasoning model from Alibaba. It is suited for coding assistance, agentic workflows, visual understanding, document and codebase analysis, desktop interaction, chart analysis, and long-video analysis.
Qwen3-Coder-480B-A35B-Instruct is a Mixture-of-Experts (MoE) code generation model developed by the Qwen team. It is optimized for agentic coding tasks such as function calling, tool use, and long-context reasoning over...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Fugu Ultra v2 is the higher-performance model in Sakana AI's Fugu family. Rather than a single monolithic model, Fugu is a learned multi-agent orchestration system: a language model trained to...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
OpenAI o3-mini is a cost-efficient language model optimized for STEM reasoning tasks, particularly excelling in science, mathematics, and coding. This model supports the `reasoning_effort` parameter, which can be set to...
No provider description is available for this model yet.
GPT-5.5 is OpenAI’s frontier model designed for complex professional workloads, building on GPT-5.4 with stronger reasoning, higher reliability, and improved token efficiency on hard tasks. It features a 1M+ token...
No provider description is available for this model yet.
No provider description is available for this model yet.
MiniMax-M2.7 is a next-generation large language model designed for autonomous, real-world productivity and continuous improvement. Built to actively participate in its own evolution, M2.7 integrates advanced agentic capabilities through multi-agent...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
GLM-5.3 is a large-scale reasoning model from Z.ai, built for complex software engineering and long-horizon agent tasks. It supports text input and output with a 1M-token context window, and improves...
No provider description is available for this model yet.
No provider description is available for this model yet.
Gemma 4 31B Instruct is Google DeepMind's 30.7B dense multimodal model supporting text and image input with text output. Features a 256K token context window, configurable thinking/reasoning mode, native function...
A balanced model in the Ministral 3 family, Ministral 3 8B is a powerful, efficient tiny language model with vision capabilities.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
GLM 5.2 is a large-scale reasoning model from Z.ai. It supports text input and output with a 1M-token context window, and is suited for long-horizon agent workflows, project-level software engineering,...
No provider description is available for this model yet.
No provider description is available for this model yet.
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| qwen3.5-plusqwencloud/qwen3.5-plus | 991.808K | — | — | — | |||
| qwen3.7-maxqwencloud/qwen3.7-max | 991.808K | $2.5 | $7.5 | — | |||
| qwen3.7-plusqwencloud/qwen3.7-plus | 991.808K | — | — | — | |||
| Qwen: Qwen3.8 Flashqwen/qwen3.8-flash | 1M | $0.15 | $0.47 | — | |||
| Qwen: Qwen3 Coder 480B A35Bqwen/qwen3-coder | 262.144K | $0.3 | $1 | — | |||
| qwen3.8-maxqwencloud/qwen3.8-max | 991.808K | $2 | $6 | — | |||
| qwq-plusqwencloud/qwq-plus | 98.304K | $0.8 | $2.4 | — | |||
| deepseek-v4-flashqwen_ai_platform/deepseek-v4-flash | 1M | $0.2 | $0.4 | — | |||
| deepseek-v4-flash-0731qwen_ai_platform/deepseek-v4-flash-0731 | 1M | $0.2 | $0.4 | — | |||
| deepseek-v4-proqwen_ai_platform/deepseek-v4-pro | 1M | $2.4 | $4.8 | — | |||
| glm-5.1qwen_ai_platform/glm-5.1 | 202.745K | $1.4 | $4.4 | — | |||
| glm-5.2qwen_ai_platform/glm-5.2 | 1.04858M | $1.4 | $4.4 | — | |||
| kimi-k2.7-codeqwen_ai_platform/kimi-k2.7-code | 229.376K | $0.95 | $4 | — | |||
| qwen-coderqwen_ai_platform/qwen-coder | 1M | $0.3 | $1.5 | — | |||
| qwen-flashqwen_ai_platform/qwen-flash | 997.952K | — | — | — | |||
| qwen-flash-2025-07-28qwen_ai_platform/qwen-flash-2025-07-28 | 997.952K | — | — | — | |||
| qwen-maxqwen_ai_platform/qwen-max | 30.72K | $1.6 | $6.4 | — | |||
| Sakana: Fugu Ultra v2sakana/fugu-ultra-v2 | 1M | $5 | $30 | — | |||
| qwen-plusqwen_ai_platform/qwen-plus | 129.024K | $0.4 | $1.2 | — | |||
| qwen-plus-2025-01-25qwen_ai_platform/qwen-plus-2025-01-25 | 129.024K | $0.4 | $1.2 | — | |||
| qwen-plus-2025-04-28qwen_ai_platform/qwen-plus-2025-04-28 | 129.024K | $0.4 | $1.2 | — | |||
| qwen-plus-2025-07-14qwen_ai_platform/qwen-plus-2025-07-14 | 129.024K | $0.4 | $1.2 | — | |||
| qwen-plus-2025-07-28qwen_ai_platform/qwen-plus-2025-07-28 | 997.952K | — | — | — | |||
| qwen-plus-2025-09-11qwen_ai_platform/qwen-plus-2025-09-11 | 997.952K | — | — | — | |||
| qwen-plus-latestqwen_ai_platform/qwen-plus-latest | 997.952K | — | — | — | |||
| qwen-turboqwen_ai_platform/qwen-turbo | 129.024K | $0.05 | $0.2 | — | |||
| qwen-turbo-2024-11-01qwen_ai_platform/qwen-turbo-2024-11-01 | 1M | $0.05 | $0.2 | — | |||
| qwen-turbo-2025-04-28qwen_ai_platform/qwen-turbo-2025-04-28 | 1M | $0.05 | $0.2 | — | |||
| qwen-turbo-latestqwen_ai_platform/qwen-turbo-latest | 1M | $0.05 | $0.2 | — | |||
| qwen3-30b-a3bqwen_ai_platform/qwen3-30b-a3b | 129.024K | — | — | — | |||
| OpenAI: o3 Mini (batch)openai/o3-mini:batch | 200K | $0.55 | $2.2 | — | |||
| qwen3-coder-flashqwen_ai_platform/qwen3-coder-flash | 997.952K | — | — | — | |||
| OpenAI: GPT-5.5 (batch)openai/gpt-5.5:batch | 1.05M | $2.5 | $15 | — | |||
| qwen3-coder-flash-2025-07-28qwen_ai_platform/qwen3-coder-flash-2025-07-28 | 997.952K | — | — | — | |||
| qwen3-coder-plusqwen_ai_platform/qwen3-coder-plus | 997.952K | — | — | — | |||
| MiniMax: MiniMax M2.7minimax/minimax-m2.7 | 204.8K | $0.3 | $1.2 | — | |||
| qwen3-coder-plus-2025-07-22qwen_ai_platform/qwen3-coder-plus-2025-07-22 | 997.952K | — | — | — | |||
| qwen3-max-previewqwen_ai_platform/qwen3-max-preview | 258.048K | — | — | — | |||
| qwen3-maxqwen_ai_platform/qwen3-max | 258.048K | — | — | — | |||
| Z.ai: GLM 5.3z-ai/glm-5.3 | 1.04858M | $1.4 | $4.4 | — | |||
| qwen3-max-2026-01-23qwen_ai_platform/qwen3-max-2026-01-23 | 258.048K | — | — | — | |||
| qwen3-next-80b-a3b-instructqwen_ai_platform/qwen3-next-80b-a3b-instruct | 262.144K | $0.15 | $1.2 | — | |||
| Google: Gemma 4 31B (batch)google/gemma-4-31b-it:batch | 262.144K | $0.39 | $0.97 | — | |||
| Mistral: Ministral 3 8B 2512 (batch)mistralai/ministral-8b-2512:batch | 262.144K | $0.075 | $0.075 | — | |||
| qwen3-next-80b-a3b-thinkingqwen_ai_platform/qwen3-next-80b-a3b-thinking | 262.144K | $0.15 | $1.2 | — | |||
| qwen3-vl-235b-a22b-instructqwen_ai_platform/qwen3-vl-235b-a22b-instruct | 131.072K | $0.4 | $1.6 | — | |||
| qwen3-vl-235b-a22b-thinkingqwen_ai_platform/qwen3-vl-235b-a22b-thinking | 131.072K | $0.4 | $4 | — | |||
| Z.ai: GLM 5.2 (batch)z-ai/glm-5.2:batch | 1.04858M | $0.7 | $2.2 | — | |||
| qwen3-vl-32b-instructqwen_ai_platform/qwen3-vl-32b-instruct | 131.072K | $0.16 | $0.64 | — | |||
| qwen3-vl-32b-thinkingqwen_ai_platform/qwen3-vl-32b-thinking | 131.072K | $0.16 | $2.87 | — |