No provider description is available for this model yet.
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
GLM-5.3-Flash is a native multimodal model from Z.ai. It is suited for efficient coding and long-horizon agent tasks. Its hybrid sparse and linear attention architecture maintains accurate long-context behavior while...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Qwen3.8 Flash is a multimodal reasoning model from Alibaba. It is suited for coding assistance, agentic workflows, visual understanding, document and codebase analysis, desktop interaction, chart analysis, and long-video analysis.
Fugu Ultra v2 is the higher-performance model in Sakana AI's Fugu family. Rather than a single monolithic model, Fugu is a learned multi-agent orchestration system: a language model trained to...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Grok 4.20 Multi-Agent is a variant of SpaceXAI’s Grok 4.20 designed for collaborative, agent-based workflows. Multiple agents operate in parallel to conduct deep research, coordinate tool use, and synthesize information...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
MiniMax-M2.7 is a next-generation large language model designed for autonomous, real-world productivity and continuous improvement. Built to actively participate in its own evolution, M2.7 integrates advanced agentic capabilities through multi-agent...
No provider description is available for this model yet.
No provider description is available for this model yet.
GPT-6 Astra is OpenAI's flagship model for demanding end-to-end work. It is suited for advanced analysis, software engineering, deep research, scientific work, and document creation, with particular strengths in long-horizon...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
GPT-5 Pro is OpenAI’s most advanced model, offering major improvements in reasoning, code quality, and user experience. It is optimized for complex tasks that require step-by-step reasoning, instruction following, and...
Mistral's cutting-edge language model for coding released end of July 2025. Codestral specializes in low-latency, high-frequency tasks such as fill-in-the-middle (FIM), code correction and test generation. [Blog Post](https://mistral.ai/news/codestral-25-08)
No provider description is available for this model yet.
No provider description is available for this model yet.
Qwen Plus 0728, based on the Qwen3 foundation model, is a 1 million context hybrid reasoning model with a balanced performance, speed, and cost combination.
For tasks that demand low latency, GPT‑4.1 nano is the fastest and cheapest model in the GPT-4.1 series. It delivers exceptional performance at a small size with its 1 million...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| qwen3-maxqwencloud/qwen3-max | 258.048K | — | — | — | |||
| qwen3-max-2026-01-23qwencloud/qwen3-max-2026-01-23 | 258.048K | — | — | — | |||
| qwen3-next-80b-a3b-instructqwencloud/qwen3-next-80b-a3b-instruct | 262.144K | $0.15 | $1.2 | — | |||
| qwen3-next-80b-a3b-thinkingqwencloud/qwen3-next-80b-a3b-thinking | 262.144K | $0.15 | $1.2 | — | |||
| Z.ai: GLM 5.3 Flash (batch)z-ai/glm-5.3-flash:batch | 1.04858M | $0.075 | $0.25 | — | |||
| qwen3-vl-plusqwencloud/qwen3-vl-plus | 260.096K | — | — | — | |||
| qwen3.5-plusqwencloud/qwen3.5-plus | 991.808K | — | — | — | |||
| qwen3.7-maxqwencloud/qwen3.7-max | 991.808K | $2.5 | $7.5 | — | |||
| qwen3.7-plusqwencloud/qwen3.7-plus | 991.808K | — | — | — | |||
| qwen3.8-maxqwencloud/qwen3.8-max | 991.808K | $2 | $6 | — | |||
| Qwen: Qwen3.8 Flashqwen/qwen3.8-flash | 1M | $0.15 | $0.47 | — | |||
| Sakana: Fugu Ultra v2sakana/fugu-ultra-v2 | 1M | $5 | $30 | — | |||
| deepseek-v4-flashqwen_ai_platform/deepseek-v4-flash | 1M | $0.2 | $0.4 | — | |||
| deepseek-v4-flash-0731qwen_ai_platform/deepseek-v4-flash-0731 | 1M | $0.2 | $0.4 | — | |||
| deepseek-v4-proqwen_ai_platform/deepseek-v4-pro | 1M | $2.4 | $4.8 | — | |||
| glm-5.1qwen_ai_platform/glm-5.1 | 202.745K | $1.4 | $4.4 | — | |||
| glm-5.2qwen_ai_platform/glm-5.2 | 1.04858M | $1.4 | $4.4 | — | |||
| kimi-k2.7-codeqwen_ai_platform/kimi-k2.7-code | 229.376K | $0.95 | $4 | — | |||
| qwen-coderqwen_ai_platform/qwen-coder | 1M | $0.3 | $1.5 | — | |||
| qwen-flashqwen_ai_platform/qwen-flash | 997.952K | — | — | — | |||
| qwen-flash-2025-07-28qwen_ai_platform/qwen-flash-2025-07-28 | 997.952K | — | — | — | |||
| SpaceXAI: Grok 4.20 Multi-Agentx-ai/grok-4.20-multi-agent | 2M | $1.25 | $2.5 | — | |||
| qwen-plus-2025-07-28qwen_ai_platform/qwen-plus-2025-07-28 | 997.952K | — | — | — | |||
| qwen-plus-2025-09-11qwen_ai_platform/qwen-plus-2025-09-11 | 997.952K | — | — | — | |||
| qwen-plus-latestqwen_ai_platform/qwen-plus-latest | 997.952K | — | — | — | |||
| qwen-turbo-2024-11-01qwen_ai_platform/qwen-turbo-2024-11-01 | 1M | $0.05 | $0.2 | — | |||
| qwen-turbo-2025-04-28qwen_ai_platform/qwen-turbo-2025-04-28 | 1M | $0.05 | $0.2 | — | |||
| qwen-turbo-latestqwen_ai_platform/qwen-turbo-latest | 1M | $0.05 | $0.2 | — | |||
| qwen3-coder-flashqwen_ai_platform/qwen3-coder-flash | 997.952K | — | — | — | |||
| MiniMax: MiniMax M2.7minimax/minimax-m2.7 | 204.8K | $0.3 | $1.2 | — | |||
| qwen3-coder-flash-2025-07-28qwen_ai_platform/qwen3-coder-flash-2025-07-28 | 997.952K | — | — | — | |||
| qwen3-coder-plusqwen_ai_platform/qwen3-coder-plus | 997.952K | — | — | — | |||
| OpenAI: GPT-6 Astra (batch)openai/gpt-6-astra:batch | 1.05M | $5 | $25 | — | |||
| qwen3-coder-plus-2025-07-22qwen_ai_platform/qwen3-coder-plus-2025-07-22 | 997.952K | — | — | — | |||
| qwen3-max-previewqwen_ai_platform/qwen3-max-preview | 258.048K | — | — | — | |||
| qwen3-maxqwen_ai_platform/qwen3-max | 258.048K | — | — | — | |||
| qwen3-max-2026-01-23qwen_ai_platform/qwen3-max-2026-01-23 | 258.048K | — | — | — | |||
| qwen3-next-80b-a3b-instructqwen_ai_platform/qwen3-next-80b-a3b-instruct | 262.144K | $0.15 | $1.2 | — | |||
| OpenAI: GPT-5 Pro (batch)openai/gpt-5-pro:batch | 400K | $7.5 | $60 | — | |||
| Mistral: Codestral 2508 (batch)mistralai/codestral-2508:batch | 256K | $0.15 | $0.45 | — | |||
| qwen3-next-80b-a3b-thinkingqwen_ai_platform/qwen3-next-80b-a3b-thinking | 262.144K | $0.15 | $1.2 | — | |||
| qwen3-vl-plusqwen_ai_platform/qwen3-vl-plus | 260.096K | — | — | — | |||
| Qwen: Qwen Plus 0728qwen/qwen-plus-2025-07-28 | 1M | $0.26 | $0.78 | — | |||
| OpenAI: GPT-4.1 Nano (batch)openai/gpt-4.1-nano:batch | 1.04758M | $0.05 | $0.2 | — | |||
| qwen3.5-plusqwen_ai_platform/qwen3.5-plus | 991.808K | — | — | — | |||
| qwen3.7-maxqwen_ai_platform/qwen3.7-max | 991.808K | $2.5 | $7.5 | — | |||
| qwen3.7-plusqwen_ai_platform/qwen3.7-plus | 991.808K | — | — | — | |||
| qwen3.8-maxqwen_ai_platform/qwen3.8-max | 991.808K | $2 | $6 | — | |||
| databricks-deepseek-v4-flash-0731databricks/databricks-deepseek-v4-flash-0731 | 1M | $0.14 | $0.28 | — | |||
| databricks-deepseek-v4-pro-0813databricks/databricks-deepseek-v4-pro-0813 | 1M | $1.32 | $3.96 | — |