No provider description is available for this model yet.
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
No provider description is available for this model yet.
No provider description is available for this model yet.
Kimi K3 is a 2.8T parameter open-weight multimodal reasoning model from Moonshot AI. It is suited for complex coding, knowledge work, and long-horizon agentic workflows, and is particularly strong at...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
This model always redirects to the latest GLM model from Z.ai.
No provider description is available for this model yet.
No provider description is available for this model yet.
The o-series of models are trained with reinforcement learning to think before they answer and perform complex reasoning. The o3-pro model uses more compute to think harder and provide consistently...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
GLM-5.3-Flash is a native multimodal model from Z.ai. It is suited for efficient coding and long-horizon agent tasks. Its hybrid sparse and linear attention architecture maintains accurate long-context behavior while...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Qwen3.8 Flash is a multimodal reasoning model from Alibaba. It is suited for coding assistance, agentic workflows, visual understanding, document and codebase analysis, desktop interaction, chart analysis, and long-video analysis.
No provider description is available for this model yet.
No provider description is available for this model yet.
Fugu Ultra v2 is the higher-performance model in Sakana AI's Fugu family. Rather than a single monolithic model, Fugu is a learned multi-agent orchestration system: a language model trained to...
GLM-5 is Z.ai’s flagship open-source foundation model engineered for complex systems design and long-horizon agent workflows. Built for expert developers, it delivers production-grade performance on large-scale programming tasks, rivaling leading...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Grok 4.20 Multi-Agent is a variant of SpaceXAI’s Grok 4.20 designed for collaborative, agent-based workflows. Multiple agents operate in parallel to conduct deep research, coordinate tool use, and synthesize information...
No provider description is available for this model yet.
No provider description is available for this model yet.
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| qwen-turboqwencloud/qwen-turbo | 129.024K | $0.05 | $0.2 | — | |||
| qwen-turbo-2024-11-01qwencloud/qwen-turbo-2024-11-01 | 1M | $0.05 | $0.2 | — | |||
| qwen-turbo-2025-04-28qwencloud/qwen-turbo-2025-04-28 | 1M | $0.05 | $0.2 | — | |||
| MoonshotAI: Kimi K3 (batch)moonshotai/kimi-k3:batch | 1.04858M | $3 | $15 | — | |||
| qwen-turbo-latestqwencloud/qwen-turbo-latest | 1M | $0.05 | $0.2 | — | |||
| qwen3-30b-a3bqwencloud/qwen3-30b-a3b | 129.024K | — | — | — | |||
| qwen3-coder-flashqwencloud/qwen3-coder-flash | 997.952K | — | — | — | |||
| qwen3-coder-flash-2025-07-28qwencloud/qwen3-coder-flash-2025-07-28 | 997.952K | — | — | — | |||
| Z.ai: GLM Latest~z-ai/glm-latest | 262.144K | $0.936 | $3.168 | — | |||
| qwen3-coder-plusqwencloud/qwen3-coder-plus | 997.952K | — | — | — | |||
| qwen3.6-27blibertai/qwen3.6-27b | 262.144K | $0.15 | $0.5 | — | |||
| OpenAI: o3 Pro (batch)openai/o3-pro:batch | 200K | $10 | $40 | — | |||
| qwen3-coder-plus-2025-07-22qwencloud/qwen3-coder-plus-2025-07-22 | 997.952K | — | — | — | |||
| qwen3-max-previewqwencloud/qwen3-max-preview | 258.048K | — | — | — | |||
| qwen3-maxqwencloud/qwen3-max | 258.048K | — | — | — | |||
| Meta-Llama-3_1-70B-Instructovhcloud/meta-llama-3_1-70b-instruct | 131K | $0.67 | $0.67 | — | |||
| qwen3-max-2026-01-23qwencloud/qwen3-max-2026-01-23 | 258.048K | — | — | — | |||
| qwen3-next-80b-a3b-instructqwencloud/qwen3-next-80b-a3b-instruct | 262.144K | $0.15 | $1.2 | — | |||
| qwen3-next-80b-a3b-thinkingqwencloud/qwen3-next-80b-a3b-thinking | 262.144K | $0.15 | $1.2 | — | |||
| qwen3-vl-235b-a22b-instructqwencloud/qwen3-vl-235b-a22b-instruct | 131.072K | $0.4 | $1.6 | — | |||
| Z.ai: GLM 5.3 Flash (batch)z-ai/glm-5.3-flash:batch | 1.04858M | $0.075 | $0.25 | — | |||
| jamba-1.5-miniai21/jamba-1.5-mini | 256K | $0.2 | $0.4 | — | |||
| qwen3-vl-235b-a22b-thinkingqwencloud/qwen3-vl-235b-a22b-thinking | 131.072K | $0.4 | $4 | — | |||
| gemma-4-31b-it-thinkinglibertai/gemma-4-31b-it-thinking | 262.144K | $0.15 | $0.4 | — | |||
| qwen3-vl-32b-instructqwencloud/qwen3-vl-32b-instruct | 131.072K | $0.16 | $0.64 | — | |||
| qwen3-vl-32b-thinkingqwencloud/qwen3-vl-32b-thinking | 131.072K | $0.16 | $2.87 | — | |||
| gemma-4-31b-itlibertai/gemma-4-31b-it | 262.144K | $0.15 | $0.4 | — | |||
| qwen3-vl-plusqwencloud/qwen3-vl-plus | 260.096K | — | — | — | |||
| qwen3.5-plusqwencloud/qwen3.5-plus | 991.808K | — | — | — | |||
| qwen3.7-maxqwencloud/qwen3.7-max | 991.808K | $2.5 | $7.5 | — | |||
| qwen3.7-plusqwencloud/qwen3.7-plus | 991.808K | — | — | — | |||
| Llama-3.1-8B-Instructovhcloud/llama-3.1-8b-instruct | 131K | $0.1 | $0.1 | — | |||
| Qwen: Qwen3.8 Flashqwen/qwen3.8-flash | 1M | $0.15 | $0.47 | — | |||
| qwen3.8-maxqwencloud/qwen3.8-max | 991.808K | $2 | $6 | — | |||
| qwq-plusqwencloud/qwq-plus | 98.304K | $0.8 | $2.4 | — | |||
| Sakana: Fugu Ultra v2sakana/fugu-ultra-v2 | 1M | $5 | $30 | — | |||
| Z.ai: GLM 5z-ai/glm-5 | 198K | $0.6 | $1.92 | — | |||
| gemma3-4bllamagate/gemma3-4b | 128K | $0.03 | $0.08 | — | |||
| deepseek-v4-flashqwen_ai_platform/deepseek-v4-flash | 1M | $0.2 | $0.4 | — | |||
| deepseek-v4-flash-0731qwen_ai_platform/deepseek-v4-flash-0731 | 1M | $0.2 | $0.4 | — | |||
| deepseek-v4-proqwen_ai_platform/deepseek-v4-pro | 1M | $2.4 | $4.8 | — | |||
| glm-5.1qwen_ai_platform/glm-5.1 | 202.745K | $1.4 | $4.4 | — | |||
| glm-5.2qwen_ai_platform/glm-5.2 | 1.04858M | $1.4 | $4.4 | — | |||
| kimi-k2.7-codeqwen_ai_platform/kimi-k2.7-code | 229.376K | $0.95 | $4 | — | |||
| qwen-coderqwen_ai_platform/qwen-coder | 1M | $0.3 | $1.5 | — | |||
| qwen-flashqwen_ai_platform/qwen-flash | 997.952K | — | — | — | |||
| qwen-flash-2025-07-28qwen_ai_platform/qwen-flash-2025-07-28 | 997.952K | — | — | — | |||
| SpaceXAI: Grok 4.20 Multi-Agentx-ai/grok-4.20-multi-agent | 2M | $1.25 | $2.5 | — | |||
| qwen-plusqwen_ai_platform/qwen-plus | 129.024K | $0.4 | $1.2 | — | |||
| qwen3-vl-8bllamagate/qwen3-vl-8b | 32.768K | $0.15 | $0.55 | — |