No provider description is available for this model yet.
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
MiniMax-M1 is a large-scale, open-weight reasoning model designed for extended context and high-efficiency inference. It leverages a hybrid Mixture-of-Experts (MoE) architecture paired with a custom "lightning attention" mechanism, allowing it...
Claude Fable 5 is a Mythos-class model from Anthropic, built for autonomous knowledge work and coding. It supports text, image, and file inputs with text output, with reasoning support and...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Opus 4.6 is Anthropic’s strongest model for coding and long-running professional tasks. It is built for agents that operate across entire workflows rather than single prompts, making it especially effective...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
GPT-4.1 Mini is a mid-sized model delivering performance competitive with GPT-4o at substantially lower latency and cost. It retains a 1 million token context window and scores 45.1% on hard...
Gemini 3.8 Flash is Google's most intelligent Flash model with significant gains from 3.7 Flash across software engineering, agentic tasks, and multi-step reasoning.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Gemini 2.5 Flash-Lite is a lightweight reasoning model in the Gemini 2.5 family, optimized for ultra-low latency and cost efficiency. It offers improved throughput, faster token generation, and better performance...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Claude Opus 5 is Anthropic’s flagship model for demanding reasoning, coding, and long-horizon agentic work. It is particularly strong at end-to-end software tasks, code review and bug finding, visual analysis...
GPT-5.6 Sol is the flagship model in OpenAI's GPT-5.6 series. It is suited for complex reasoning, coding, and agentic workflows, and is particularly strong at command-line and multi-step coding tasks...
Qwen3.7-Plus is a cost-effective model in Alibaba's Qwen3.7 series. It supports text and image input with text output, building on the series' text capabilities with a comprehensive upgrade to its...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Grok 4.3 is a reasoning model from SpaceXAI. It accepts text and image inputs with text output, and is suited for agentic workflows, instruction-following tasks, and applications requiring high factual...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| glm-5.3zai/glm-5.3 | 1M | $1.4 | $4.4 | — | |||
| minimax-m3tencent/minimax-m3 | 1M | $0.3 | $1.2 | — | |||
| zai-org/glm-5.3novita/zai-org/glm-5.3 | 1.04858M | $1.4 | $4.4 | — | |||
| deepseek/deepseek-v4-pro-0813novita/deepseek/deepseek-v4-pro-0813 | 1.04858M | $1.32 | $3.96 | — | |||
| moonshotai/kimi-k3novita/moonshotai/kimi-k3 | 1.04858M | $3 | $15 | — | |||
| zai-org/glm-5.2novita/zai-org/glm-5.2 | 1.04858M | $1.4 | $4.4 | — | |||
| MiniMax: MiniMax M1minimax/minimax-m1 | 1M | $0.55 | $2.2 | — | |||
| Anthropic: Claude Fable 5 (batch)anthropic/claude-fable-5:batch | 1M | $5 | $25 | — | |||
| deepseek/deepseek-v4-flash-vision-expnovita/deepseek/deepseek-v4-flash-vision-exp | 1.04858M | $0.44 | $1.32 | — | |||
| deepseek/deepseek-v4-flash-0731novita/deepseek/deepseek-v4-flash-0731 | 1.04858M | $0.44 | $1.32 | — | |||
| mindai/macaron-v1-ventinovita/mindai/macaron-v1-venti | 1.04858M | $1.5 | $4.5 | — | |||
| minimax/minimax-m3novita/minimax/minimax-m3 | 1M | $0.3 | $1.2 | — | |||
| deepseek/deepseek-v4-flashnovita/deepseek/deepseek-v4-flash | 1.04858M | $0.14 | $0.28 | — | |||
| deepseek/deepseek-v4-pronovita/deepseek/deepseek-v4-pro | 1.04858M | $1.6 | $3.2 | — | |||
| qwen/qwen3.8-maxnovita/qwen/qwen3.8-max | 1M | $2 | $6 | — | |||
| Anthropic: Claude Opus 4.6 (batch)anthropic/claude-opus-4.6:batch | 1M | $2.5 | $12.5 | — | |||
| xiaomimimo/mimo-v2.5novita/xiaomimimo/mimo-v2.5 | 1.04858M | $0.168 | $0.336 | — | |||
| qwen/qwen3.7-maxnovita/qwen/qwen3.7-max | 1M | $1.25 | $3.75 | — | |||
| xiaomimimo/mimo-v2.5-pronovita/xiaomimimo/mimo-v2.5-pro | 1.04858M | $0.522 | $1.044 | — | |||
| OpenAI: GPT-4.1 Mini (batch)openai/gpt-4.1-mini:batch | 1.04758M | $0.2 | $0.8 | — | |||
| Google: Gemini 3.8 Flash (batch)google/gemini-3.8-flash:batch | 1.04858M | $0.375 | $1.875 | — | |||
| databricks-claude-opus-4-6databricks/databricks-claude-opus-4-6 | 1M | $5 | $25 | — | |||
| databricks-claude-sonnet-4-6databricks/databricks-claude-sonnet-4-6 | 1M | $3 | $15 | — | |||
| databricks-gemini-3-1-flash-litedatabricks/databricks-gemini-3-1-flash-lite | 1.04858M | $0.312 | $1.875 | — | |||
| databricks-gemini-3-1-prodatabricks/databricks-gemini-3-1-pro | 1.04858M | $2.5 | $15 | — | |||
| databricks-gemini-3-flashdatabricks/databricks-gemini-3-flash | 1.04858M | $0.625 | $3.75 | — | |||
| databricks-gemini-3-prodatabricks/databricks-gemini-3-pro | 1.04858M | $2.5 | $15 | — | |||
| Google: Gemini 2.5 Flash Lite (batch)google/gemini-2.5-flash-lite:batch | 1.04858M | $0.05 | $0.2 | — | |||
| deepseek-ai/DeepSeek-V4-Flashwandb/deepseek-ai/deepseek-v4-flash | 1.04858M | $0.14 | $0.28 | — | |||
| deepseek-ai/DeepSeek-V4-Prowandb/deepseek-ai/deepseek-v4-pro | 1.04858M | $1.15 | $2.55 | — | |||
| anthropic/claude-opus-4-8deepinfra/anthropic/claude-opus-4-8 | 1M | $5 | $25 | — | |||
| anthropic/claude-sonnet-4-6deepinfra/anthropic/claude-sonnet-4-6 | 1M | $3 | $15 | — | |||
| google/gemini-3.5-flashdeepinfra/google/gemini-3.5-flash | 1M | $1.5 | $9 | — | |||
| databricks-glm-5-3-flashdatabricks/databricks-glm-5-3-flash | 1.04858M | — | — | — | |||
| zai-org/GLM-5.3together_ai/zai-org/glm-5.3 | 1.04858M | $1.4 | $4.4 | — | |||
| glm-5.3-flashzai/glm-5.3-flash | 1.04858M | $0.15 | $0.5 | — | |||
| grok-4.20xai/grok-4.20 | 1M | $1.25 | $2.5 | — | |||
| grok-4.20-reasoningxai/grok-4.20-reasoning | 1M | $1.25 | $2.5 | — | |||
| Claude Opus 5 (batch)anthropic/claude-opus-5:batch | 1M | $2.5 | $12.5 | — | |||
| OpenAI: GPT-5.6 Sol (batch)openai/gpt-5.6-sol:batch | 1.05M | $1 | $5 | — | |||
| Qwen: Qwen3.7 Plusqwen/qwen3.7-plus | 1M | $0.32 | $1.28 | — | |||
| grok-4.20-reasoning-latestxai/grok-4.20-reasoning-latest | 1M | $1.25 | $2.5 | — | |||
| grok-4.20-non-reasoningxai/grok-4.20-non-reasoning | 1M | $1.25 | $2.5 | — | |||
| grok-4.20-non-reasoning-latestxai/grok-4.20-non-reasoning-latest | 1M | $1.25 | $2.5 | — | |||
| anthropic/claude-opus-5deepinfra/anthropic/claude-opus-5 | 1M | $5 | $25 | — | |||
| SpaceXAI: Grok 4.3x-ai/grok-4.3 | 1M | $1.25 | $2.5 | — | |||
| deepseek-ai/DeepSeek-V4-Pro-0813deepinfra/deepseek-ai/deepseek-v4-pro-0813 | 1.04858M | $1.3 | $2.6 | — | |||
| google/gemini-3.1-flash-litedeepinfra/google/gemini-3.1-flash-lite | 1M | $0.25 | $1.5 | — | |||
| google/gemini-3.7-flashdeepinfra/google/gemini-3.7-flash | 1M | $0.75 | $3.75 | — | |||
| anthropic/claude-sonnet-5deepinfra/anthropic/claude-sonnet-5 | 1M | $2 | $10 | — |