No provider description is available for this model yet.
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
Mercury 2 is an extremely fast reasoning LLM, and the first reasoning diffusion LLM (dLLM). Instead of generating tokens sequentially, Mercury 2 produces and refines multiple tokens in parallel, achieving...
No provider description is available for this model yet.
GPT-4o mini is OpenAI's newest model after [GPT-4 Omni](/models/openai/gpt-4o), supporting both text and image inputs with text outputs. As their most advanced small model, it is many multiples more affordable...
Nex-N2.5 is an agentic model built to turn goals into working, verified outcomes. Its core strength is agentic coding within a visual feedback loop: it can explore codebases, implement multi-file...
No provider description is available for this model yet.
No provider description is available for this model yet.
As a 30B-class SOTA model, GLM-4.7-Flash offers a new option that balances performance and efficiency. It is further optimized for agentic coding use cases, strengthening coding capabilities, long-horizon task planning,...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
GPT-5.5 Pro is OpenAI’s high-capability model optimized for deep reasoning and accuracy on complex, high-stakes workloads. It features a 1M+ token context window (922K input, 128K output) with support for...
No provider description is available for this model yet.
Gemini 3 Flash Preview is a high speed, high value thinking model designed for agentic workflows, multi turn chat, and coding assistance. It delivers near Pro level reasoning and tool...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Claude Opus 4.8 is Anthropic's most capable generally available model in the Opus family. It supports text, image, and file inputs with text output, with reasoning support and a 1M-token...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
MiniMax-M1 is a large-scale, open-weight reasoning model designed for extended context and high-efficiency inference. It leverages a hybrid Mixture-of-Experts (MoE) architecture paired with a custom "lightning attention" mechanism, allowing it...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Claude Opus 4.5 is Anthropic’s frontier reasoning model optimized for complex software engineering, agentic workflows, and long-horizon computer use. It offers strong multimodal capabilities, competitive performance across real-world coding and...
Fugu Max is the cost-performance model in Sakana AI's Fugu family. Rather than a single monolithic model, Fugu is a learned multi-agent orchestration system: a language model trained to route...
Grok 4.3 is a reasoning model from SpaceXAI. It accepts text and image inputs with text output, and is suited for agentic workflows, instruction-following tasks, and applications requiring high factual...
No provider description is available for this model yet.
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| gpt-5.6-sol-discaihubmix/gpt-5.6-sol-disc | 1.05M | $4 | $20 | — | |||
| Inception: Mercury 2inception/mercury-2 | 128K | $0.25 | $0.75 | — | |||
| gpt-5.6-terraaihubmix/gpt-5.6-terra | 1.05M | $2 | $12 | — | |||
| OpenAI: GPT-4o-mini (batch)openai/gpt-4o-mini:batch | 128K | $0.075 | $0.3 | — | |||
| Nex AGI: Nex-N2.5-Pro (free)nex-agi/nex-n2.5-pro:free | 262.144K | Free | Free | — | |||
| gpt-chat-latestaihubmix/gpt-chat-latest | 400K | $5 | $30 | — | |||
| grok-4-20-non-reasoningaihubmix/grok-4-20-non-reasoning | 1M | $2 | $6 | — | |||
| Z.ai: GLM 4.7 Flashz-ai/glm-4.7-flash | 131.072K | $0.061 | $0.4 | — | |||
| grok-4-20-reasoningaihubmix/grok-4-20-reasoning | 1M | $2 | $6 | — | |||
| grok-4.6aihubmix/grok-4.6 | 500K | $2 | $6 | — | |||
| grok-build-0.1aihubmix/grok-build-0.1 | 256K | $1 | $2 | — | |||
| hy3aihubmix/hy3 | 256K | $0.156 | $0.625 | — | |||
| hy4-previewaihubmix/hy4-preview | 1.04858M | $0.845 | $2.535 | — | |||
| kimi-k2.6aihubmix/kimi-k2.6 | 262.144K | $0.95 | $4 | — | |||
| kimi-k2.7-code-highspeedaihubmix/kimi-k2.7-code-highspeed | 262.144K | $1.9 | $7.999 | — | |||
| kimi-k3aihubmix/kimi-k3 | 1.04858M | $3 | $15 | — | |||
| longcat-2.0aihubmix/longcat-2.0 | 1M | $0.775 | $3.098 | — | |||
| OpenAI: GPT-5.5 Pro (batch)openai/gpt-5.5-pro:batch | 1.05M | $15 | $90 | — | |||
| mai-thinking-1aihubmix/mai-thinking-1 | 256K | $2 | $8 | — | |||
| Google: Gemini 3 Flash Preview (batch)google/gemini-3-flash-preview:batch | 1.04858M | $0.25 | $1.5 | — | |||
| mimo-v2-omniaihubmix/mimo-v2-omni | 256K | $0.44 | $2.2 | — | |||
| mimo-v2-proaihubmix/mimo-v2-pro | 1M | $1.1 | $3.3 | — | |||
| minimax-m2.7aihubmix/minimax-m2.7 | 204.8K | $0.296 | $1.183 | — | |||
| databricks-glm-5-2databricks/databricks-glm-5-2 | 1M | $1.4 | $4.4 | — | |||
| databricks-kimi-k3databricks/databricks-kimi-k3 | 1M | $3 | $15 | — | |||
| minimax-m3aihubmix/minimax-m3 | 1M | $0.288 | $1.152 | — | |||
| Anthropic: Claude Opus 4.8 (batch)anthropic/claude-opus-4.8:batch | 1M | $2.5 | $12.5 | — | |||
| accounts/fireworks/models/deepseek-v4-pro-0813fireworks_ai/accounts/fireworks/models/deepseek-v4-pro-0813 | 1.04858M | $1.32 | $3.96 | — | |||
| ministral-14b-2512mistral/ministral-14b-2512 | 262.144K | $0.2 | $0.2 | — | |||
| ministral-14b-latestmistral/ministral-14b-latest | 262.144K | $0.2 | $0.2 | — | |||
| muse-spark-1.2aihubmix/muse-spark-1.2 | 1.04858M | $1.375 | $4.675 | — | |||
| ministral-3b-2512mistral/ministral-3b-2512 | 131.072K | $0.1 | $0.1 | — | |||
| ministral-3b-latestmistral/ministral-3b-latest | 131.072K | $0.1 | $0.1 | — | |||
| mistral-medium-3mistral/mistral-medium-3 | 262.144K | $1.5 | $7.5 | — | |||
| voxtral-small-2507mistral/voxtral-small-2507 | 32.768K | $0.1 | $0.4 | — | |||
| zai-org/GLM-5.3-Flashtogether_ai/zai-org/glm-5.3-flash | 1.04858M | $0.15 | $0.5 | — | |||
| MiniMax: MiniMax M1minimax/minimax-m1 | 1M | $0.4 | $2.2 | — | |||
| glm-5.3zai/glm-5.3 | 1M | $1.4 | $4.4 | — | |||
| minimax-m3tencent/minimax-m3 | 1M | $0.3 | $1.2 | — | |||
| zai-org/glm-5.3novita/zai-org/glm-5.3 | 1.04858M | $1.4 | $4.4 | — | |||
| deepseek/deepseek-v4-pro-0813novita/deepseek/deepseek-v4-pro-0813 | 1.04858M | $1.32 | $3.96 | — | |||
| moonshotai/kimi-k3novita/moonshotai/kimi-k3 | 1.04858M | $3 | $15 | — | |||
| tencent/hy3novita/tencent/hy3 | 262.144K | $0.14 | $0.58 | — | |||
| zai-org/glm-5.2novita/zai-org/glm-5.2 | 1.04858M | $1.4 | $4.4 | — | |||
| qwen3-coder-nextaihubmix/qwen3-coder-next | 262.144K | $0.137 | $0.548 | — | |||
| moonshotai/kimi-k2.7-codenovita/moonshotai/kimi-k2.7-code | 262.144K | $0.95 | $4 | — | |||
| Anthropic: Claude Opus 4.5 (batch)anthropic/claude-opus-4.5:batch | 200K | $2.5 | $12.5 | — | |||
| Sakana: Fugu Maxsakana/fugu-max | 1M | $2 | $6 | — | |||
| SpaceXAI: Grok 4.3x-ai/grok-4.3 | 1M | $1.25 | $2.5 | — | |||
| deepseek/deepseek-v4-flash-vision-expnovita/deepseek/deepseek-v4-flash-vision-exp | 1.04858M | $0.44 | $1.32 | — |