No provider description is available for this model yet.
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Schematron V2 Turbo is a 3B-parameter HTML-to-JSON extraction model from Inference.net. It prioritizes throughput for high-volume extraction workloads. Extraction instructions must be supplied through a JSON schema in response_format rather...
Nex-N2.5 is an agentic model built to turn goals into working, verified outcomes. Its core strength is agentic coding within a visual feedback loop: it can explore codebases, implement multi-file...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Inkling is an open-weight multimodal mixture-of-experts model from Thinking Machines Lab, with 41B active parameters out of 975B total. It is designed for general-purpose reasoning, coding, agentic and tool-use systems,...
GPT-5.5 Pro is OpenAI’s high-capability model optimized for deep reasoning and accuracy on complex, high-stakes workloads. It features a 1M+ token context window (922K input, 128K output) with support for...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Claude Opus 4.8 is Anthropic's most capable generally available model in the Opus family. It supports text, image, and file inputs with text output, with reasoning support and a 1M-token...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
MiniMax-M1 is a large-scale, open-weight reasoning model designed for extended context and high-efficiency inference. It leverages a hybrid Mixture-of-Experts (MoE) architecture paired with a custom "lightning attention" mechanism, allowing it...
No provider description is available for this model yet.
No provider description is available for this model yet.
Fugu Max is the cost-performance model in Sakana AI's Fugu family. Rather than a single monolithic model, Fugu is a learned multi-agent orchestration system: a language model trained to route...
No provider description is available for this model yet.
No provider description is available for this model yet.
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| gpt-5.4-nanoaihubmix/gpt-5.4-nano | 400K | $0.2 | $1.25 | — | |||
| gpt-5.5aihubmix/gpt-5.5 | 1.05M | $5 | $30 | — | |||
| gpt-5.5-proaihubmix/gpt-5.5-pro | 1.05M | $30 | $180 | — | |||
| gpt-5.6-lunaaihubmix/gpt-5.6-luna | 1.05M | $0.2 | $1.2 | — | |||
| gpt-5.6-sol-discaihubmix/gpt-5.6-sol-disc | 1.05M | $4 | $20 | — | |||
| gpt-5.6-terraaihubmix/gpt-5.6-terra | 1.05M | $2 | $12 | — | |||
| Inference.net: Schematron V2 Turboinference-net/schematron-v2-turbo | 128K | $0.03 | $0.15 | — | |||
| Nex AGI: Nex-N2.5-Pro (free)nex-agi/nex-n2.5-pro:free | 262.144K | Free | Free | — | |||
| gpt-chat-latestaihubmix/gpt-chat-latest | 400K | $5 | $30 | — | |||
| grok-4-20-non-reasoningaihubmix/grok-4-20-non-reasoning | 1M | $2 | $6 | — | |||
| grok-4-20-reasoningaihubmix/grok-4-20-reasoning | 1M | $2 | $6 | — | |||
| grok-4.6aihubmix/grok-4.6 | 500K | $2 | $6 | — | |||
| grok-build-0.1aihubmix/grok-build-0.1 | 256K | $1 | $2 | — | |||
| hy3aihubmix/hy3 | 256K | $0.156 | $0.625 | — | |||
| hy4-previewaihubmix/hy4-preview | 1.04858M | $0.845 | $2.535 | — | |||
| kimi-k2.6aihubmix/kimi-k2.6 | 262.144K | $0.95 | $4 | — | |||
| kimi-k2.7-code-highspeedaihubmix/kimi-k2.7-code-highspeed | 262.144K | $1.9 | $7.999 | — | |||
| kimi-k3aihubmix/kimi-k3 | 1.04858M | $3 | $15 | — | |||
| longcat-2.0aihubmix/longcat-2.0 | 1M | $0.775 | $3.098 | — | |||
| mai-thinking-1aihubmix/mai-thinking-1 | 256K | $2 | $8 | — | |||
| Thinking Machines: Inkling (batch)thinkingmachines/inkling:batch | 524.288K | $1 | $4.05 | — | |||
| OpenAI: GPT-5.5 Pro (batch)openai/gpt-5.5-pro:batch | 1.05M | $15 | $90 | — | |||
| mimo-v2-omniaihubmix/mimo-v2-omni | 256K | $0.44 | $2.2 | — | |||
| mimo-v2-proaihubmix/mimo-v2-pro | 1M | $1.1 | $3.3 | — | |||
| minimax-m2.7aihubmix/minimax-m2.7 | 204.8K | $0.296 | $1.183 | — | |||
| databricks-glm-5-2databricks/databricks-glm-5-2 | 1M | $1.4 | $4.4 | — | |||
| databricks-kimi-k3databricks/databricks-kimi-k3 | 1M | $3 | $15 | — | |||
| minimax-m3aihubmix/minimax-m3 | 1M | $0.288 | $1.152 | — | |||
| Anthropic: Claude Opus 4.8 (batch)anthropic/claude-opus-4.8:batch | 1M | $2.5 | $12.5 | — | |||
| accounts/fireworks/models/deepseek-v4-pro-0813fireworks_ai/accounts/fireworks/models/deepseek-v4-pro-0813 | 1.04858M | $1.32 | $3.96 | — | |||
| ministral-14b-2512mistral/ministral-14b-2512 | 262.144K | $0.2 | $0.2 | — | |||
| ministral-14b-latestmistral/ministral-14b-latest | 262.144K | $0.2 | $0.2 | — | |||
| muse-spark-1.2aihubmix/muse-spark-1.2 | 1.04858M | $1.375 | $4.675 | — | |||
| ministral-3b-2512mistral/ministral-3b-2512 | 131.072K | $0.1 | $0.1 | — | |||
| ministral-3b-latestmistral/ministral-3b-latest | 131.072K | $0.1 | $0.1 | — | |||
| mistral-medium-3mistral/mistral-medium-3 | 262.144K | $1.5 | $7.5 | — | |||
| zai-org/GLM-5.3-Flashtogether_ai/zai-org/glm-5.3-flash | 1.04858M | $0.15 | $0.5 | — | |||
| glm-5.3zai/glm-5.3 | 1M | $1.4 | $4.4 | — | |||
| minimax-m3tencent/minimax-m3 | 1M | $0.3 | $1.2 | — | |||
| zai-org/glm-5.3novita/zai-org/glm-5.3 | 1.04858M | $1.4 | $4.4 | — | |||
| deepseek/deepseek-v4-pro-0813novita/deepseek/deepseek-v4-pro-0813 | 1.04858M | $1.32 | $3.96 | — | |||
| moonshotai/kimi-k3novita/moonshotai/kimi-k3 | 1.04858M | $3 | $15 | — | |||
| tencent/hy3novita/tencent/hy3 | 262.144K | $0.14 | $0.58 | — | |||
| zai-org/glm-5.2novita/zai-org/glm-5.2 | 1.04858M | $1.4 | $4.4 | — | |||
| MiniMax: MiniMax M1minimax/minimax-m1 | 1M | $0.4 | $2.2 | — | |||
| qwen3-coder-nextaihubmix/qwen3-coder-next | 262.144K | $0.137 | $0.548 | — | |||
| moonshotai/kimi-k2.7-codenovita/moonshotai/kimi-k2.7-code | 262.144K | $0.95 | $4 | — | |||
| Sakana: Fugu Maxsakana/fugu-max | 1M | $2 | $6 | — | |||
| deepseek/deepseek-v4-flash-vision-expnovita/deepseek/deepseek-v4-flash-vision-exp | 1.04858M | $0.44 | $1.32 | — | |||
| deepseek/deepseek-v4-flash-0731novita/deepseek/deepseek-v4-flash-0731 | 1.04858M | $0.44 | $1.32 | — |