No provider description is available for this model yet.
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Ling 3.0 Flash Fin is a finance-focused mixture-of-experts model from InclusionAI, built on Ling 3.0 Flash with 5.1B active parameters out of 124B total. It is designed for real-world investment...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Nex-N2.5 is an agentic model built to turn goals into working, verified outcomes. Its core strength is agentic coding within a visual feedback loop: it can explore codebases, implement multi-file...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Mistral Medium 3.5 is a dense 128B instruction-following model from Mistral AI. It supports text and image inputs with text output, and is designed for agentic workflows, coding, and complex...
Mercury 2.5 is the fastest reasoning LLM, and the latest diffusion LLM (dLLM) from Inception. Instead of generating tokens sequentially, Mercury 2.5 produces and refines multiple tokens in parallel, achieving...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
A balanced model in the Ministral 3 family, Ministral 3 8B is a powerful, efficient tiny language model with vision capabilities.
The smallest model in the Ministral 3 family, Ministral 3 3B is a powerful, efficient tiny language model with vision capabilities.
DeepSeek V4 Pro 0813 is a large-scale mixture-of-experts model from DeepSeek. This is the GA release of DeepSeek V4 Pro.
Qwen3.5-9B is a multimodal foundation model from the Qwen3.5 family, designed to deliver strong reasoning, coding, and visual understanding in an efficient 9B-parameter architecture. It uses a unified vision-language design...
Fusion turns your prompt into a small multi-model deliberation. A panel of expert models (see below) analyzes your prompt in parallel with web search and web fetch enabled, then a...
Qwen3.8 Max 0902 is an updated snapshot of Qwen3.8 Max from Alibaba's Qwen team. It is a 2.4-trillion-parameter mixture-of-experts model that accepts text, image, and video input and returns text,...
No provider description is available for this model yet.
No provider description is available for this model yet.
Qwen3.7 Flash is a vision-language reasoning model from Alibaba. It is suited for multimodal agents, visual coding, search, and computer interaction, with strengths in object recognition, spatial understanding, and real-world...
No provider description is available for this model yet.
No provider description is available for this model yet.
NVIDIA-Nemotron-Nano-9B-v2 is a large language model (LLM) trained from scratch by NVIDIA, and designed as a unified model for both reasoning and non-reasoning tasks. It responds to user queries and...
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| microsoft/Dayhoff-170M-UR90-HL-18000microsoft/Dayhoff-170M-UR90-HL-18000 | Not documented | — | — | — | |||
| eu/gpt-5.6-terraazure/eu/gpt-5.6-terra | 1.05M | $2.75 | $16.5 | — | |||
| eu/gpt-5.6-lunaazure/eu/gpt-5.6-luna | 1.05M | $1.1 | $6.6 | — | |||
| microsoft/Dayhoff-170M-UR90-HL-14000microsoft/Dayhoff-170M-UR90-HL-14000 | Not documented | — | — | — | |||
| us.anthropic.claude-3-5-haiku-20241022-v1:0bedrock/us.anthropic.claude-3-5-haiku-20241022-v1:0 | 200K | $0.8 | $4 | — | |||
| pplx-7b-onlineperplexity/pplx-7b-online | 4.096K | — | $0.28 | — | |||
| sonar-medium-chatperplexity/sonar-medium-chat | 16.384K | $0.6 | $1.8 | — | |||
| sonar-medium-onlineperplexity/sonar-medium-online | 12K | — | $1.8 | — | |||
| microsoft/Dayhoff-170M-UR90-HL-12000microsoft/Dayhoff-170M-UR90-HL-12000 | Not documented | — | — | — | |||
| meta-llama/Llama-2-13b-hfmeta-llama/Llama-2-13b-hf | Not documented | — | — | — | |||
| zai-glm-4.6cerebras/zai-glm-4.6 | 128K | $2.25 | $2.75 | — | |||
| microsoft/Dayhoff-170M-UR90-HL-6000microsoft/Dayhoff-170M-UR90-HL-6000 | Not documented | — | — | — | |||
| gpt-5.5azure/gpt-5.5 | 1.05M | $5 | $30 | — | |||
| microsoft/Dayhoff-170M-UR90-HL-2000microsoft/Dayhoff-170M-UR90-HL-2000 | Not documented | — | — | — | |||
| deepseek-ai/ESFT-gate-summary-litedeepseek-ai/ESFT-gate-summary-lite | Not documented | — | — | — | |||
| microsoft/Dayhoff-170M-GRS-76000microsoft/Dayhoff-170M-GRS-76000 | Not documented | — | — | — | |||
| us/gpt-5.5azure/us/gpt-5.5 | 1.05M | $5.5 | $33 | — | |||
| eu/gpt-5.5azure/eu/gpt-5.5 | 1.05M | $5.5 | $33 | — | |||
| inclusionAI: Ling 3.0 Flash Fin (free)inclusionai/ling-3.0-flash-fin:free | 262.144K | Free | Free | — | |||
| microsoft/Dayhoff-170M-GRS-26000microsoft/Dayhoff-170M-GRS-26000 | Not documented | — | — | — | |||
| microsoft/Dayhoff-170M-GRS-112000microsoft/Dayhoff-170M-GRS-112000 | Not documented | — | — | — | |||
| microsoft/Dayhoff-170M-GRS-2000microsoft/Dayhoff-170M-GRS-2000 | Not documented | — | — | — | |||
| nvidia/Nemotron-Cascade-2-30B-A3Bnvidia/Nemotron-Cascade-2-30B-A3B | Not documented | — | — | — | |||
| eu.twelvelabs.pegasus-1-2-v1:0bedrock/eu.twelvelabs.pegasus-1-2-v1:0 | Not documented | — | $7.5 | — | |||
| nvidia/Kimi-K2.6-DFlashnvidia/Kimi-K2.6-DFlash | Not documented | — | — | — | |||
| gpt-5.5-2026-04-23azure/gpt-5.5-2026-04-23 | 1.05M | $5 | $30 | — | |||
| llama-3.3-70bcerebras/llama-3.3-70b | 128K | $0.85 | $1.2 | — | |||
| Nex AGI: Nex-N2.5-Mini (free)nex-agi/nex-n2.5-mini:free | 262.144K | Free | Free | — | |||
| deepseek-ai/deepseek-coder-1.3b-basedeepseek-ai/deepseek-coder-1.3b-base | Not documented | — | — | — | |||
| us.twelvelabs.pegasus-1-2-v1:0bedrock/us.twelvelabs.pegasus-1-2-v1:0 | Not documented | — | $7.5 | — | |||
| eu/gpt-5.5-2026-04-23azure/eu/gpt-5.5-2026-04-23 | 1.05M | $5.5 | $33 | — | |||
| Mistral: Mistral Medium 3.5mistralai/mistral-medium-3-5 | 262.144K | $1.5 | $7.5 | — | |||
| Inception: Mercury 2.5 Previewinception/mercury-2.5-preview | 260K | $0.2 | $0.75 | — | |||
| gpt-5.4-miniazure/gpt-5.4-mini | 272K | $0.75 | $4.5 | — | |||
| llama3.1-70bcerebras/llama3.1-70b | 128K | $0.6 | $0.6 | — | |||
| ibm-granite/granite-4.2-30bibm-granite/granite-4.2-30b | Not documented | — | — | — | |||
| ibm-granite/granite-4.2-3bibm-granite/granite-4.2-3b | Not documented | — | — | — | |||
| gpt-5.4-mini-2026-03-17azure/gpt-5.4-mini-2026-03-17 | 272K | $0.75 | $4.5 | — | |||
| Mistral: Ministral 3 8B 2512 (batch)mistralai/ministral-8b-2512:batch | 262.144K | $0.075 | $0.075 | — | |||
| Mistral: Ministral 3 3B 2512mistralai/ministral-3b-2512 | 131.072K | $0.1 | $0.1 | — | |||
| DeepSeek: DeepSeek V4 Pro 0813 (batch)deepseek/deepseek-v4-pro-0813:batch | 1.04858M | $0.66 | $1.98 | — | |||
| Qwen: Qwen3.5-9B (batch)qwen/qwen3.5-9b:batch | 262.144K | $0.17 | $0.25 | — | |||
| OpenRouter: Fusionopenrouter/fusion | 1M | — | — | — | |||
| Qwen: Qwen3.8 Max (0902)qwen/qwen3.8-max-0902 | 1M | $2 | $6 | — | |||
| gpt-5.4-nano-2026-03-17azure/gpt-5.4-nano-2026-03-17 | 272K | $0.2 | $1.25 | — | |||
| mistral-large-2402azure/mistral-large-2402 | 32K | $8 | $24 | — | |||
| Qwen: Qwen3.7 Flashqwen/qwen3.7-flash | 1M | $0.03 | $0.13 | — | |||
| mistral-large-latestazure/mistral-large-latest | 32K | $8 | $24 | — | |||
| o1azure/o1 | 200K | $15 | $60 | — | |||
| NVIDIA: Nemotron Nano 9B V2 (free)nvidia/nemotron-nano-9b-v2:free | 128K | Free | Free | — |