No provider description is available for this model yet.
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
GLM-5.3-Flash is a native multimodal model from Z.ai. It is suited for efficient coding and long-horizon agent tasks. Its hybrid sparse and linear attention architecture maintains accurate long-context behavior while...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Grok 4.5 is a model from SpaceXAI with frontier performance on coding, knowledge work, and STEM.
Sonnet 5 is Anthropic's most capable Sonnet-class model, with frontier performance across coding, agents, and professional work. It supports adaptive thinking with selectable reasoning effort levels (low, medium, high, max,...
Qwen 3.6 Plus builds on a hybrid architecture that combines efficient linear attention with sparse mixture-of-experts routing, enabling strong scalability and high-performance inference. Compared to the 3.5 series, it delivers...
No provider description is available for this model yet.
Ministral 8B is an 8B parameter model featuring a unique interleaved sliding-window attention pattern for faster, memory-efficient inference. Designed for edge use cases, it supports up to 128k context length...
No provider description is available for this model yet.
No provider description is available for this model yet.
Fugu Max is the cost-performance model in Sakana AI's Fugu family. Rather than a single monolithic model, Fugu is a learned multi-agent orchestration system: a language model trained to route...
GLM-5 is Z.ai’s flagship open-source foundation model engineered for complex systems design and long-horizon agent workflows. Built for expert developers, it delivers production-grade performance on large-scale programming tasks, rivaling leading...
No provider description is available for this model yet.
NVIDIA Nemotron 3.5 Lightning is an open mixture-of-experts model from NVIDIA, with 3B active parameters out of 30B total. It is suited for high-throughput agentic workloads and specialized tasks that...
No provider description is available for this model yet.
GLM-4.5V is a vision-language foundation model for multimodal agent applications. Built on a Mixture-of-Experts (MoE) architecture with 106B parameters and 12B activated parameters, it achieves state-of-the-art results in video understanding,...
gpt-oss-20b is an open-weight 21B parameter model released by OpenAI under the Apache 2.0 license. It uses a Mixture-of-Experts (MoE) architecture with 3.6B active parameters per forward pass, optimized for...
Gemini 2.5 Pro is Google’s state-of-the-art AI model designed for advanced reasoning, coding, mathematics, and scientific tasks. It employs “thinking” capabilities, enabling it to reason through responses with enhanced accuracy...
No provider description is available for this model yet.
No provider description is available for this model yet.
GPT-4o mini is OpenAI's newest model after [GPT-4 Omni](/models/openai/gpt-4o), supporting both text and image inputs with text outputs. As their most advanced small model, it is many multiples more affordable...
No provider description is available for this model yet.
Granite 4.2 8B is a dense reasoning model from IBM. It is suited for mathematics, code generation, multilingual dialogue, and agentic workflows that need multi-step reasoning. It supports full, low-effort,...
Claude Fable 5 is a Mythos-class model from Anthropic, built for autonomous knowledge work and coding. It supports text, image, and file inputs with text output, with reasoning support and...
No provider description is available for this model yet.
Opus 4.6 is Anthropic’s strongest model for coding and long-running professional tasks. It is built for agents that operate across entire workflows rather than single prompts, making it especially effective...
No provider description is available for this model yet.
Gemini 3.5 Flash Lite is a high-efficiency model from Google with upgraded agentic capabilities. It is suited for subagents that execute focused tasks within complex, multi-agent workflows.
MiniMax-M3 is a multimodal foundation model from MiniMax. It supports text, image, and video inputs with text output, a 1M-token context window, and is suited for long-horizon agentic work, coding,...
GPT-5.6 Sol is the flagship model in OpenAI's GPT-5.6 series. It is suited for complex reasoning, coding, and agentic workflows, and is particularly strong at command-line and multi-step coding tasks...
Gemini 3 Flash Preview is a high speed, high value thinking model designed for agentic workflows, multi turn chat, and coding assistance. It delivers near Pro level reasoning and tool...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| ps/qwen3-30b-a3bpinstripes/ps/qwen3-30b-a3b | 131.072K | $0.09 | $0.2 | — | |||
| ps/qwen3-coder-30b-a3bpinstripes/ps/qwen3-coder-30b-a3b | 131.072K | $0.3 | $0.6 | — | |||
| ps/deepseek-v4-flashpinstripes/ps/deepseek-v4-flash | 163.84K | $0.1 | $0.2 | — | |||
| ps/minimax-m2.7pinstripes/ps/minimax-m2.7 | 1.00019M | $0.255 | $0.55 | — | |||
| gemma-4-26bdarkbloom/gemma-4-26b | 131.072K | $0.03 | $0.165 | — | |||
| gpt-oss-20bdarkbloom/gpt-oss-20b | 131.072K | $0.015 | $0.07 | — | |||
| meta-llama/llama-guard-3-11b-visionwatsonx/meta-llama/llama-guard-3-11b-vision | 128K | $0.35 | $0.35 | — | |||
| meta-llama/llama-4-maverick-17bwatsonx/meta-llama/llama-4-maverick-17b | 128K | $0.35 | $1.4 | — | |||
| mistral.magistral-small-2509bedrock_converse/mistral.magistral-small-2509 | 128K | $0.5 | $1.5 | — | |||
| accounts/fireworks/models/glm-5p1fireworks_ai/accounts/fireworks/models/glm-5p1 | 202.8K | $1.4 | $4.4 | — | |||
| meta-llama/llama-3-3-70b-instructwatsonx/meta-llama/llama-3-3-70b-instruct | 128K | $0.71 | $0.71 | — | |||
| meta-llama/llama-3-2-90b-vision-instructwatsonx/meta-llama/llama-3-2-90b-vision-instruct | 128K | $2 | $2 | — | |||
| mistral.devstral-2-123bbedrock_converse/mistral.devstral-2-123b | 256K | $0.4 | $2 | — | |||
| Z.ai: GLM 5.3 Flashz-ai/glm-5.3-flash | 1.04858M | $0.075 | $0.25 | — | |||
| meta-llama/llama-3-2-3b-instructwatsonx/meta-llama/llama-3-2-3b-instruct | 128K | $0.15 | $0.15 | — | |||
| meta-llama/llama-3-2-1b-instructwatsonx/meta-llama/llama-3-2-1b-instruct | 128K | $0.1 | $0.1 | — | |||
| MiniMax-M2.5-lightningminimax/minimax-m2.5-lightning | 1M | $0.3 | $2.4 | — | |||
| accounts/fireworks/models/glm-4p7fireworks_ai/accounts/fireworks/models/glm-4p7 | 202.8K | $0.6 | $2.2 | — | |||
| SpaceXAI: Grok 4.5x-ai/grok-4.5 | 500K | $2 | $6 | — | |||
| Anthropic: Claude Sonnet 5 (batch)anthropic/claude-sonnet-5:batch | 1M | $1 | $5 | — | |||
| Qwen: Qwen3.6 Plusqwen/qwen3.6-plus | 1M | $0.325 | $1.95 | — | |||
| qwen3-coder-flashdashscope/qwen3-coder-flash | 997.952K | — | — | — | |||
| Mistral: Ministral 8Bmistralai/ministral-8b | 128K | $0.11 | $0.11 | — | |||
| mistral-large-latestazure_ai/mistral-large-latest | 128K | $2 | $6 | — | |||
| meta-llama/llama-3-2-11b-vision-instructwatsonx/meta-llama/llama-3-2-11b-vision-instruct | 128K | $0.35 | $0.35 | — | |||
| Sakana: Fugu Maxsakana/fugu-max | 1M | $2 | $6 | — | |||
| Z.ai: GLM 5z-ai/glm-5 | 198K | $0.6 | $1.92 | — | |||
| openai/gpt-5.6-solopenrouter/openai/gpt-5.6-sol | 1.05M | $2 | $10 | — | |||
| NVIDIA: Nemotron 3.5 Lightning (free)nvidia/nemotron-3.5-lightning:free | 1M | Free | Free | — | |||
| ibm/granite-vision-3-2-2bwatsonx/ibm/granite-vision-3-2-2b | 8.192K | $0.1 | $0.1 | — | |||
| Z.ai: GLM 4.5Vz-ai/glm-4.5v | 65.536K | $0.6 | $1.8 | — | |||
| OpenAI: gpt-oss-20b (batch)openai/gpt-oss-20b:batch | 131.072K | $0.05 | $0.2 | — | |||
| Google: Gemini 2.5 Pro (batch)google/gemini-2.5-pro:batch | 1.04858M | $0.625 | $5 | — | |||
| MiniMax-M2.1-lightningminimax/minimax-m2.1-lightning | 1M | $0.3 | $2.4 | — | |||
| ibm/granite-guardian-3-3-8bwatsonx/ibm/granite-guardian-3-3-8b | 8.192K | $0.2 | $0.2 | — | |||
| OpenAI: GPT-4o-mini (batch)openai/gpt-4o-mini:batch | 128K | $0.075 | $0.3 | — | |||
| ibm/granite-guardian-3-2-2bwatsonx/ibm/granite-guardian-3-2-2b | 8.192K | $0.1 | $0.1 | — | |||
| IBM: Granite 4.2 8Bibm-granite/granite-4.2-8b | 131.072K | $0.06 | $0.25 | — | |||
| Anthropic: Claude Fable 5 (batch)anthropic/claude-fable-5:batch | 1M | $5 | $25 | — | |||
| minimax.minimax-m2.5bedrock_converse/minimax.minimax-m2.5 | 1M | $0.3 | $1.2 | — | |||
| Anthropic: Claude Opus 4.6 (batch)anthropic/claude-opus-4.6:batch | 1M | $2.5 | $12.5 | — | |||
| accounts/fireworks/models/glm-4p6fireworks_ai/accounts/fireworks/models/glm-4p6 | 202.8K | $0.55 | $2.19 | — | |||
| Google: Gemini 3.5 Flash Lite (batch)google/gemini-3.5-flash-lite:batch | 1.04858M | $0.15 | $1.25 | — | |||
| MiniMax: MiniMax M3minimax/minimax-m3 | 524.288K | $0.3 | $1.2 | — | |||
| OpenAI: GPT-5.6 Sol (batch)openai/gpt-5.6-sol:batch | 1.05M | $1 | $5 | — | |||
| Google: Gemini 3 Flash Preview (batch)google/gemini-3-flash-preview:batch | 1.04858M | $0.25 | $1.5 | — | |||
| ibm/granite-4-h-smallwatsonx/ibm/granite-4-h-small | 20.48K | $0.06 | $0.25 | — | |||
| ibm/granite-3-3-8b-instructwatsonx/ibm/granite-3-3-8b-instruct | 8.192K | $0.2 | $0.2 | — | |||
| minimax.minimax-m2.1bedrock_converse/minimax.minimax-m2.1 | 196K | $0.3 | $1.2 | — | |||
| ibm/granite-13b-instruct-v2watsonx/ibm/granite-13b-instruct-v2 | 8.192K | $0.6 | $0.6 | — |