Laguna S 2.1 is the latest coding agent model from [Poolside](<https://poolside.ai/>). Laguna S 2.1 is a 118B total parameter model with 8B active parameters, scoring 70.2% on Terminal-Bench 2.1 and...
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
No provider description is available for this model yet.
No provider description is available for this model yet.
Qwen-Plus, based on the Qwen2.5 foundation model, is a 131K context model with a balanced performance, speed, and cost combination.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
GLM 5.2 is a large-scale reasoning model from Z.ai. It supports text input and output with a 1M-token context window, and is suited for long-horizon agent workflows, project-level software engineering,...
DeepSeek-V3.1 Terminus is an update to [DeepSeek V3.1](/deepseek/deepseek-chat-v3.1) that maintains the model's original capabilities while addressing issues reported by users, including language consistency and agent capabilities, further optimizing the model's...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
This model always redirects to the latest model in the Claude Haiku family.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
NVIDIA Nemotron™ 3 Nano Omni is a 30B-A3B open multimodal model designed to function as a perception and context sub-agent in enterprise agent systems. It accepts text, image, video, and...
Solar Pro 3 is Upstage's powerful Mixture-of-Experts (MoE) language model. With 102B total parameters and 12B active parameters per forward pass, it delivers exceptional performance while maintaining computational efficiency. Optimized...
Ling 3.0 Flash VL builds on Ling 3.0 Flash (124B total / 5.5B active MoE from InclusionAI), further strengthening its language capabilities while adding native visual perception and advanced visual...
No provider description is available for this model yet.
No provider description is available for this model yet.
The o1 series of models are trained with reinforcement learning to think before they answer and perform complex reasoning. The o1-pro model uses more compute to think harder and provide...
No provider description is available for this model yet.
No provider description is available for this model yet.
Gemini 3.1 Flash Lite is Google’s GA high-efficiency multimodal model optimized for low-latency, high-volume workloads. It supports text, image, video, audio, and PDF inputs, and is designed for lightweight agentic...
No provider description is available for this model yet.
No provider description is available for this model yet.
Llama Guard 4 is a Llama 4 Scout-derived multimodal pretrained model, fine-tuned for content safety classification. Similar to previous versions, it can be used to classify content in both LLM...
The gpt-audio model is OpenAI's first generally available audio model. The new snapshot features an upgraded decoder for more natural sounding voices and maintains better voice consistency. Audio is priced...
North Mini Code is Cohere's first agentic coding model and the debut of its North family. A sparse mixture-of-experts model with 30B total parameters and 3B active, it is optimized...
No provider description is available for this model yet.
No provider description is available for this model yet.
OpenAI o4-mini-high is the same model as [o4-mini](/openai/o4-mini) with reasoning_effort set to high. OpenAI o4-mini is a compact reasoning model in the o-series, optimized for fast, cost-efficient performance while retaining...
Gemini 3.5 Flash is Google's high-efficiency multimodal model, bringing near-Pro level coding and reasoning at Flash-tier cost and speed. It is highly optimized for coding proficiency and parallel agentic execution...
No provider description is available for this model yet.
GPT-5.4 is OpenAI’s latest frontier model, unifying the Codex and GPT lines into a single system. It features a 1M+ token context window (922K input, 128K output) with support for...
Claude Opus 4.8 is Anthropic's most capable generally available model in the Opus family. It supports text, image, and file inputs with text output, with reasoning support and a 1M-token...
Muse Spark 1.2 contributor tier is a reasoning model from Meta designed for developers who want to start building at an even lower cost. It’s meaningfully cheaper than Muse Spark...
Qwen3.7-Plus is a cost-effective model in Alibaba's Qwen3.7 series. It supports text and image input with text output, building on the series' text capabilities with a comprehensive upgrade to its...
Schematron V2 Turbo is a 3B-parameter HTML-to-JSON extraction model from Inference.net. It prioritizes throughput for high-volume extraction workloads. Extraction instructions must be supplied through a JSON schema in response_format rather...
Sonnet 5 is Anthropic's most capable Sonnet-class model, with frontier performance across coding, agents, and professional work. It supports adaptive thinking with selectable reasoning effort levels (low, medium, high, max,...
The Qwen3.5 native vision-language series Plus models are built on a hybrid architecture that integrates linear attention mechanisms with sparse mixture-of-experts models, achieving higher inference efficiency. In a variety of...
Ling 3.0 Tiny is a mixture-of-experts model from InclusionAI, with 1.3B active parameters out of 7.9B total. It is designed for responsive agents, instruction following, and multi-turn conversations, with switchable...
GPT-5.6 Sol is the flagship model in OpenAI's GPT-5.6 series. It is suited for complex reasoning, coding, and agentic workflows, and is particularly strong at command-line and multi-step coding tasks...
No provider description is available for this model yet.
Claude Opus 4.5 is Anthropic’s frontier reasoning model optimized for complex software engineering, agentic workflows, and long-horizon computer use. It offers strong multimodal capabilities, competitive performance across real-world coding and...
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| Poolside: Laguna S 2.1 (free)poolside/laguna-s-2.1:free | 262.144K | Free | Free | — | |||
| gemini-3.7-flashvertex_ai-language-models/gemini-3.7-flash | 1.04858M | $0.75 | $3.75 | — | |||
| claude-3-opus-20240229anthropic/claude-3-opus-20240229 | 200K | $15 | $75 | — | |||
| Qwen: Qwen-Plusqwen/qwen-plus | 1M | $0.26 | $0.78 | — | |||
| qwen/qwen3.6-27bgroq/qwen/qwen3.6-27b | 131.072K | $0.6 | $3 | — | |||
| us-gov-west-1/anthropic.claude-sonnet-4-5-20250929-v1:0bedrock/us-gov-west-1/anthropic.claude-sonnet-4-5-20250929-v1:0 | 200K | $3.6 | $18 | — | |||
| nvidia/nemotron-3.5-lightningopenrouter/nvidia/nemotron-3.5-lightning | 262.144K | $0.05 | $0.2 | — | |||
| global.anthropic.claude-opus-4-8bedrock_converse/global.anthropic.claude-opus-4-8 | 1M | $5 | $25 | — | |||
| anthropic.claude-3-opus-20240229-v1:0bedrock/anthropic.claude-3-opus-20240229-v1:0 | 200K | $15 | $75 | — | |||
| Z.ai: GLM 5.2 (free)z-ai/glm-5.2:free | 256K | Free | Free | — | |||
| DeepSeek: DeepSeek V3.1 Terminusdeepseek/deepseek-v3.1-terminus | 131.072K | $0.27 | $1 | — | |||
| anthropic.claude-3-haiku-20240307-v1:0bedrock/anthropic.claude-3-haiku-20240307-v1:0 | 200K | $0.25 | $1.25 | — | |||
| anthropic.claude-3-7-sonnet-20250219-v1:0bedrock_converse/anthropic.claude-3-7-sonnet-20250219-v1:0 | 200K | $3 | $15 | — | |||
| eu-west-1/minimax.minimax-m2.5bedrock/eu-west-1/minimax.minimax-m2.5 | 1M | $0.36 | $1.44 | — | |||
| Anthropic: Claude Haiku Latest~anthropic/claude-haiku-latest | 200K | $1 | $5 | — | |||
| anthropic.claude-3-7-sonnet-20240620-v1:0bedrock/anthropic.claude-3-7-sonnet-20240620-v1:0 | 200K | $3.6 | $18 | — | |||
| anthropic.claude-3-5-sonnet-20241022-v2:0bedrock/anthropic.claude-3-5-sonnet-20241022-v2:0 | 1M | $3 | $15 | — | |||
| eu-west-1/minimax.minimax-m2.1bedrock/eu-west-1/minimax.minimax-m2.1 | 196K | $0.36 | $1.44 | — | |||
| anthropic.claude-3-5-sonnet-20240620-v1:0bedrock/anthropic.claude-3-5-sonnet-20240620-v1:0 | 1M | $3 | $15 | — | |||
| eu-central-1/qwen.qwen3-coder-nextbedrock/eu-central-1/qwen.qwen3-coder-next | 262.144K | $0.6 | $1.44 | — | |||
| NVIDIA: Nemotron 3 Nano Omni (free)nvidia/nemotron-3-nano-omni-30b-a3b-reasoning:free | 256K | Free | Free | — | |||
| Upstage: Solar Pro 3upstage/solar-pro-3 | 131.072K | $0.15 | $0.6 | — | |||
| inclusionAI: Ling 3.0 Flash VLinclusionai/ling-3.0-flash-vl | 131.072K | $0.06 | $0.18 | — | |||
| eu-central-1/minimax.minimax-m2.5bedrock/eu-central-1/minimax.minimax-m2.5 | 1M | $0.36 | $1.44 | — | |||
| anthropic.claude-haiku-4-5@20251001bedrock_converse/anthropic.claude-haiku-4-5@20251001 | 200K | $1 | $5 | — | |||
| OpenAI: o1-pro (batch)openai/o1-pro:batch | 200K | $75 | $300 | — | |||
| eu-central-1/minimax.minimax-m2.1bedrock/eu-central-1/minimax.minimax-m2.1 | 196K | $0.36 | $1.44 | — | |||
| anthropic.claude-haiku-4-5-20251001-v1:0bedrock_converse/anthropic.claude-haiku-4-5-20251001-v1:0 | 200K | $1 | $5 | — | |||
| Google: Gemini 3.1 Flash Lite (batch)google/gemini-3.1-flash-lite:batch | 1.04858M | $0.125 | $0.75 | — | |||
| @cf/zai-org/glm-5.2cloudflare/@cf/zai-org/glm-5.2 | 262.144K | $1.4 | $4.4 | — | |||
| us-gov-east-1/amazon.nova-pro-v1:0bedrock/us-gov-east-1/amazon.nova-pro-v1:0 | 300K | $0.96 | $3.84 | — | |||
| Meta: Llama Guard 4 12Bmeta-llama/llama-guard-4-12b | 163.84K | $0.18 | $0.18 | — | |||
| OpenAI: GPT Audioopenai/gpt-audio | 128K | $2.5 | $10 | — | |||
| Cohere: North Mini Code (free)cohere/north-mini-code:free | 256K | Free | Free | — | |||
| anthropic.claude-3-5-haiku-20241022-v1:0bedrock/anthropic.claude-3-5-haiku-20241022-v1:0 | 200K | $0.8 | $4 | — | |||
| deepseek-ai/DeepSeek-V3.2gmi/deepseek-ai/deepseek-v3.2 | 163.84K | $0.28 | $0.4 | — | |||
| OpenAI: o4 Mini Highopenai/o4-mini-high | 200K | $1.1 | $4.4 | — | |||
| Google: Gemini 3.5 Flash (batch)google/gemini-3.5-flash:batch | 1.04858M | $0.75 | $4.5 | — | |||
| openai/gpt-4o-minigmi/openai/gpt-4o-mini | 131.072K | $0.15 | $0.6 | — | |||
| OpenAI: GPT-5.4 (batch)openai/gpt-5.4:batch | 1.05M | $1.25 | $7.5 | — | |||
| Anthropic: Claude Opus 4.8 (batch)anthropic/claude-opus-4.8:batch | 1M | $2.5 | $12.5 | — | |||
| Meta: Muse Spark 1.2 Contributormeta/muse-spark-1.2-contributor | 1.04858M | $0.1 | $0.2 | — | |||
| Qwen: Qwen3.7 Plusqwen/qwen3.7-plus | 1M | $0.32 | $1.28 | — | |||
| Inference.net: Schematron V2 Turboinference-net/schematron-v2-turbo | 128K | $0.03 | $0.15 | — | |||
| Anthropic: Claude Sonnet 5 (batch)anthropic/claude-sonnet-5:batch | 1M | $1 | $5 | — | |||
| Qwen: Qwen3.5 Plus 2026-02-15qwen/qwen3.5-plus-02-15 | 1M | $0.26 | $1.56 | — | |||
| inclusionAI: Ling 3.0 Tiny (free)inclusionai/ling-3.0-tiny:free | 262.144K | Free | Free | — | |||
| OpenAI: GPT-5.6 Sol (batch)openai/gpt-5.6-sol:batch | 1.05M | $1 | $5 | — | |||
| amazon.nova-pro-v1:0bedrock_converse/amazon.nova-pro-v1:0 | 300K | $0.8 | $3.2 | — | |||
| Anthropic: Claude Opus 4.5 (batch)anthropic/claude-opus-4.5:batch | 200K | $2.5 | $12.5 | — |