No provider description is available for this model yet.
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
o3 is a well-rounded and powerful model across domains. It sets a new standard for math, science, coding, and visual reasoning tasks. It also excels at technical writing and instruction-following....
Claude Fable 5.1 improves on Claude Fable 5 across the board, with the biggest gains in agentic coding, long-running agentic workflows, and knowledge work: long code refactors, front-end and visual...
GLM-5.3-Flash is a native multimodal model from Z.ai. It is suited for efficient coding and long-horizon agent tasks. Its hybrid sparse and linear attention architecture maintains accurate long-context behavior while...
Fugu Max is the cost-performance model in Sakana AI's Fugu family. Rather than a single monolithic model, Fugu is a learned multi-agent orchestration system: a language model trained to route...
GPT-5.4 Pro is OpenAI's most advanced model, building on GPT-5.4's unified architecture with enhanced reasoning capabilities for complex, high-stakes tasks. It features a 1M+ token context window (922K input, 128K...
No provider description is available for this model yet.
NVIDIA Nemotron 3.5 Lightning is an open mixture-of-experts model from NVIDIA, with 3B active parameters out of 30B total. It is suited for high-throughput agentic workloads and specialized tasks that...
Claude Fable 5 is a Mythos-class model from Anthropic, built for autonomous knowledge work and coding. It supports text, image, and file inputs with text output, with reasoning support and...
Gemini 3.8 Flash is Google's most intelligent Flash model with significant gains from 3.7 Flash across software engineering, agentic tasks, and multi-step reasoning.
Claude Opus 4.8 is Anthropic's most capable generally available model in the Opus family. It supports text, image, and file inputs with text output, with reasoning support and a 1M-token...
Claude Opus 4.8 is Anthropic's most capable generally available model in the Opus family. It supports text, image, and file inputs with text output, with reasoning support and a 1M-token...
GLM-5.1 delivers a major leap in coding capability, with particularly significant gains in handling long-horizon tasks. Unlike previous models built around minute-level interactions, GLM-5.1 can work independently and continuously on...
GPT-4.1 is a flagship large language model optimized for advanced instruction following, real-world software engineering, and long-context reasoning. It supports a 1 million token context window and outperforms GPT-4o and...
GLM-5V-Turbo is Z.ai’s first native multimodal agent foundation model, built for vision-based coding and agent-driven tasks. It natively handles image, video, and text inputs, excels at long-horizon planning, complex coding,...
GLM-5.3 is a large-scale reasoning model from Z.ai, built for complex software engineering and long-horizon agent tasks. It supports text input and output with a 1M-token context window, and improves...
GPT-5.6 Sol is the flagship model in OpenAI's GPT-5.6 series. It is suited for complex reasoning, coding, and agentic workflows, and is particularly strong at command-line and multi-step coding tasks...
Grok 4.20 is a reasoning model from SpaceXAI with industry-leading speed and agentic tool calling capabilities. It combines the lowest hallucination rate on the market with strict prompt adherance, delivering...
No provider description is available for this model yet.
No provider description is available for this model yet.
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| gemini-2.5-flash-native-audio-preview-12-2025gemini/gemini-2.5-flash-native-audio-preview-12-2025 | 1.04858M | $0.3 | $2.5 | — | |||
| gemini-pro-latestgemini/gemini-pro-latest | 1.04858M | $1.25 | $10 | — | |||
| claude-sonnet-5@defaultvertex_ai-anthropic_models/claude-sonnet-5@default | 1M | $2 | $10 | — | |||
| claude-sonnet-4-6@defaultvertex_ai-anthropic_models/claude-sonnet-4-6@default | 1M | $3 | $15 | — | |||
| google.gemma-4-31bbedrock_mantle/google.gemma-4-31b | 256K | $0.14 | $0.4 | — | |||
| google.gemma-4-26b-a4bbedrock_mantle/google.gemma-4-26b-a4b | 256K | $0.13 | $0.4 | — | |||
| doubao-seed-2-0-pro-260215volcengine/doubao-seed-2-0-pro-260215 | 256K | — | — | — | |||
| doubao-seed-2-0-lite-260215volcengine/doubao-seed-2-0-lite-260215 | 256K | — | — | — | |||
| claude-4-sonnetsnowflake/claude-4-sonnet | 200K | $3 | $15 | — | |||
| doubao-seed-2-0-mini-260215volcengine/doubao-seed-2-0-mini-260215 | 256K | — | — | — | |||
| doubao-seed-2-0-code-preview-260215volcengine/doubao-seed-2-0-code-preview-260215 | 256K | — | — | — | |||
| us-east-1/zai.glm-5bedrock/us-east-1/zai.glm-5 | 200K | $1 | $3.2 | — | |||
| us-west-2/zai.glm-5bedrock/us-west-2/zai.glm-5 | 200K | $1 | $3.2 | — | |||
| us-gov-east-1/anthropic.claude-haiku-4-5-20251001-v1:0bedrock/us-gov-east-1/anthropic.claude-haiku-4-5-20251001-v1:0 | 200K | $1.2 | $6 | — | |||
| us-gov-west-1/anthropic.claude-haiku-4-5-20251001-v1:0bedrock/us-gov-west-1/anthropic.claude-haiku-4-5-20251001-v1:0 | 200K | $1.2 | $6 | — | |||
| claude-sonnet-4-5snowflake/claude-sonnet-4-5 | 200K | $3 | $15 | — | |||
| claude-sonnet-4-6snowflake/claude-sonnet-4-6 | 200K | $3 | $15 | — | |||
| claude-4-opussnowflake/claude-4-opus | 200K | $5 | $25 | — | |||
| claude-haiku-4-5snowflake/claude-haiku-4-5 | 200K | $1 | $5 | — | |||
| claude-3-7-sonnetsnowflake/claude-3-7-sonnet | 200K | $3 | $15 | — | |||
| openai-gpt-4.1snowflake/openai-gpt-4.1 | 300K | $2 | $8 | — | |||
| openai-gpt-5snowflake/openai-gpt-5 | 300K | $1.25 | $10 | — | |||
| openai-gpt-5-minisnowflake/openai-gpt-5-mini | 1M | $0.3 | $1.2 | — | |||
| openai-gpt-5-nanosnowflake/openai-gpt-5-nano | 5M | $0.15 | $0.6 | — | |||
| Qwen/Qwen3.5-397B-A17B-FP8tensormesh/qwen/qwen3.5-397b-a17b-fp8 | 262.144K | $0.6 | $3.6 | — | |||
| Qwen/Qwen3-Coder-480B-A35B-Instruct-FP8tensormesh/qwen/qwen3-coder-480b-a35b-instruct-fp8 | 262.144K | $0.45 | $1.8 | — | |||
| Qwen/Qwen3.6-27B-FP8tensormesh/qwen/qwen3.6-27b-fp8 | 262.144K | $0.32 | $3.2 | — | |||
| lukealonso/GLM-5.1-NVFP4-MTPtensormesh/lukealonso/glm-5.1-nvfp4-mtp | 202.752K | $1.4 | $4.4 | — | |||
| deepseek-v4-protencent/deepseek-v4-pro | 1M | $0.435 | $0.87 | — | |||
| deepseek-v4-flashtencent/deepseek-v4-flash | 1M | $0.14 | $0.28 | — | |||
| ps/minimax-m2.7pinstripes/ps/minimax-m2.7 | 1.00019M | $0.255 | $0.55 | — | |||
| OpenAI: o3 (batch)openai/o3:batch | 200K | $1 | $4 | — | |||
| Anthropic: Claude Fable 5.1anthropic/claude-fable-5.1 | 1M | $10 | $50 | — | |||
| Z.ai: GLM 5.3 Flashz-ai/glm-5.3-flash | 1.04858M | $0.15 | $0.5 | — | |||
| Sakana: Fugu Maxsakana/fugu-max | 1M | $2 | $6 | — | |||
| OpenAI: GPT-5.4 Pro (batch)openai/gpt-5.4-pro:batch | 1.05M | $15 | $90 | — | |||
| openai/gpt-5.6-solopenrouter/openai/gpt-5.6-sol | 1.05M | $2 | $10 | — | |||
| NVIDIA: Nemotron 3.5 Lightning (free)nvidia/nemotron-3.5-lightning:free | 1M | Free | Free | — | |||
| Anthropic: Claude Fable 5 (batch)anthropic/claude-fable-5:batch | 1M | $5 | $25 | — | |||
| Google: Gemini 3.8 Flash (batch)google/gemini-3.8-flash:batch | 1.04858M | $0.375 | $1.875 | — | |||
| Anthropic: Claude Opus 4.8anthropic/claude-opus-4.8 | 1M | $5 | $25 | — | |||
| Anthropic: Claude Opus 4.8 (batch)anthropic/claude-opus-4.8:batch | 1M | $2.5 | $12.5 | — | |||
| Z.ai: GLM 5.1z-ai/glm-5.1 | 200K | $0.966 | $3.036 | — | |||
| OpenAI: GPT-4.1 (batch)openai/gpt-4.1:batch | 1.04758M | $1 | $4 | — | |||
| Z.ai: GLM 5V Turboz-ai/glm-5v-turbo | 202.752K | $1.2 | $4 | — | |||
| Z.ai: GLM 5.3z-ai/glm-5.3 | 1.04858M | $1.092 | $3.432 | — | |||
| OpenAI: GPT-5.6 Sol (batch)openai/gpt-5.6-sol:batch | 1.05M | $1 | $5 | — | |||
| SpaceXAI: Grok 4.20x-ai/grok-4.20 | 2M | $1.25 | $2.5 | — | |||
| databricks-glm-5-2databricks/databricks-glm-5-2 | 1M | $1.4 | $4.4 | — | |||
| databricks-kimi-k3databricks/databricks-kimi-k3 | 1M | $3 | $15 | — |