No provider description is available for this model yet.
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
GLM-5V-Turbo is Z.ai’s first native multimodal agent foundation model, built for vision-based coding and agent-driven tasks. It natively handles image, video, and text inputs, excels at long-horizon planning, complex coding,...
Compared with GLM-4.5, this generation brings several key improvements: Longer context window: The context window has been expanded from 128K to 200K tokens, enabling the model to handle more complex...
No provider description is available for this model yet.
No provider description is available for this model yet.
Claude Sonnet 4.5 is Anthropic’s most advanced Sonnet model to date, optimized for real-world agents and coding workflows. It delivers state-of-the-art performance on coding benchmarks such as SWE-bench Verified, with...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Mistral Medium 3.1 is an updated version of Mistral Medium 3, which is a high-performance enterprise-grade language model designed to deliver frontier-level capabilities at significantly reduced operational cost. It balances...
Gemini 2.5 Flash is Google's state-of-the-art workhorse model, specifically designed for advanced reasoning, coding, mathematics, and scientific tasks. It includes built-in "thinking" capabilities, enabling it to provide responses with greater...
Gemini 2.5 Pro is Google’s state-of-the-art AI model designed for advanced reasoning, coding, mathematics, and scientific tasks. It employs “thinking” capabilities, enabling it to reason through responses with enhanced accuracy...
Gemini 3 Flash Preview is a high speed, high value thinking model designed for agentic workflows, multi turn chat, and coding assistance. It delivers near Pro level reasoning and tool...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
LFM2.5-1.2B-Thinking is a lightweight reasoning-focused model optimized for agentic tasks, data extraction, and RAG—while still running comfortably on edge devices. It supports long context (up to 32K tokens) and is...
No provider description is available for this model yet.
Mercury 2.5 is the fastest reasoning LLM, and the latest diffusion LLM (dLLM) from Inception. Instead of generating tokens sequentially, Mercury 2.5 produces and refines multiple tokens in parallel, achieving...
GLM-5.3 is a large-scale reasoning model from Z.ai, built for complex software engineering and long-horizon agent tasks. It supports text input and output with a 1M-token context window, and improves...
LFM2.5-1.2B-Instruct is a compact, high-performance instruction-tuned model built for fast on-device AI. It delivers strong chat quality in a 1.2B parameter footprint, with efficient edge inference and broad runtime support.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| ap-southeast-2/minimax.minimax-m2.5bedrock/ap-southeast-2/minimax.minimax-m2.5 | 1M | $0.309 | $1.236 | — | |||
| Z.ai: GLM 5V Turboz-ai/glm-5v-turbo | 202.752K | $1.2 | $4 | — | |||
| Z.ai: GLM 4.6z-ai/glm-4.6 | 198K | $0.43 | $1.75 | — | |||
| ap-southeast-3/deepseek.v3.2bedrock/ap-southeast-3/deepseek.v3.2 | 163.84K | $0.74 | $2.22 | — | |||
| ap-southeast-3/minimax.minimax-m2.1bedrock/ap-southeast-3/minimax.minimax-m2.1 | 196K | $0.36 | $1.44 | — | |||
| Anthropic: Claude Sonnet 4.5 (batch)anthropic/claude-sonnet-4.5:batch | 1M | $1.5 | $7.5 | — | |||
| ap-southeast-3/minimax.minimax-m2.5bedrock/ap-southeast-3/minimax.minimax-m2.5 | 1M | $0.36 | $1.44 | — | |||
| ap-southeast-3/moonshotai.kimi-k2.5bedrock/ap-southeast-3/moonshotai.kimi-k2.5 | 262.144K | $0.72 | $3.6 | — | |||
| ap-southeast-3/qwen.qwen3-coder-nextbedrock/ap-southeast-3/qwen.qwen3-coder-next | 262.144K | $0.6 | $1.44 | — | |||
| ca-central-1/meta.llama3-70b-instruct-v1:0bedrock/ca-central-1/meta.llama3-70b-instruct-v1:0 | 8.192K | $3.05 | $4.03 | — | |||
| accounts/fireworks/models/kimi-k2-instruct-0905fireworks_ai/accounts/fireworks/models/kimi-k2-instruct-0905 | 262.144K | $0.6 | $2.5 | — | |||
| ca-central-1/meta.llama3-8b-instruct-v1:0bedrock/ca-central-1/meta.llama3-8b-instruct-v1:0 | 8.192K | $0.35 | $0.69 | — | |||
| eu-north-1/deepseek.v3.2bedrock/eu-north-1/deepseek.v3.2 | 163.84K | $0.74 | $2.22 | — | |||
| eu-north-1/minimax.minimax-m2.1bedrock/eu-north-1/minimax.minimax-m2.1 | 196K | $0.36 | $1.44 | — | |||
| eu-north-1/minimax.minimax-m2.5bedrock/eu-north-1/minimax.minimax-m2.5 | 1M | $0.36 | $1.44 | — | |||
| eu-north-1/moonshotai.kimi-k2.5bedrock/eu-north-1/moonshotai.kimi-k2.5 | 262.144K | $0.72 | $3.6 | — | |||
| eu-central-1/1-month-commitment/anthropic.claude-instant-v1bedrock/eu-central-1/1-month-commitment/anthropic.claude-instant-v1 | 100K | — | — | — | |||
| eu-central-1/1-month-commitment/anthropic.claude-v1bedrock/eu-central-1/1-month-commitment/anthropic.claude-v1 | 100K | — | — | — | |||
| eu-central-1/1-month-commitment/anthropic.claude-v2:1bedrock/eu-central-1/1-month-commitment/anthropic.claude-v2:1 | 100K | — | — | — | |||
| Mistral: Mistral Medium 3.1 (batch)mistralai/mistral-medium-3.1:batch | 131.072K | $0.2 | $1 | — | |||
| Google: Gemini 2.5 Flash (batch)google/gemini-2.5-flash:batch | 1.04858M | $0.15 | $1.25 | — | |||
| Google: Gemini 2.5 Pro Preview 05-06google/gemini-2.5-pro-preview-05-06 | 1.04858M | $1.25 | $10 | — | |||
| Google: Gemini 3 Flash Preview (batch)google/gemini-3-flash-preview:batch | 1.04858M | $0.25 | $1.5 | — | |||
| qwen/qwen3.5-122b-a10bnovita/qwen/qwen3.5-122b-a10b | 262.144K | $0.4 | $3.2 | — | |||
| openai/gpt-3.5-turbo-16kopenrouter/openai/gpt-3.5-turbo-16k | 16.385K | $3 | $4 | — | |||
| qwen/qwen3.5-27bnovita/qwen/qwen3.5-27b | 262.144K | $0.3 | $2.4 | — | |||
| eu-central-1/6-month-commitment/anthropic.claude-instant-v1bedrock/eu-central-1/6-month-commitment/anthropic.claude-instant-v1 | 100K | — | — | — | |||
| minimax/minimax-m2.5-highspeednovita/minimax/minimax-m2.5-highspeed | 204.8K | $0.6 | $2.4 | — | |||
| eu-central-1/6-month-commitment/anthropic.claude-v1bedrock/eu-central-1/6-month-commitment/anthropic.claude-v1 | 100K | — | — | — | |||
| eu-central-1/6-month-commitment/anthropic.claude-v2:1bedrock/eu-central-1/6-month-commitment/anthropic.claude-v2:1 | 100K | — | — | — | |||
| eu-central-1/anthropic.claude-instant-v1bedrock/eu-central-1/anthropic.claude-instant-v1 | 100K | $2.48 | $8.38 | — | |||
| eu-central-1/anthropic.claude-v1bedrock/eu-central-1/anthropic.claude-v1 | 100K | $8 | $24 | — | |||
| eu-central-1/anthropic.claude-v2:1bedrock/eu-central-1/anthropic.claude-v2:1 | 100K | $8 | $24 | — | |||
| @cf/zai-org/glm-5.2cloudflare/@cf/zai-org/glm-5.2 | 262.144K | $1.4 | $4.4 | — | |||
| openai/gpt-3.5-turboopenrouter/openai/gpt-3.5-turbo | 16.385K | $1.5 | $2 | — | |||
| eu-central-1/minimax.minimax-m2.1bedrock/eu-central-1/minimax.minimax-m2.1 | 196K | $0.36 | $1.44 | — | |||
| accounts/fireworks/models/kimi-k2-instructfireworks_ai/accounts/fireworks/models/kimi-k2-instruct | 131.072K | $0.6 | $2.5 | — | |||
| eu-central-1/qwen.qwen3-coder-nextbedrock/eu-central-1/qwen.qwen3-coder-next | 262.144K | $0.6 | $1.44 | — | |||
| eu-west-1/meta.llama3-70b-instruct-v1:0bedrock/eu-west-1/meta.llama3-70b-instruct-v1:0 | 8.192K | $2.86 | $3.78 | — | |||
| eu-west-1/meta.llama3-8b-instruct-v1:0bedrock/eu-west-1/meta.llama3-8b-instruct-v1:0 | 8.192K | $0.32 | $0.65 | — | |||
| eu-west-1/minimax.minimax-m2.1bedrock/eu-west-1/minimax.minimax-m2.1 | 196K | $0.36 | $1.44 | — | |||
| eu-west-1/minimax.minimax-m2.5bedrock/eu-west-1/minimax.minimax-m2.5 | 1M | $0.36 | $1.44 | — | |||
| LiquidAI: LFM2.5-1.2B-Thinking (free)liquid/lfm-2.5-1.2b-thinking:free | 32.768K | Free | Free | — | |||
| minimax/minimax-m2.7novita/minimax/minimax-m2.7 | 204.8K | $0.3 | $1.2 | — | |||
| Inception: Mercury 2.5inception/mercury-2.5 | 260K | $0.04 | $0.15 | — | |||
| Z.ai: GLM 5.3z-ai/glm-5.3 | 1.04858M | $1.4 | $4.4 | — | |||
| LiquidAI: LFM2.5-1.2B-Instruct (free)liquid/lfm-2.5-1.2b-instruct:free | 32.768K | Free | Free | — | |||
| eu-west-1/qwen.qwen3-coder-nextbedrock/eu-west-1/qwen.qwen3-coder-next | 262.144K | $0.6 | $1.44 | — | |||
| eu-west-2/meta.llama3-70b-instruct-v1:0bedrock/eu-west-2/meta.llama3-70b-instruct-v1:0 | 8.192K | $3.45 | $4.55 | — | |||
| moonshotai/kimi-k2.5openrouter/moonshotai/kimi-k2.5 | 262.144K | $0.6 | $3 | — |