No provider description is available for this model yet.
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
LFM2.5-1.2B-Thinking is a lightweight reasoning-focused model optimized for agentic tasks, data extraction, and RAG—while still running comfortably on edge devices. It supports long context (up to 32K tokens) and is...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Mercury 2.5 is the fastest reasoning LLM, and the latest diffusion LLM (dLLM) from Inception. Instead of generating tokens sequentially, Mercury 2.5 produces and refines multiple tokens in parallel, achieving...
LFM2.5-1.2B-Instruct is a compact, high-performance instruction-tuned model built for fast on-device AI. It delivers strong chat quality in a 1.2B parameter footprint, with efficient edge inference and broad runtime support.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Claude Sonnet 4.5 is Anthropic’s most advanced Sonnet model to date, optimized for real-world agents and coding workflows. It delivers state-of-the-art performance on coding benchmarks such as SWE-bench Verified, with...
No provider description is available for this model yet.
Qwen3-Max is an updated release built on the Qwen3 series, offering major improvements in reasoning, instruction following, multilingual support, and long-tail knowledge coverage compared to the January 2025 version. It...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
GPT-5.6 Luna is a fast, cost-efficient model in OpenAI's GPT-5.6 series. It is suited for high-volume, latency-sensitive tasks such as chat, classification, and lightweight agentic workflows, providing capable reasoning for...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Qwen3-Next-80B-A3B-Instruct is an instruction-tuned chat model in the Qwen3-Next series optimized for fast, stable responses without “thinking” traces. It targets complex tasks across reasoning, code generation, knowledge QA, and multilingual...
Mistral's cutting-edge language model for coding released end of July 2025. Codestral specializes in low-latency, high-frequency tasks such as fill-in-the-middle (FIM), code correction and test generation. [Blog Post](https://mistral.ai/news/codestral-25-08)
Qwen3-Coder-480B-A35B-Instruct is a Mixture-of-Experts (MoE) code generation model developed by the Qwen team. It is optimized for agentic coding tasks such as function calling, tool use, and long-context reasoning over...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Mistral Large 3 2512 is Mistral’s most capable model to date, featuring a sparse mixture-of-experts architecture with 41B active parameters (675B total), and released under the Apache 2.0 license.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| eu-west-1/minimax.minimax-m2.5bedrock/eu-west-1/minimax.minimax-m2.5 | 1M | $0.36 | $1.44 | — | |||
| LiquidAI: LFM2.5-1.2B-Thinking (free)liquid/lfm-2.5-1.2b-thinking:free | 32.768K | Free | Free | — | |||
| qwen/qwen2.5-7b-instructnovita/qwen/qwen2.5-7b-instruct | 32K | $0.07 | $0.07 | — | |||
| magistral-medium-1-2-2509mistral/magistral-medium-1-2-2509 | 40K | $2 | $5 | — | |||
| @cf/qwen/qwq-32bcloudflare/@cf/qwen/qwq-32b | 24K | $0.66 | $1 | — | |||
| Inception: Mercury 2.5inception/mercury-2.5 | 260K | $0.04 | $0.15 | — | |||
| LiquidAI: LFM2.5-1.2B-Instruct (free)liquid/lfm-2.5-1.2b-instruct:free | 32.768K | Free | Free | — | |||
| eu-west-1/qwen.qwen3-coder-nextbedrock/eu-west-1/qwen.qwen3-coder-next | 262.144K | $0.6 | $1.44 | — | |||
| eu-west-2/meta.llama3-70b-instruct-v1:0bedrock/eu-west-2/meta.llama3-70b-instruct-v1:0 | 8.192K | $3.45 | $4.55 | — | |||
| eu-west-2/meta.llama3-8b-instruct-v1:0bedrock/eu-west-2/meta.llama3-8b-instruct-v1:0 | 8.192K | $0.39 | $0.78 | — | |||
| eu-west-2/minimax.minimax-m2.1bedrock/eu-west-2/minimax.minimax-m2.1 | 196K | $0.47 | $1.86 | — | |||
| eu-west-2/minimax.minimax-m2.5bedrock/eu-west-2/minimax.minimax-m2.5 | 1M | $0.47 | $1.86 | — | |||
| eu-west-2/qwen.qwen3-coder-nextbedrock/eu-west-2/qwen.qwen3-coder-next | 262.144K | $0.78 | $1.86 | — | |||
| eu-west-3/mistral.mistral-7b-instruct-v0:2bedrock/eu-west-3/mistral.mistral-7b-instruct-v0:2 | 32K | $0.2 | $0.26 | — | |||
| eu-west-3/mistral.mistral-large-2402-v1:0bedrock/eu-west-3/mistral.mistral-large-2402-v1:0 | 32K | $10.4 | $31.2 | — | |||
| eu-west-3/mistral.mixtral-8x7b-instruct-v0:1bedrock/eu-west-3/mistral.mixtral-8x7b-instruct-v0:1 | 32K | $0.59 | $0.91 | — | |||
| microsoft/Dayhoff-170M-GRS-SS-86000microsoft/Dayhoff-170M-GRS-SS-86000 | Not documented | — | — | — | |||
| eu-south-1/minimax.minimax-m2.1bedrock/eu-south-1/minimax.minimax-m2.1 | 196K | $0.36 | $1.44 | — | |||
| eu-south-1/minimax.minimax-m2.5bedrock/eu-south-1/minimax.minimax-m2.5 | 1M | $0.36 | $1.44 | — | |||
| eu-south-1/qwen.qwen3-coder-nextbedrock/eu-south-1/qwen.qwen3-coder-next | 262.144K | $0.6 | $1.44 | — | |||
| llama3.2-1bsnowflake/llama3.2-1b | 128K | — | — | — | |||
| databricks-claude-sonnet-5databricks/databricks-claude-sonnet-5 | 1M | $3 | $15 | — | |||
| Anthropic: Claude Sonnet 4.5 (batch)anthropic/claude-sonnet-4.5:batch | 1M | $1.5 | $7.5 | — | |||
| zai-glm-4.6cerebras/zai-glm-4.6 | 128K | $2.25 | $2.75 | — | |||
| Qwen: Qwen3 Maxqwen/qwen3-max | 262.144K | $0.78 | $3.9 | — | |||
| upstage/solar-pro-preview-instructupstage/solar-pro-preview-instruct | Not documented | — | — | — | |||
| qwen/qwen3-4b-fp8novita/qwen/qwen3-4b-fp8 | 128K | $0.03 | $0.03 | — | |||
| qwen/qwen3-8b-fp8novita/qwen/qwen3-8b-fp8 | 128K | $0.035 | $0.138 | — | |||
| OpenAI: GPT-5.6 Luna (batch)openai/gpt-5.6-luna:batch | 1.05M | $0.1 | $0.6 | — | |||
| gemini-robotics-er-2-previewgemini/gemini-robotics-er-2-preview | 131.072K | $2 | $10 | — | |||
| gemini-robotics-er-1.6-previewgemini/gemini-robotics-er-1.6-preview | 131.072K | $1 | $5 | — | |||
| nvidia/Ising-Calibration-1-35B-A3Bnvidia/Ising-Calibration-1-35B-A3B | Not documented | — | — | — | |||
| invoke/anthropic.claude-3-5-sonnet-20240620-v1:0bedrock/invoke/anthropic.claude-3-5-sonnet-20240620-v1:0 | 200K | $3 | $15 | — | |||
| Qwen: Qwen3 Next 80B A3B Instruct (free)qwen/qwen3-next-80b-a3b-instruct:free | 262.144K | Free | Free | — | |||
| Mistral: Codestral 2508mistralai/codestral-2508 | 256K | $0.3 | $0.9 | — | |||
| Qwen: Qwen3 Coder 480B A35Bqwen/qwen3-coder | 262.144K | $0.3 | $1 | — | |||
| magistral-medium-2509mistral/magistral-medium-2509 | 40K | $2 | $5 | — | |||
| baidu/ernie-4.5-21B-a3bnovita/baidu/ernie-4.5-21b-a3b | 120K | $0.07 | $0.28 | — | |||
| baidu/ernie-4.5-vl-28b-a3bnovita/baidu/ernie-4.5-vl-28b-a3b | 30K | $0.14 | $0.56 | — | |||
| Mistral: Mistral Large 3 2512 (batch)mistralai/mistral-large-2512:batch | 262.144K | $0.25 | $0.75 | — | |||
| sa-east-1/meta.llama3-70b-instruct-v1:0bedrock/sa-east-1/meta.llama3-70b-instruct-v1:0 | 8.192K | $4.45 | $5.88 | — | |||
| sa-east-1/meta.llama3-8b-instruct-v1:0bedrock/sa-east-1/meta.llama3-8b-instruct-v1:0 | 8.192K | $0.5 | $1.01 | — | |||
| llama2:70bollama/llama2:70b | 4.096K | — | — | — | |||
| sa-east-1/deepseek.v3.2bedrock/sa-east-1/deepseek.v3.2 | 163.84K | $0.74 | $2.22 | — | |||
| llama2:7bollama/llama2:7b | 4.096K | — | — | — | |||
| llama3ollama/llama3 | 8.192K | — | — | — | |||
| magistral-medium-2506mistral/magistral-medium-2506 | 40K | $2 | $5 | — | |||
| google/gemma-4-31B-itfriendliai/google/gemma-4-31b-it | 262.144K | $0.14 | $0.4 | — | |||
| zai-org/GLM-5.2friendliai/zai-org/glm-5.2 | 1.04858M | $1.4 | $4.4 | — | |||
| @cf/meta/llama-4-scout-17b-16e-instructcloudflare/@cf/meta/llama-4-scout-17b-16e-instruct | 131K | $0.27 | $0.85 | — |