No provider description is available for this model yet.
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
No provider description is available for this model yet.
Uncensored and creative writing model based on Mistral Small 3.2 24B with good recall, prompt adherence, and intelligence.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Claude Opus 5 is Anthropic’s flagship model for demanding reasoning, coding, and long-horizon agentic work. It is particularly strong at end-to-end software tasks, code review and bug finding, visual analysis...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
The Meta Llama 3.3 multilingual large language model (LLM) is a pretrained and instruction tuned generative model in 70B (text in/text out). The Llama 3.3 instruction tuned text only model...
GLM 5.2 is a large-scale reasoning model from Z.ai. It supports text input and output with a 1M-token context window, and is suited for long-horizon agent workflows, project-level software engineering,...
No provider description is available for this model yet.
LFM2.5-2.6B is a compact reasoning model from Liquid AI. It is suited for agent workflows, data extraction, RAG, and long-context processing. Liquid advises against using it for agentic coding or...
No provider description is available for this model yet.
Qwen3-Coder-30B-A3B-Instruct is a 30.5B parameter Mixture-of-Experts (MoE) model with 128 experts (8 active per forward pass), designed for advanced code generation, repository-scale understanding, and agentic tool use. Built on the...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
As a 30B-class SOTA model, GLM-4.7-Flash offers a new option that balances performance and efficiency. It is further optimized for agentic coding use cases, strengthening coding capabilities, long-horizon task planning,...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Qwen3-VL-8B-Thinking is the reasoning-optimized variant of the Qwen3-VL-8B multimodal model, designed for advanced visual and textual reasoning across complex scenes, documents, and temporal sequences. It integrates enhanced multimodal alignment and...
Hunyuan-A13B is a 13B active parameter Mixture-of-Experts (MoE) language model developed by Tencent, with a total parameter count of 80B and support for reasoning via Chain-of-Thought. It offers competitive benchmark...
No provider description is available for this model yet.
No provider description is available for this model yet.
This model always redirects to the latest model in the GLM Flash family.
No provider description is available for this model yet.
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| gemini-3.7-flashvertex_ai/gemini-3.7-flash | 1.04858M | $0.75 | $3.75 | — | |||
| upstage/Solar-Open2-250Bupstage/Solar-Open2-250B | Not documented | — | — | — | |||
| TheDrummer: Cydonia 24B V4.1thedrummer/cydonia-24b-v4.1 | 131.072K | $0.3 | $0.5 | — | |||
| gemini-3.7-flashgemini/gemini-3.7-flash | 1.04858M | $0.75 | $3.75 | — | |||
| us-gov-west-1/anthropic.claude-3-haiku-20240307-v1:0bedrock/us-gov-west-1/anthropic.claude-3-haiku-20240307-v1:0 | 200K | $0.3 | $1.5 | — | |||
| swe-1.7-lightningcognition/swe-1.7-lightning | Not documented | $2.5 | $12.5 | — | |||
| microsoft/UniRG-CXRmicrosoft/UniRG-CXR | Not documented | — | — | — | |||
| eu.anthropic.claude-opus-5bedrock_converse/eu.anthropic.claude-opus-5 | 1M | $5.5 | $27.5 | — | |||
| au.anthropic.claude-opus-5bedrock_converse/au.anthropic.claude-opus-5 | 1M | $5.5 | $27.5 | — | |||
| google/gemma-4-26B-A4B-it-assistantgoogle/gemma-4-26B-A4B-it-assistant | Not documented | — | — | — | |||
| us-gov-east-1/claude-sonnet-4-5-20250929-v1:0bedrock/us-gov-east-1/claude-sonnet-4-5-20250929-v1:0 | 200K | $3.6 | $18 | — | |||
| Claude Opus 5 (batch)anthropic/claude-opus-5:batch | 1M | $2.5 | $12.5 | — | |||
| @cf/qwen/qwen3-30b-a3b-fp8cloudflare/@cf/qwen/qwen3-30b-a3b-fp8 | 32.768K | $0.051 | $0.335 | — | |||
| jp.anthropic.claude-opus-5bedrock_converse/jp.anthropic.claude-opus-5 | 1M | $5.5 | $27.5 | — | |||
| anthropic.claude-opus-4-8bedrock_converse/anthropic.claude-opus-4-8 | 1M | $5 | $25 | — | |||
| gemini-3.7-flashvertex_ai-language-models/gemini-3.7-flash | 1.04858M | $0.75 | $3.75 | — | |||
| chat-latestopenai/chat-latest | 400K | $5 | $30 | — | |||
| claude-3-opus-20240229anthropic/claude-3-opus-20240229 | 200K | $15 | $75 | — | |||
| meta-llama/llama-prompt-guard-2-86mgroq/meta-llama/llama-prompt-guard-2-86m | 512 | $0.04 | $0.04 | — | |||
| qwen/qwen3.6-27bgroq/qwen/qwen3.6-27b | 131.072K | $0.6 | $3 | — | |||
| us-gov-west-1/anthropic.claude-sonnet-4-5-20250929-v1:0bedrock/us-gov-west-1/anthropic.claude-sonnet-4-5-20250929-v1:0 | 200K | $3.6 | $18 | — | |||
| meta-llama/llama-prompt-guard-2-22mgroq/meta-llama/llama-prompt-guard-2-22m | 512 | $0.03 | $0.03 | — | |||
| nvidia/nemotron-3.5-lightningopenrouter/nvidia/nemotron-3.5-lightning | 262.144K | $0.05 | $0.2 | — | |||
| google/gemma-4-E4B-it-assistantgoogle/gemma-4-E4B-it-assistant | Not documented | — | — | — | |||
| global.anthropic.claude-opus-4-8bedrock_converse/global.anthropic.claude-opus-4-8 | 1M | $5 | $25 | — | |||
| Meta: Llama 3.3 70B Instruct (free)meta-llama/llama-3.3-70b-instruct:free | 65.536K | Free | Free | — | |||
| Z.ai: GLM 5.2 (free)z-ai/glm-5.2:free | 256K | Free | Free | — | |||
| us.anthropic.claude-opus-4-8bedrock_converse/us.anthropic.claude-opus-4-8 | 1M | $5.5 | $27.5 | — | |||
| LiquidAI: LFM2.5-2.6B (free)liquid/lfm-2.5-2.6b:free | 65.536K | Free | Free | — | |||
| us-east-1/6-month-commitment/anthropic.claude-instant-v1bedrock/us-east-1/6-month-commitment/anthropic.claude-instant-v1 | 100K | — | — | — | |||
| Qwen: Qwen3 Coder 30B A3B Instructqwen/qwen3-coder-30b-a3b-instruct | 262.144K | $0.07 | $0.28 | — | |||
| @cf/mistral/mistral-7b-instruct-v0.1cloudflare/@cf/mistral/mistral-7b-instruct-v0.1 | 8.192K | $1.923 | $1.923 | — | |||
| microsoft/Dayhoff-170M-GRS-SS-134000microsoft/Dayhoff-170M-GRS-SS-134000 | Not documented | — | — | — | |||
| google/gemma-4-E2Bgoogle/gemma-4-E2B | Not documented | — | — | — | |||
| us-east-1/1-month-commitment/anthropic.claude-v2:1bedrock/us-east-1/1-month-commitment/anthropic.claude-v2:1 | 100K | — | — | — | |||
| au.anthropic.claude-opus-4-8bedrock_converse/au.anthropic.claude-opus-4-8 | 1M | $5.5 | $27.5 | — | |||
| google/gemma-4-E4B-it-qat-w4a16-ctgoogle/gemma-4-E4B-it-qat-w4a16-ct | Not documented | — | — | — | |||
| google/gemma-4-26B-A4Bgoogle/gemma-4-26B-A4B | Not documented | — | — | — | |||
| jp.anthropic.claude-opus-4-8bedrock_converse/jp.anthropic.claude-opus-4-8 | 1M | $5.5 | $27.5 | — | |||
| Z.ai: GLM 4.7 Flashz-ai/glm-4.7-flash | 131.072K | $0.061 | $0.4 | — | |||
| us-gov-east-1/anthropic.claude-sonnet-4-5-20250929-v1:0bedrock/us-gov-east-1/anthropic.claude-sonnet-4-5-20250929-v1:0 | 200K | $3.6 | $18 | — | |||
| daybreak-blue-latestopenai/daybreak-blue-latest | 1.05M | $5 | $30 | — | |||
| google/gemma-4-26B-A4B-itgoogle/gemma-4-26B-A4B-it | Not documented | — | — | — | |||
| jp.anthropic.claude-opus-4-7bedrock_converse/jp.anthropic.claude-opus-4-7 | 1M | $5.5 | $27.5 | — | |||
| Qwen: Qwen3 VL 8B Thinkingqwen/qwen3-vl-8b-thinking | 131.072K | $0.18 | $2.1 | — | |||
| Tencent: Hunyuan A13B Instructtencent/hunyuan-a13b-instruct | 131.072K | $0.14 | $0.57 | — | |||
| global.anthropic.claude-sonnet-5bedrock_converse/global.anthropic.claude-sonnet-5 | 1M | $2 | $10 | — | |||
| google/gemma-4-31Bgoogle/gemma-4-31B | Not documented | — | — | — | |||
| Z.ai: GLM Flash Latest~z-ai/glm-flash-latest | 1.04858M | $0.075 | $0.25 | — | |||
| microsoft/Dayhoff-3b-UR90-30000microsoft/Dayhoff-3b-UR90-30000 | Not documented | — | — | — |