No provider description is available for this model yet.
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Qwen3 Coder Flash is Alibaba's fast and cost efficient version of their proprietary Qwen3 Coder Plus. It is a powerful coding agent model specializing in autonomous programming via tool calling...
Euryale L3.1 70B v2.2 is a model focused on creative roleplay from [Sao10k](https://ko-fi.com/sao10k). It is the successor of [Euryale L3 70B v2.1](/models/sao10k/l3-euryale-70b).
Mistral Saba is a 24B-parameter language model specifically designed for the Middle East and South Asia, delivering accurate and contextually relevant responses while maintaining efficient performance. Trained on curated regional...
Morph's high-accuracy apply model for complex code edits. ~4,500 tokens/sec with 98% accuracy for precise code transformations. The model requires the prompt to be in the following format: <instruction>{instruction}</instruction> <code>{initial_code}</code>...
No provider description is available for this model yet.
Qwen3-Coder-480B-A35B-Instruct is a Mixture-of-Experts (MoE) code generation model developed by the Qwen team. It is optimized for agentic coding tasks such as function calling, tool use, and long-context reasoning over...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
The largest model in the Ministral 3 family, Ministral 3 14B offers frontier capabilities and performance comparable to its larger Mistral Small 3.2 24B counterpart. A powerful and efficient language...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Venice Uncensored Dolphin Mistral 24B Venice Edition is a fine-tuned variant of Mistral-Small-24B-Instruct-2501, developed by dphn.ai in collaboration with Venice.ai. This model is designed as an “uncensored” instruct-tuned LLM, preserving...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
NVIDIA Nemotron 3 Nano 30B A3B is a small language MoE model with highest compute efficiency and accuracy for developers to build specialized agentic AI systems. The model is fully...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| us-gov-west-1/anthropic.claude-opus-5bedrock/us-gov-west-1/anthropic.claude-opus-5 | 1M | $6 | $30 | — | |||
| us.openai.gpt-5.6-terrabedrock_converse/us.openai.gpt-5.6-terra | 1M | $2.2 | $13.2 | — | |||
| @cf/nvidia/nemotron-3-120b-a12bcloudflare/@cf/nvidia/nemotron-3-120b-a12b | 256K | $0.5 | $1.5 | — | |||
| us-gov-west-1/nvidia.nemotron-nano-9b-v2bedrock/us-gov-west-1/nvidia.nemotron-nano-9b-v2 | 128K | $0.072 | $0.276 | — | |||
| eu-central-1/1-month-commitment/anthropic.claude-instant-v1bedrock/eu-central-1/1-month-commitment/anthropic.claude-instant-v1 | 100K | — | — | — | |||
| global.openai.gpt-5.6-solbedrock_converse/global.openai.gpt-5.6-sol | 1M | $5 | $30 | — | |||
| google/gemma-4-12B-it-assistantgoogle/gemma-4-12B-it-assistant | Not documented | — | — | — | |||
| muse-glimmer-30bfireworks_ai/muse-glimmer-30b | 131.072K | $0.35 | $1.5 | — | |||
| deepseek-ai/ESFT-gate-intent-litedeepseek-ai/ESFT-gate-intent-lite | Not documented | — | — | — | |||
| us-gov-west-1/amazon.titan-text-premier-v1:0bedrock/us-gov-west-1/amazon.titan-text-premier-v1:0 | 42K | $0.5 | $1.5 | — | |||
| qwen3p8-maxfireworks_ai/qwen3p8-max | 262.144K | $2 | $6 | — | |||
| eu-north-1/moonshotai.kimi-k2.5bedrock/eu-north-1/moonshotai.kimi-k2.5 | 262.144K | $0.72 | $3.6 | — | |||
| Qwen: Qwen3 Coder Flashqwen/qwen3-coder-flash | 1M | $0.195 | $0.975 | — | |||
| Sao10K: Llama 3.1 Euryale 70B v2.2sao10k/l3.1-euryale-70b | 131.072K | $0.85 | $0.85 | — | |||
| Mistral: Sabamistralai/mistral-saba | 32.768K | $0.2 | $0.6 | — | |||
| Morph: Morph V3 Largemorph/morph-v3-large | 262.144K | $0.9 | $1.9 | — | |||
| nvidia/Cosmos3-Super-Text2Image-4Stepnvidia/Cosmos3-Super-Text2Image-4Step | Not documented | — | — | — | |||
| Qwen: Qwen3 Coder 480B A35B (free)qwen/qwen3-coder:free | 262K | Free | Free | — | |||
| kimi-k3-usfireworks_ai/kimi-k3-us | 1.04858M | $3.3 | $16.5 | — | |||
| Qwen/Qwen2.5-72B-Instruct-Turbotogether_ai/qwen/qwen2.5-72b-instruct-turbo | Not documented | — | — | — | |||
| together-ai-up-to-4btogether_ai/together-ai-up-to-4b | Not documented | $0.1 | $0.1 | — | |||
| FW-Kimi-K2.7-Codeazure_ai/fw-kimi-k2.7-code | 262.144K | $1.05 | $4.4 | — | |||
| eu-north-1/minimax.minimax-m2.5bedrock/eu-north-1/minimax.minimax-m2.5 | 1M | $0.36 | $1.44 | — | |||
| together-ai-81.1b-110btogether_ai/together-ai-81.1b-110b | Not documented | $1.8 | $1.8 | — | |||
| us-gov-west-1/amazon.titan-text-lite-v1bedrock/us-gov-west-1/amazon.titan-text-lite-v1 | 42K | $0.3 | $0.4 | — | |||
| FW-Kimi-K2.6azure_ai/fw-kimi-k2.6 | 262.144K | $1.045 | $4.4 | — | |||
| us.openai.gpt-5.6-solbedrock_converse/us.openai.gpt-5.6-sol | 1M | $5.5 | $33 | — | |||
| together-ai-8.1b-21btogether_ai/together-ai-8.1b-21b | 1K | $0.3 | $0.3 | — | |||
| kimi-k3-fastfireworks_ai/kimi-k3-fast | 1.04858M | $4.5 | $22.5 | — | |||
| eu-north-1/minimax.minimax-m2.1bedrock/eu-north-1/minimax.minimax-m2.1 | 196K | $0.36 | $1.44 | — | |||
| Mistral: Ministral 3 14B 2512mistralai/ministral-14b-2512 | 262.144K | $0.2 | $0.2 | — | |||
| google/gemma-4-12B-itgoogle/gemma-4-12B-it | Not documented | — | — | — | |||
| us-gov-west-1/amazon.titan-text-express-v1bedrock/us-gov-west-1/amazon.titan-text-express-v1 | 42K | $1.3 | $1.7 | — | |||
| us-gov.xai.grok-4.6bedrock_converse/us-gov.xai.grok-4.6 | 500K | $2.64 | $7.92 | — | |||
| Qwen3.8-Maxscx-ai/qwen3.8-max | 1M | $1.65 | $4.99 | — | |||
| us-gov.openai.gpt-oss-120b-1:0bedrock_converse/us-gov.openai.gpt-oss-120b-1:0 | 128K | $0.18 | $0.72 | — | |||
| eu-north-1/deepseek.v3.2bedrock/eu-north-1/deepseek.v3.2 | 163.84K | $0.74 | $2.22 | — | |||
| anthropic.claude-fable-5bedrock_converse/anthropic.claude-fable-5 | 1M | $10 | $50 | — | |||
| GLM-5.2scx-ai/glm-5.2 | 1.04858M | $0.61 | $1.98 | — | |||
| Venice: Uncensored (free)cognitivecomputations/dolphin-mistral-24b-venice-edition:free | 32.768K | Free | Free | — | |||
| microsoft/Mage-Flow-Basemicrosoft/Mage-Flow-Base | Not documented | — | — | — | |||
| google/gemma-4-12Bgoogle/gemma-4-12B | Not documented | — | — | — | |||
| nvidia/Riva-Translate-4B-Instruct-v2nvidia/Riva-Translate-4B-Instruct-v2 | Not documented | — | — | — | |||
| ca-central-1/meta.llama3-8b-instruct-v1:0bedrock/ca-central-1/meta.llama3-8b-instruct-v1:0 | 8.192K | $0.35 | $0.69 | — | |||
| NVIDIA: Nemotron 3 Nano 30B A3B (free)nvidia/nemotron-3-nano-30b-a3b:free | 256K | Free | Free | — | |||
| us-gov-west-1/amazon.nova-pro-v1:0bedrock/us-gov-west-1/amazon.nova-pro-v1:0 | 300K | $0.96 | $3.84 | — | |||
| gpt-audio-miniazure/gpt-audio-mini | 128K | $0.6 | $2.4 | — | |||
| us-gov.openai.gpt-oss-20b-1:0bedrock_converse/us-gov.openai.gpt-oss-20b-1:0 | 128K | $0.084 | $0.36 | — | |||
| au.anthropic.claude-opus-4-7bedrock_converse/au.anthropic.claude-opus-4-7 | 1M | $5.5 | $27.5 | — | |||
| us-gov.nvidia.nemotron-super-3-120bbedrock_converse/us-gov.nvidia.nemotron-super-3-120b | 256K | $0.18 | $0.78 | — |