No provider description is available for this model yet.
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
GPT-5.1 is the latest frontier-grade model in the GPT-5 series, offering stronger general-purpose reasoning, improved instruction adherence, and a more natural conversational style compared to GPT-5. It uses adaptive reasoning...
Claude Opus 4.8 is Anthropic's most capable generally available model in the Opus family. It supports text, image, and file inputs with text output, with reasoning support and a 1M-token...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Muse Spark 1.2 contributor tier is a reasoning model from Meta designed for developers who want to start building at an even lower cost. It’s meaningfully cheaper than Muse Spark...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Schematron V2 Turbo is a 3B-parameter HTML-to-JSON extraction model from Inference.net. It prioritizes throughput for high-volume extraction workloads. Extraction instructions must be supplied through a JSON schema in response_format rather...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| us-gov.nvidia.nemotron-nano-9b-v2bedrock_converse/us-gov.nvidia.nemotron-nano-9b-v2 | 128K | $0.072 | $0.276 | — | |||
| gpt-35-turbo-0125azure/gpt-35-turbo-0125 | 16.384K | $0.5 | $1.5 | — | |||
| gpt-35-turboazure/gpt-35-turbo | 4.097K | $0.5 | $1.5 | — | |||
| @cf/qwen/qwen3-30b-a3b-fp8cloudflare/@cf/qwen/qwen3-30b-a3b-fp8 | 32.768K | $0.051 | $0.335 | — | |||
| gpt-3.5-turbo-instruct-0914azure_text/gpt-3.5-turbo-instruct-0914 | 4.097K | $1.5 | $2 | — | |||
| gpt-3.5-turbo-0125azure/gpt-3.5-turbo-0125 | 16.384K | $0.5 | $1.5 | — | |||
| eu.twelvelabs.pegasus-1-2-v1:0bedrock/eu.twelvelabs.pegasus-1-2-v1:0 | Not documented | — | $7.5 | — | |||
| @cf/google/gemma-7b-it-loracloudflare/@cf/google/gemma-7b-it-lora | 3.5K | — | — | — | |||
| stabilityai/stable-diffusion-3.5-controlnets-tensorrtstabilityai/stable-diffusion-3.5-controlnets-tensorrt | Not documented | — | — | — | |||
| stabilityai/stable-diffusion-3.5-medium-tensorrtstabilityai/stable-diffusion-3.5-medium-tensorrt | Not documented | — | — | — | |||
| gpt-3.5-turboazure/gpt-3.5-turbo | 4.097K | $0.5 | $1.5 | — | |||
| OpenAI: GPT-5.1 (batch)openai/gpt-5.1:batch | 400K | $0.625 | $5 | — | |||
| Anthropic: Claude Opus 4.8 (batch)anthropic/claude-opus-4.8:batch | 1M | $2.5 | $12.5 | — | |||
| stabilityai/stable-diffusion-3-medium-tensorrtstabilityai/stable-diffusion-3-medium-tensorrt | Not documented | — | — | — | |||
| stabilityai/sdxl-turbo-tensorrtstabilityai/sdxl-turbo-tensorrt | Not documented | — | — | — | |||
| au.anthropic.claude-opus-5bedrock_converse/au.anthropic.claude-opus-5 | 1M | $5.5 | $27.5 | — | |||
| databricks-claude-fable-5databricks/databricks-claude-fable-5 | 1M | $10 | $50 | — | |||
| eu.anthropic.claude-opus-5bedrock_converse/eu.anthropic.claude-opus-5 | 1M | $5.5 | $27.5 | — | |||
| Meta: Muse Spark 1.2 Contributormeta/muse-spark-1.2-contributor | 1.04858M | $0.1 | $0.2 | — | |||
| nvidia/Nemotron-Labs-Diffusion-8Bnvidia/Nemotron-Labs-Diffusion-8B | Not documented | — | — | — | |||
| nvidia/Nemotron-Labs-Diffusion-3Bnvidia/Nemotron-Labs-Diffusion-3B | Not documented | — | — | — | |||
| swe-1.7-lightningcognition/swe-1.7-lightning | Not documented | $2.5 | $12.5 | — | |||
| nvidia/Nemotron-Labs-Diffusion-14Bnvidia/Nemotron-Labs-Diffusion-14B | Not documented | — | — | — | |||
| nvidia/Nemotron-Labs-Diffusion-14B-Basenvidia/Nemotron-Labs-Diffusion-14B-Base | Not documented | — | — | — | |||
| us-gov-west-1/anthropic.claude-3-haiku-20240307-v1:0bedrock/us-gov-west-1/anthropic.claude-3-haiku-20240307-v1:0 | 200K | $0.3 | $1.5 | — | |||
| meta-llama/Llama-4-Maverick-17B-128E-Instructmeta-llama/Llama-4-Maverick-17B-128E-Instruct | Not documented | — | — | — | |||
| accounts/fireworks/models/muse-glimmer-30bfireworks_ai/accounts/fireworks/models/muse-glimmer-30b | 131.072K | $0.35 | $1.5 | — | |||
| nvidia/Nemotron-Labs-Diffusion-8B-Basenvidia/Nemotron-Labs-Diffusion-8B-Base | Not documented | — | — | — | |||
| nvidia/Nemotron-Labs-Diffusion-3B-Basenvidia/Nemotron-Labs-Diffusion-3B-Base | Not documented | — | — | — | |||
| gemini-3.7-flashgemini/gemini-3.7-flash | 1.04858M | $0.75 | $3.75 | — | |||
| nvidia/Nemotron-Labs-Diffusion-VLM-8Bnvidia/Nemotron-Labs-Diffusion-VLM-8B | Not documented | — | — | — | |||
| nvidia/NVIDIA-Nemotron-3-Ultra-550B-A55B-GenRMnvidia/NVIDIA-Nemotron-3-Ultra-550B-A55B-GenRM | Not documented | — | — | — | |||
| upstage/Solar-Open2-250Bupstage/Solar-Open2-250B | Not documented | — | — | — | |||
| Inference.net: Schematron V2 Turboinference-net/schematron-v2-turbo | 128K | $0.03 | $0.15 | — | |||
| global/gpt-5.1-chatazure/global/gpt-5.1-chat | 128K | $1.25 | $10 | — | |||
| global/gpt-5.1azure/global/gpt-5.1 | 272K | $1.25 | $10 | — | |||
| gemini-3.7-flashvertex_ai/gemini-3.7-flash | 1.04858M | $0.75 | $3.75 | — | |||
| nvidia/CUDA-Autocompletenvidia/CUDA-Autocomplete | Not documented | — | — | — | |||
| nvidia/Kimi-K2.5-Thinking-Eagle3nvidia/Kimi-K2.5-Thinking-Eagle3 | Not documented | — | — | — | |||
| nvidia/NVIDIA-Nemotron-3.5-Lightningdeepinfra/nvidia/nvidia-nemotron-3.5-lightning | 262.144K | $0.05 | $0.2 | — | |||
| us.twelvelabs.pegasus-1-2-v1:0bedrock/us.twelvelabs.pegasus-1-2-v1:0 | Not documented | — | $7.5 | — | |||
| gemini-robotics-er-2-streaming-previewgemini/gemini-robotics-er-2-streaming-preview | Not documented | $2 | $10 | — | |||
| nvidia/Kimi-K2.6-Eagle3nvidia/Kimi-K2.6-Eagle3 | Not documented | — | — | — | |||
| eu.amazon.nova-micro-v1:0bedrock_converse/eu.amazon.nova-micro-v1:0 | 128K | $0.046 | $0.184 | — | |||
| nvidia/Nemotron-3-Content-Safetynvidia/Nemotron-3-Content-Safety | Not documented | — | — | — | |||
| qwen3.8-maxdashscope/qwen3.8-max | 991.808K | $2 | $6 | — | |||
| eu.amazon.nova-lite-v1:0bedrock_converse/eu.amazon.nova-lite-v1:0 | 300K | $0.078 | $0.312 | — | |||
| kimi-k2-thinking-251104volcengine/kimi-k2-thinking-251104 | 229.376K | — | — | — | |||
| nvidia/LocateAnything-3Bnvidia/LocateAnything-3B | Not documented | — | — | — | |||
| glm-4-7-251222volcengine/glm-4-7-251222 | 204.8K | — | — | — |