No provider description is available for this model yet.
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
No provider description is available for this model yet.
Ling 3.0 Flash Fin is a finance-focused mixture-of-experts model from InclusionAI, built on Ling 3.0 Flash with 5.1B active parameters out of 124B total. It is designed for real-world investment...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Ox Alpha is a reasoning model designed for coding, sustained agentic work, and production workloads. It is suited for long-horizon software engineering, complex reasoning, and workflows that combine text with...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Qwen3-Coder-Next is an open-weight causal language model optimized for coding agents and local development workflows. It uses a sparse MoE design with 80B total parameters and only 3B activated per...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Amazon Nova Premier is the most capable of Amazon’s multimodal models for complex reasoning tasks and for use as the best teacher for distilling custom models.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
A cost-efficient version of GPT Audio. The new snapshot features an upgraded decoder for more natural sounding voices and maintains better voice consistency. Input is priced at $0.60 per million...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
MiniMax-M3 is a multimodal foundation model from MiniMax. It supports text, image, and video inputs with text output, a 1M-token context window, and is suited for long-horizon agentic work, coding,...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| nvidia/Kimi-K2.7-Code-DFlashnvidia/Kimi-K2.7-Code-DFlash | Not documented | — | — | — | |||
| nvidia/Nemotron-Labs-Audex-30B-A3Bnvidia/Nemotron-Labs-Audex-30B-A3B | Not documented | — | — | — | |||
| inclusionAI: Ling 3.0 Flash Fin (free)inclusionai/ling-3.0-flash-fin:free | 262.144K | Free | Free | — | |||
| global/gpt-4o-2024-08-06azure/global/gpt-4o-2024-08-06 | 128K | $2.5 | $10 | — | |||
| global/gpt-4o-2024-11-20azure/global/gpt-4o-2024-11-20 | 128K | $2.5 | $10 | — | |||
| nvidia/Privasis-Cleaner-4Bnvidia/Privasis-Cleaner-4B | Not documented | — | — | — | |||
| nvidia/Privasis-Cleaner-0.6Bnvidia/Privasis-Cleaner-0.6B | Not documented | — | — | — | |||
| nvidia/LocateAnything-3Bnvidia/LocateAnything-3B | Not documented | — | — | — | |||
| nvidia/Nemotron-3-Content-Safetynvidia/Nemotron-3-Content-Safety | Not documented | — | — | — | |||
| nvidia/Kimi-K2.6-Eagle3nvidia/Kimi-K2.6-Eagle3 | Not documented | — | — | — | |||
| Ox Alphastealth/ox-alpha | 1.04858M | — | — | — | |||
| nvidia/CUDA-Autocompletenvidia/CUDA-Autocomplete | Not documented | — | — | — | |||
| google.gemma-3-12b-itbedrock_converse/google.gemma-3-12b-it | 128K | $0.09 | $0.29 | — | |||
| global/gpt-5.1-chatazure/global/gpt-5.1-chat | 128K | $1.25 | $10 | — | |||
| nvidia/NVIDIA-Nemotron-3-Ultra-550B-A55B-GenRMnvidia/NVIDIA-Nemotron-3-Ultra-550B-A55B-GenRM | Not documented | — | — | — | |||
| nvidia/Nemotron-Labs-Diffusion-VLM-8Bnvidia/Nemotron-Labs-Diffusion-VLM-8B | Not documented | — | — | — | |||
| zai-org/GLM-4.7-FP8gmi/zai-org/glm-4.7-fp8 | 202.752K | $0.4 | $2 | — | |||
| nvidia/Nemotron-Labs-Diffusion-8B-Basenvidia/Nemotron-Labs-Diffusion-8B-Base | Not documented | — | — | — | |||
| Qwen: Qwen3 Coder Nextqwen/qwen3-coder-next | 262.144K | $0.12 | $0.8 | — | |||
| nvidia/Nemotron-Labs-Diffusion-14Bnvidia/Nemotron-Labs-Diffusion-14B | Not documented | — | — | — | |||
| nvidia/Nemotron-Labs-Diffusion-3Bnvidia/Nemotron-Labs-Diffusion-3B | Not documented | — | — | — | |||
| nvidia/Nemotron-Labs-Diffusion-8Bnvidia/Nemotron-Labs-Diffusion-8B | Not documented | — | — | — | |||
| Amazon: Nova Premier 1.0amazon/nova-premier-v1 | 1M | $2.5 | $12.5 | — | |||
| Qwen/Qwen3-VL-235B-A22B-Instruct-FP8gmi/qwen/qwen3-vl-235b-a22b-instruct-fp8 | 262.144K | $0.3 | $1.4 | — | |||
| us-gov-east-1/meta.llama3-8b-instruct-v1:0bedrock/us-gov-east-1/meta.llama3-8b-instruct-v1:0 | 8K | $0.3 | $2.65 | — | |||
| stabilityai/stable-diffusion-3-medium-tensorrtstabilityai/stable-diffusion-3-medium-tensorrt | Not documented | — | — | — | |||
| OpenAI: GPT Audio Miniopenai/gpt-audio-mini | 128K | $0.6 | $2.4 | — | |||
| stabilityai/stable-diffusion-3.5-medium-tensorrtstabilityai/stable-diffusion-3.5-medium-tensorrt | Not documented | — | — | — | |||
| stabilityai/stable-diffusion-3.5-controlnets-tensorrtstabilityai/stable-diffusion-3.5-controlnets-tensorrt | Not documented | — | — | — | |||
| @cf/google/gemma-7b-it-loracloudflare/@cf/google/gemma-7b-it-lora | 3.5K | — | — | — | |||
| gpt-3.5-turbo-0125azure/gpt-3.5-turbo-0125 | 16.384K | $0.5 | $1.5 | — | |||
| gpt-3.5-turbo-instruct-0914azure_text/gpt-3.5-turbo-instruct-0914 | 4.097K | $1.5 | $2 | — | |||
| gpt-35-turboazure/gpt-35-turbo | 4.097K | $0.5 | $1.5 | — | |||
| gpt-35-turbo-0125azure/gpt-35-turbo-0125 | 16.384K | $0.5 | $1.5 | — | |||
| stabilityai/stable-diffusion-3.5-large-tensorrtstabilityai/stable-diffusion-3.5-large-tensorrt | Not documented | — | — | — | |||
| stabilityai/stable-diffusion-3-medium_amdgpustabilityai/stable-diffusion-3-medium_amdgpu | Not documented | — | — | — | |||
| stabilityai/stable-diffusion-3.5-medium_amdgpustabilityai/stable-diffusion-3.5-medium_amdgpu | Not documented | — | — | — | |||
| stabilityai/stable-diffusion-3.5-large-turbo_amdgpustabilityai/stable-diffusion-3.5-large-turbo_amdgpu | Not documented | — | — | — | |||
| stabilityai/sdxl-turbo_amdgpustabilityai/sdxl-turbo_amdgpu | Not documented | — | — | — | |||
| MiniMax: MiniMax M3minimax/minimax-m3 | 524.288K | $0.3 | $1.2 | — | |||
| gpt-35-turbo-1106azure/gpt-35-turbo-1106 | 16.384K | $1 | $2 | — | |||
| stabilityai/stable-diffusion-3.5-large_amdgpustabilityai/stable-diffusion-3.5-large_amdgpu | Not documented | — | — | — | |||
| stabilityai/stable-diffusion-3.5-controlnetsstabilityai/stable-diffusion-3.5-controlnets | Not documented | — | — | — | |||
| stabilityai/sdxl-turbo-ryzen-aistabilityai/sdxl-turbo-ryzen-ai | Not documented | — | — | — | |||
| stabilityai/ar-stablelm-2-basestabilityai/ar-stablelm-2-base | Not documented | — | — | — | |||
| gpt-35-turbo-16kazure/gpt-35-turbo-16k | 16.385K | $3 | $4 | — | |||
| gpt-35-turbo-16k-0613azure/gpt-35-turbo-16k-0613 | 16.385K | $3 | $4 | — | |||
| gpt-35-turbo-instructazure_text/gpt-35-turbo-instruct | 4.097K | $1.5 | $2 | — | |||
| deepseek-ai/DeepSeek-V3-0324baseten/deepseek-ai/deepseek-v3-0324 | Not documented | $0.77 | $0.77 | — | |||
| zai-org/GLM-5.2friendliai/zai-org/glm-5.2 | 1.04858M | $1.4 | $4.4 | — |