NVIDIA Nemotron™ 3 Nano Omni is a 30B-A3B open multimodal model designed to function as a perception and context sub-agent in enterprise agent systems. It accepts text, image, video, and...
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
NVIDIA Nemotron 3.5 Content Safety is a compact 4B-parameter multimodal guardrail model from NVIDIA, fine-tuned from Google Gemma-3-4B. It moderates both inputs to and responses from LLMs and VLMs, accepting...
NVIDIA Nemotron 3 Nano 30B A3B is a small language MoE model with highest compute efficiency and accuracy for developers to build specialized agentic AI systems. The model is fully...
NVIDIA Nemotron Nano 2 VL is a 12-billion-parameter open multimodal reasoning model designed for video understanding and document intelligence. It introduces a hybrid Transformer-Mamba architecture, combining transformer-level accuracy with Mamba’s...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
NVIDIA Nemotron 3 Ultra is an open frontier-reasoning and orchestration model from NVIDIA, with 55B active parameters out of 550B total (MoE). Built on a hybrid Transformer-Mamba mixture-of-experts architecture, it...
NVIDIA Nemotron 3.5 Lightning is an open mixture-of-experts model from NVIDIA, with 3B active parameters out of 30B total. It is suited for high-throughput agentic workloads and specialized tasks that...
NVIDIA Nemotron 3 Ultra is an open frontier-reasoning and orchestration model from NVIDIA, with 55B active parameters out of 550B total (MoE). Built on a hybrid Transformer-Mamba mixture-of-experts architecture, it...
No provider description is available for this model yet.
NVIDIA-Nemotron-Nano-9B-v2 is a large language model (LLM) trained from scratch by NVIDIA, and designed as a unified model for both reasoning and non-reasoning tasks. It responds to user queries and...
NVIDIA Nemotron 3 Super is a 120B-parameter open hybrid MoE model, activating just 12B parameters for maximum compute efficiency and accuracy in complex multi-agent applications. Built on a hybrid Mamba-Transformer...
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| NVIDIA: Nemotron 3 Nano Omni (free)nvidia/nemotron-3-nano-omni-30b-a3b-reasoning:free | 256K | Free | Free | — | |||
| NVIDIA: Nemotron 3.5 Content Safety (free)nvidia/nemotron-3.5-content-safety:free | 128K | Free | Free | — | |||
| NVIDIA: Nemotron 3 Nano 30B A3B (free)nvidia/nemotron-3-nano-30b-a3b:free | 256K | Free | Free | — | |||
| NVIDIA: Nemotron Nano 12B 2 VL (free)nvidia/nemotron-nano-12b-v2-vl:free | 128K | Free | Free | — | |||
| nvidia/Riva-Translate-4B-Instruct-v1.1nvidia/Riva-Translate-4B-Instruct-v1.1 | Not documented | — | — | — | |||
| nvidia/Nemotron-3-Labs-Ultra-Math-SFTnvidia/Nemotron-3-Labs-Ultra-Math-SFT | Not documented | — | — | — | |||
| nvidia/Nemotron-3-Labs-Ultra-Math-RLnvidia/Nemotron-3-Labs-Ultra-Math-RL | Not documented | — | — | — | |||
| nvidia/Ising-Calibration-1-35B-A3Bnvidia/Ising-Calibration-1-35B-A3B | Not documented | — | — | — | |||
| nvidia/Riva-Translate-4B-Instruct-v2nvidia/Riva-Translate-4B-Instruct-v2 | Not documented | — | — | — | |||
| nvidia/Qwen-Image-Flashnvidia/Qwen-Image-Flash | Not documented | — | — | — | |||
| nvidia/Riva-Translate-4B-Instructnvidia/Riva-Translate-4B-Instruct | Not documented | — | — | — | |||
| nvidia/Cosmos3-Super-Text2Image-4Stepnvidia/Cosmos3-Super-Text2Image-4Step | Not documented | — | — | — | |||
| nvidia/Kimi-K2.6-DFlashnvidia/Kimi-K2.6-DFlash | Not documented | — | — | — | |||
| nvidia/Nemotron-Cascade-2-30B-A3Bnvidia/Nemotron-Cascade-2-30B-A3B | Not documented | — | — | — | |||
| nvidia/Nemotron-Labs-Audex-2Bnvidia/Nemotron-Labs-Audex-2B | Not documented | — | — | — | |||
| nvidia/Cosmos3-Super-Text2Imagenvidia/Cosmos3-Super-Text2Image | Not documented | — | — | — | |||
| nvidia/Kimi-K2.7-Code-DFlashnvidia/Kimi-K2.7-Code-DFlash | Not documented | — | — | — | |||
| nvidia/Nemotron-Labs-Audex-30B-A3Bnvidia/Nemotron-Labs-Audex-30B-A3B | Not documented | — | — | — | |||
| nvidia/MiniMax-M2.7-DFlashnvidia/MiniMax-M2.7-DFlash | Not documented | — | — | — | |||
| nvidia/Privasis-Cleaner-4Bnvidia/Privasis-Cleaner-4B | Not documented | — | — | — | |||
| nvidia/Privasis-Cleaner-0.6Bnvidia/Privasis-Cleaner-0.6B | Not documented | — | — | — | |||
| nvidia/LocateAnything-3Bnvidia/LocateAnything-3B | Not documented | — | — | — | |||
| nvidia/Nemotron-3-Content-Safetynvidia/Nemotron-3-Content-Safety | Not documented | — | — | — | |||
| nvidia/Kimi-K2.6-Eagle3nvidia/Kimi-K2.6-Eagle3 | Not documented | — | — | — | |||
| nvidia/Kimi-K2.5-Thinking-Eagle3nvidia/Kimi-K2.5-Thinking-Eagle3 | Not documented | — | — | — | |||
| nvidia/CUDA-Autocompletenvidia/CUDA-Autocomplete | Not documented | — | — | — | |||
| nvidia/NVIDIA-Nemotron-3-Ultra-550B-A55B-GenRMnvidia/NVIDIA-Nemotron-3-Ultra-550B-A55B-GenRM | Not documented | — | — | — | |||
| nvidia/Nemotron-Labs-Diffusion-VLM-8Bnvidia/Nemotron-Labs-Diffusion-VLM-8B | Not documented | — | — | — | |||
| nvidia/Nemotron-Labs-Diffusion-3B-Basenvidia/Nemotron-Labs-Diffusion-3B-Base | Not documented | — | — | — | |||
| nvidia/Nemotron-Labs-Diffusion-8B-Basenvidia/Nemotron-Labs-Diffusion-8B-Base | Not documented | — | — | — | |||
| nvidia/Nemotron-Labs-Diffusion-14B-Basenvidia/Nemotron-Labs-Diffusion-14B-Base | Not documented | — | — | — | |||
| nvidia/Nemotron-Labs-Diffusion-14Bnvidia/Nemotron-Labs-Diffusion-14B | Not documented | — | — | — | |||
| nvidia/Nemotron-Labs-Diffusion-3Bnvidia/Nemotron-Labs-Diffusion-3B | Not documented | — | — | — | |||
| nvidia/Nemotron-Labs-Diffusion-8Bnvidia/Nemotron-Labs-Diffusion-8B | Not documented | — | — | — | |||
| nvidia/NVIDIA-Nemotron-Labs-Teacher-STEMnvidia/NVIDIA-Nemotron-Labs-Teacher-STEM | Not documented | — | — | — | |||
| nvidia/NVIDIA-Nemotron-Labs-Teacher-Instruction-Followingnvidia/NVIDIA-Nemotron-Labs-Teacher-Instruction-Following | Not documented | — | — | — | |||
| nvidia/NVIDIA-Nemotron-Labs-Teacher-Competition-Codingnvidia/NVIDIA-Nemotron-Labs-Teacher-Competition-Coding | Not documented | — | — | — | |||
| nvidia/NVIDIA-Nemotron-Labs-Teacher-Chatnvidia/NVIDIA-Nemotron-Labs-Teacher-Chat | Not documented | — | — | — | |||
| nvidia/NVIDIA-Nemotron-Labs-Teacher-General-Reasoningnvidia/NVIDIA-Nemotron-Labs-Teacher-General-Reasoning | Not documented | — | — | — | |||
| NVIDIA: Nemotron 3 Ultra (batch)nvidia/nemotron-3-ultra-550b-a55b:batch | 512.288K | $0.6 | $3.6 | — | |||
| NVIDIA: Nemotron 3.5 Lightning (free)nvidia/nemotron-3.5-lightning:free | 1M | Free | Free | — | |||
| NVIDIA: Nemotron 3 Ultra (free)nvidia/nemotron-3-ultra-550b-a55b:free | 1M | Free | Free | — | |||
| nvidia/NVIDIA-NemotronLabs-AI-for-Media-Sports-Tennisnvidia/NVIDIA-NemotronLabs-AI-for-Media-Sports-Tennis | Not documented | — | — | — | |||
| NVIDIA: Nemotron Nano 9B V2 (free)nvidia/nemotron-nano-9b-v2:free | 128K | Free | Free | — | |||
| NVIDIA: Nemotron 3 Super (free)nvidia/nemotron-3-super-120b-a12b:free | 262.144K | Free | Free | — |