Fast NVIDIA Nemotron MoE for reliable agentic tasks across enterprise workloads
10 providers
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
Fast NVIDIA Nemotron MoE for reliable agentic tasks across enterprise workloads
Safety model for policy screening, moderation, and risk-aware routing workflows
Largest Nemotron 3 model for maximum open-weight reasoning and agent accuracy
Nemotron middle tier for collaborative agents and high-volume reasoning workloads
Small Nemotron 3 MoE for efficient coding, math, and long-context agents
Nemotron model for efficient reasoning, coding, and specialized AI agents
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| Nemotron 3.5 Lightning 30B A3Bnvidia/nemotron-3.5-lightning | 262.144K | $0.05 | $0.2 | 2026-08-11 | |||
| Nemotron 3.5 Content Safetynvidia/nemotron-3.5-content-safety | 128K | $0.2 | $0.2 | 2026-06-04 | |||
| Nemotron 3 Ultra 550B A55Bnvidia/nemotron-3-ultra-550b-a55b | 1M | $0.5 | $2.5 | 2026-06-04 | |||
| Nemotron 3 Super 120B A12Bnvidia/nemotron-3-super-120b-a12b | 262.144K | $0.2 | $0.8 | 2026-03-11 | |||
| Nemotron 3 Nano 30B A3Bnvidia/nemotron-3-nano-30b-a3b | 262.144K | $0.05 | $0.2 | 2025-12-15 | |||
| Llama 3.3 Nemotron Super 49B v1.5nvidia/llama-3.3-nemotron-super-49b-v1.5 | 131.072K | $0.4 | $0.4 | 2025-07-25 |