Fast NVIDIA Nemotron MoE for reliable agentic tasks across enterprise workloads
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
Fast NVIDIA Nemotron MoE for reliable agentic tasks across enterprise workloads
Largest Nemotron 3 model for maximum open-weight reasoning and agent accuracy
Open Nemotron omni model combining reasoning with text, vision, and audio
Nemotron multimodal model for visual reasoning and agentic AI workflows
Nemotron middle tier for collaborative agents and high-volume reasoning workloads
Small Nemotron 3 MoE for efficient coding, math, and long-context agents
Nemotron multimodal model for visual reasoning and agentic AI workflows
Nemotron model for efficient reasoning, coding, and specialized AI agents
Nemotron model for efficient reasoning, coding, and specialized AI agents
Flagship Nemotron model for high-throughput reasoning and complex agents
Nemotron model for efficient reasoning, coding, and specialized AI agents
Compact Nemotron model for efficient reasoning and deployable AI agents
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| Nemotron 3.5 Lightning 30B A3Bnvidia/nemotron-3.5-lightning | 262.144K | $0.05 | $0.2 | 2026-08-11 | |||
| Nemotron 3.5 Lightning 30B A3Bnvidia/nemotron-3.5-lightning-30b-a3b | 262.144K | — | — | 2026-08-11 | |||
| Nemotron 3 Ultra 550B A55Bnvidia/nemotron-3-ultra-550b-a55b | 1M | $0.5 | $2.5 | 2026-06-04 | |||
| Nemotron 3 Nano Omni 30B A3B Reasoningnvidia/nemotron-3-nano-omni-30b-a3b-reasoning | 256K | $0.2 | $0.8 | 2026-04-28 | |||
| Nemotron VoiceChatnvidia/nemotron-voicechat | 128K | — | — | 2026-03-16 | |||
| Nemotron 3 Super 120B A12Bnvidia/nemotron-3-super-120b-a12b | 262.144K | $0.2 | $0.8 | 2026-03-11 | |||
| Nemotron 3 Nano 30B A3Bnvidia/nemotron-3-nano-30b-a3b | 262.144K | $0.05 | $0.2 | 2025-12-15 | |||
| Nemotron Nano 12B v2 VLnvidia/nemotron-nano-12b-v2-vl | 128K | $0.2 | $0.6 | 2025-10-28 | |||
| Llama 3.3 Nemotron Super 49B v1.5nvidia/llama-3.3-nemotron-super-49b-v1.5 | 131.072K | $0.4 | $0.4 | 2025-07-25 | |||
| Llama 3.1 Nemotron 70B Instructnvidia/llama-3.1-nemotron-70b-instruct | 128K | — | — | 2025-04-15 | |||
| Llama 3.1 Nemotron Ultra 253Bnvidia/llama-3.1-nemotron-ultra-253b | 128K | — | — | 2025-04-07 | |||
| Llama 3.3 Nemotron Super 49B v1nvidia/llama-3.3-nemotron-super-49b-v1 | 131.072K | — | — | 2025-04-07 | |||
| Nemotron Mini 4B Instructnvidia/nemotron-mini-4b-instruct | 128K | — | — | 2024-08-21 |