Fast NVIDIA Nemotron MoE for reliable agentic tasks across enterprise workloads
10 providers
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
Fast NVIDIA Nemotron MoE for reliable agentic tasks across enterprise workloads
Largest Nemotron 3 model for maximum open-weight reasoning and agent accuracy
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| Nemotron 3.5 Lightning 30B A3Bnvidia/nemotron-3.5-lightning | 262.144K | $0.05 | $0.2 | 2026-08-11 | |||
| Nemotron 3 Ultra 550B A55Bnvidia/nemotron-3-ultra-550b-a55b | 1M | $0.5 | $2.5 | 2026-06-04 |