NVIDIA Releases Nemotron 3.5 Lightning Open Model
Open 30B MoE model with 3B active parameters targets always-on agents.
NVIDIA announced updated weights for its Nemotron 3.5 Lightning model. Director of Research Pavlo Molchanov stated the changes apply distillation from larger Super and Ultra models during continuous pretraining and post-training. Bryan Catanzaro described the result as combining the earlier Nano architecture with Super-level intelligence plus speculative decoding. The open model is positioned for high-volume specialized agent tasks and appears on platforms including Together AI and Prime Intellect.
Introducing NVIDIA Nemotron 3.5 Lightning⚡ An open 30B MoE model with 3B active parameters, built for always-on agents to complete high-volume, specialized tasks faster. It delivers up to 4x the output speed of similar-sized models.


