Nemotron 3.5 Lightning from @NVIDIAAI is out today and available on Tinker. With just 3B active parameters and optimized for throughput speed, 3.5 Lightning is designed for work where latency and cost matter.
Introducing NVIDIA Nemotron 3.5 Lightning⚡
An open 30B MoE model with 3B active parameters, built for always-on agents to complete high-volume, specialized tasks faster.
It delivers up to 4x the output speed of similar-sized models.