Nemotron 3.5 Lightning
NVIDIA · Aug 2026NVIDIA's newest open reasoning model - a 30B mixture-of-experts with 3B active per token and a 256K context. Strong reasoning, coding, and tool use; runs on a 32 GB machine with the experts in main memory.
from 18.9 GBruns from 32 GB RAM256K contextOpenMDW 1.1