Nvidia's Open-Weight Nemotron 3.5 Lightning Trades Model Size for Speed

Nvidia's new open-weight Nemotron 3.5 Lightning packs just 3.6B active parameters yet matches gpt-oss-120b's Intelligence Index score at ~670 tokens/sec.