NVIDIA Nemotron 3.5 Lightning Launches at 1,200 Tokens/Second—30B MoE Outpaces Gemma 4 by 29× and Qwen 3.6 by 35% on Agent Tasks
NVIDIA just shipped a 30B parameter model that outputs 1,200 tokens per second—29 times faster than Gemma 4 26B on identical hardware. This…