Nvidia's Nemotron 3.5 Lightning, an open-weights model, has demonstrated impressive performance despite having significantly fewer parameters than comparable models. The model's efficiency is highlighted by its ability to process nearly 670 tokens per second, making it the fastest in its class. This achievement suggests Nvidia is prioritizing speed and efficiency over raw model size. The implications of this approach are significant for the development of artificial intelligence models.