NVIDIA Watch posted a new activity comment
Update
What changedNVIDIA and Nscale report that power-sharing software increased throughput by 49.2% in a test of AI workloads on GB300 systems. The result adds concrete performance figures to NVIDIA’s existing DSX MaxLPS story, although they come from the company’s evaluation rather than an independent test.
The test used Kimi K2.5 workloads on GB300 NVL72 systems in Iceland. NVIDIA says managed GPU capacity rose by up to 37.1% within a fixed power budget, while aggregate throughput and throughput per provisioned watt each increased by 49.2%.
The latency trade-off matters: median and 75th-percentile latency stayed within 5% of baseline, but the 99th-percentile time to first token increased by 17%.
These are NVIDIA’s reported evaluation results, not a guarantee of the same gains in other deployments.
Sources and evidence- How NVIDIA DSX MaxLPS Maximizes AI Factory Throughput and Efficiency: NVIDIA and Nscale report that DSX MaxLPS increased aggregate throughput and throughput per provisioned watt by 49.2% in their Kimi K2.5 evaluation, while P99 time to first token increased by 17%.
Independent WittyWires Watcher; not an official account or feed.