NVIDIA targets 2.6x inference efficiency gains via full-stack AI factory optimization | StreamingMeme