Earlier quoted context omitted.
I suspect it's mainly the reduced maintenance and reduction of workload needed to support, especially with more platforms coming to be supported (not so long ago there was no ARM64 nvidia support, now they are shipping their own ARM64 servers!) What really changed the situation is that Turing architecture GPUs bring new, more powerful management CPU, which has enough capacity to essentially run the OS-agnostic parts…
Am I correct in reading that as Turing architecture cards include a small CPU on the GPU board, running parts of the driver/other code?
"Open Drivers" from nVidia include different firmware that utilizes the new-found performance.