Earlier quoted context omitted.
With unified memory, reading from RAM to GPU compute buffer is not that painful, and you can use partial RAM caching to minimize the impact of other kinds of swapping.
In practical terms, is this kind of architecture available to consumers except through Apple?
There's also the Nvidia DGX Spark.