This is great news. The more projects do this, the less of a moat CUDA is, and the less of a competitive advantage Nvidia has.
Please check out https://github.com/geniusrise - tool for running llms and other stuff, behaves like docker compose, works with whatever is supported by underlying engines:
Huggingface - MPS, cuda VLLM - cuda, ROCm llama.cpp, whisper.cpp - cuda, mps, rocm
Also coming up integration with spark (TorchDistributor), kafka and airflow.