On the same machine I experimented with IOMMU and a lightweight VM to run Nvidia Pascal era GPUs in a VM with 580 drivers, while Nvidia Blackwell on the host uses 590. I vibe coded that too
When exposed via openai compatible endpoints with GPUStack https://github.com/gpustack/gpustack I could use the combined compute power from both generations on a single machine.
Comments
Funny you should say that
On the same machine I experimented with IOMMU and a lightweight VM to run Nvidia Pascal era GPUs in a VM with 580 drivers, while Nvidia Blackwell on the host uses 590. I vibe coded that too
When exposed via openai compatible endpoints with GPUStack https://github.com/gpustack/gpustack I could use the combined compute power from both generations on a single machine.