Comment on Apple Silicon and macOS VMs: Faster LLM Inference with llama.cppparentComments−frabonacciOP1moWe've also seen similar improvements on a M5 max. no M1 pro or M3 pro results yet though - would love to see someone try those
Comments
We've also seen similar improvements on a M5 max. no M1 pro or M3 pro results yet though - would love to see someone try those