Skip to content

Comment on State-of-the-Art Chatbot, Vicuna-7B, now runs on MacBook with GPU accelerationparent

Comments

The shared ram and neural engine make for an interesting/powerful platform if people are willing to port to it.

Are the neural engines able to be leveraged by 3rd parties yet? I thought there was no API available yet.

They are leveraging Apple’s Metal Performance Shaders[1] not the neural engine. From the chart, it looks like you might get ~20x max boost on inference over plain CPU. Obviously, it's not like having RTX 4090 but better than nothing.

[1] https://pytorch.org/blog/introducing-accelerated-pytorch-tra...

CoreML is the API.

AboutSource Built by g1lg1l

Hackerly is an independent reader for Hacker News, built on the public HN API. Not affiliated with Y Combinator.