Skip to content

Comment on What do you think of people buying Mac mini's to run AI?

Comments

The premise that 'barely any decent size models can run on it' misses the biggest advantage of Apple Silicon: Unified Memory. Where else can you get a machine with 64GB or 128GB of VRAM for running quantized models at this price point? Buying the equivalent VRAM in Nvidia GPUs (like multiple RTX 3090s/4090s) would cost thousands of dollars, draw massive power, and sound like a jet engine. The Mac Mini is dead silent, sips power, and lets you run 70B+ parameter models locally via llama.cpp. It's currently the undisputed king of VRAM-per-dollar for local inference.

When released unified memory was a issue since they are not swappable with a higher memory sticks.

Now it's game changer when it comes to AI, you're right, it delivers performance for local inference

I believe AMD Ryzen AI Max 395+ with 128GB of RAM is the "undisputed king of VRAM-per-dollar", but Macs are slightly better performing for local inference.

This is interesting.

What kind of models have you run on this configuration?

AboutSource Built by g1lg1l

Hackerly is an independent reader for Hacker News, built on the public HN API. Not affiliated with Y Combinator.