Skip to content

Comment on Show HN: Samosa Chat - Run Qwen3.6-35B-A3B Locally on a 16 GB Macparent

Comments

Thanks for testing. Actually if you open the samosa app; it will show you memory consumption, tokens/sec etc. Let me know what you think.

Great, yes with samosa app can see details, it is 6.52tok/sec, 4.05GB memory. Use case, I have been waiting was to run some agents in my Mac continuously with local models, but without my Mac getting heated. Seems I can experiment with samosa inference engine.

AboutSource Built by g1lg1l

Hackerly is an independent reader for Hacker News, built on the public HN API. Not affiliated with Y Combinator.