Skip to content

Comment on Show HN: Fine-tune an 8B model on a 4 GB laptop GPUparent

Comments

If you are into small local models I highly recommend vibe thinker. It's a model trained specifically for reasoning. Basically a problem solver. When compared with other models, on math problems benchmarks, it's closer to models hundred times its size than ten times its size which it beats comfortably.

It supports long contexts on limited VRAM and is blazing fast.

https://github.com/WeiboAI/VibeThinker

this is great! thanks for sharing

Wat. Those are some crazy benchmark scores for a 3B model

AboutSource Built by g1lg1l

Hackerly is an independent reader for Hacker News, built on the public HN API. Not affiliated with Y Combinator.