Skip to content

Comment on Is anyone using self hosted LLM day to day and training it like a new employeeparent

Comments

True, and makes sense that the logic is closing in but the breadth of the data is too narrow in 7GB's to ask questions about niche topics.

Mistral hasn't released their own official 13B/30B's yet, but i'm really looking forward to what they can do.

What is crazy is that Ultrafastbert, Speculative, Jacobi, or lookahead decoding could potentially speed up by up to 80x depending on size which could make GPT-4 like models feasible on entry level macs / Phones if similar wizardry is done memory wise.

..Yes im very optimistic after the insane progress over the last months with models like Mistral, Deepseek etc.

AboutSource Built by g1lg1l

Hackerly is an independent reader for Hacker News, built on the public HN API. Not affiliated with Y Combinator.