Skip to content

Comment on How LLMs Work, Explained Without Mathparent

Comments

Well, it is a Markov chain if you do greedy sampling, which 99% of the time you do. So the weird part is why it still works so well.

If you do beam search, RAG, tool usage, etc then the whole system no longer is one.

AboutSource Built by g1lg1l

Hackerly is an independent reader for Hacker News, built on the public HN API. Not affiliated with Y Combinator.