Skip to content

Comment on Ask HN: Kind of data needed to build a language model?

Comments

I shudder at the complexity of the task, but I almost think you need to manually tweak the vector space of a model that already works, and I have no idea how to do that in practice.

The amount of text required for a machine to grind through it millions of times to tease out the shape of a language doesn't sound like something you have. If you have the time of native speakers, it might be possible to build tools for them to correct the most "off" parts of the model interactively.

AboutSource Built by g1lg1l

Hackerly is an independent reader for Hacker News, built on the public HN API. Not affiliated with Y Combinator.