Skip to content

Comment on Classifying all of the pdfs on the internetparent

Comments

Nice article, thanks for sharing.

I can imagine mining all of these articles was a ton of work. I’d be curious to know how quickly the computation could be done today vs. the 13 hour 2009 benchmark :)

Nowadays people would be slamming those data through UMAP!

Thanks! Yes, I'm sure using modern hardware and modern techniques would improve both computation time and results.

AboutSource Built by g1lg1l

Hackerly is an independent reader for Hacker News, built on the public HN API. Not affiliated with Y Combinator.