Skip to content

Comment on The Guantanamo Filesparent

Comments

So I was fiddling a bit. Sadly, despite being able to skip the OCR step, there's obviously a lot of OCR errors in the text as is, which would make reliably stripping out data somewhat more difficult.

The reality is that I don't really have the time to commit to this unless everything worked perfect. Ah well, maybe someone else will do it.

AboutSource Built by g1lg1l

Hackerly is an independent reader for Hacker News, built on the public HN API. Not affiliated with Y Combinator.