Skip to content

Comment on OCR by uploading images to Google Docs

Comments

Incidentally, I noticed that if you try to use tesseract on an image taken from a Google Books page, you get terrible OCR accuracy. Anyone know why that is?

I recall that on some google-scanned books, there was some metadata from abbyy finereader. So that may be why.

Also, tesseract often needs to be configured.

AboutSource Built by g1lg1l

Hackerly is an independent reader for Hacker News, built on the public HN API. Not affiliated with Y Combinator.