logoalt Hacker News

OCR It – pull text out of un-copyable documents for your LLM

31 pointsby thiagolimatoday at 6:25 AM8 commentsview on HN

Comments

Barbingtoday at 7:06 AM

  “Pin a region once. Hit a hotkey on every page. Get the whole book as text.”
Much better than the old definition of “region lock”, nice.

HN isn’t a fan of the generated readmes though, though vibed software (thoroughly used) can be all good.

show 1 reply
tobinfekkestoday at 7:01 AM

Also available natively to the OS (Windows) with PowerToys, if you want an alternative to a browser extension. One of the unsung heroes of that library.

Jury is still out on which is more trustworthy handling any personal data, Microsoft or Google. Neither.

jbverschoortoday at 8:35 AM

Is this similar to CleanshotX?

harsh_patel14today at 6:32 AM

This is handy — I've hit this exact issue prepping documents for LLM context. How's the accuracy on lower quality scans?

show 1 reply
thiagolimatoday at 6:26 AM

[flagged]